Applied AI master's student · Oslo

The model is
the easy part.

I'm Idris. I fine-tune models, build the pipelines around them, and write up what the results actually support. Currently a master's student in Oslo, looking for an internship where that work is useful.

Open to internship conversations

01 / Selected work

Six projects, including the parts that did not work.

Each one lists what I actually measured. Where a project has no numbers yet, it says so.

Fine-tuning study · Feb 2025

Teaching Llama 2 to answer questions about one university.

Students at Abdul Wali Khan University Mardan ask the same admissions and programme questions every term. I wanted to know how small a dataset could be and still shift a 7B model's behaviour, so I wrote 41 question–answer pairs by hand and fine-tuned on those alone.

It works for questions close to the training set and invents details for anything outside it. 41 pairs buys you tone and format, not knowledge. The honest fix is retrieval over the university's actual documents, which is what I would build if I started again.

Base model
Llama-2-7b-chat, loaded in 4-bit NF4
Method
QLoRA at rank 64, alpha 16, dropout 0.1
Training run
1 epoch, batch size 4, lr 2e-4, on a single Colab T4
Evaluation
Manual inspection only. There was no held-out set, which is the main gap
Training configuration
Dataset
41 hand-written pairs
Adapter
LoRA rank 64, alpha 16
Precision
4-bit NF4 · fp16 compute
Libraries
transformers · peft · trl
Hardware
Single T4, free tier
  1. 2026

    Human Insight AI: one video interface for five detectors

    A Flask app (main-app.py) that runs several vision models over an uploaded video behind one interface. Each detector is its own module (human, fire, gender, ethnicity, fight) and they all return the same prediction shape, so the streaming pipeline does not care which one is selected. Fight detection is the exception: it needs the whole clip. I would drop gender and ethnicity for the reasons set out here.

    Flask · OpenCV · YOLO

    Repository

  2. 2024

    Reconstructing a complete fingerprint from a partial image

    A GAN trained on SOCOFing. Average SSIM 0.64 across two runs, which is moderate structural similarity. The discriminator still spots most reconstructions as fake.

    PyTorch · OpenCV

    Case study

  3. 2024

    Counting people in video, then deciding not to ship half of it

    YOLOv8 tracking with line-crossing counts. The second half of the brief was a race classifier, and I wrote up why it should not be deployed.

    YOLOv8 · OpenCV · DeepFace

    Case study

  4. 2024

    Sorting property photos for a paying client

    A Fiverr client needed property photos sorted into clearly visible and obscured, where obscured means a tree, a parked car or a wall in front of the house. The interesting part was that the client's definition, not a public label set, decided what counted as correct.

    Client work · computer vision

    Client-owned

  5. 2023

    Recognising hand signs from a webcam in real time

    MediaPipe landmarks into a Random Forest over 8 sign classes. 100% on the held-out split, which says more about the split than the classifier.

    OpenCV · MediaPipe · scikit-learn

    Case study

  6. Now

    Marginalia, retrieval over real documents

    Improving my portfolio assistant with a production-minded RAG pipeline: FastAPI searches structured Markdown knowledge with FAISS and all-MiniLM-L6-v2 embeddings, then uses OpenAI gpt-oss-20b through Groq to generate grounded, useful answers without exposing private credentials.

    gpt-oss-20b · RAG · FAISS · FastAPI

    In progress

02 / About

Four years of computer vision, now moving into language models.

Idris Khattak

I did my bachelor's in AI in Pakistan, where most of my work was computer vision: segmentation, detection, tracking. My final-year project counted people in video using YOLOv8. Since then I have moved toward language models, and I am now a master's student in Applied AI at OsloMet.

The thing I keep relearning is that the model is the easy part. The fingerprint project stalled on image preprocessing for weeks. The Llama fine-tune taught me that 41 examples will change how a model sounds and nothing about what it knows. I am more interested in that gap than in the training loop.

Right now I am working on evaluation, because I have shipped too many projects where "it looks right" was the only test I ran.

Based inOslo, Norway
LanguagesEnglish, Urdu & learning Norwegian
Looking forAI, ML and data internships

Tools, and where I have actually used them

PyTorch · Hugging Face transformers, PEFT, TRL
QLoRA fine-tuning of Llama 2 7B for the AWKUM assistant
Ultralytics YOLOv8 · OpenCV · DeepFace
Human counting and ethnicity detection for my final-year project
Python · Pandas · NumPy · scikit-learn
The everyday work: cleaning data, baselines, and figuring out what the data will not support
Retrieval pipelines · REST APIs · Vercel
The assistant answering questions on this page right now
Git · Jupyter · AWS SageMaker
Coursework and self-study. I am comfortable here, not expert yet

03 / Journey

Where I have been.

Pakistan to Oslo, computer vision to language models.

Aug 2026 – now

Current chapter

Master's in Applied Artificial Intelligence

Oslo Metropolitan University · Oslo

I am beginning advanced study in machine learning, intelligent systems and practical AI while developing a stronger engineering and research foundation.

2024 – 2026

Freelance and self-directed work

Pakistan

Freelance work across AI and web projects: an image-classification job for a Fiverr client, and Shopify store design and web-scraping workflows for Loomlan. I also wrote and taught an introductory Python course.

2020 – 2024

BS Artificial Intelligence

Hazara University Mansehra

Four years across machine learning, computer vision and digital image processing. Final-year project: people counting and attribute estimation from video.

Jun – Sep 2023

AI Programming with Python Nanodegree

Udacity · AWS-sponsored scholarship

Strengthened my practical foundation in Python, PyTorch, mathematics and end-to-end machine learning.

View credential

05 / Contact

Looking for a summer 2027 internship.

Computer vision, LLM work, or the data engineering underneath either. I am in Oslo, I can work in English, and my Norwegian is improving. If you have something you think I could contribute to, write to me.