INSERT COIN — AI ENGINEER ARCADE

MUDASSAR

AWAN

AI / ML ENGINEER

VOICE◆VISION◆LLMs

AI/ML engineer who designs, trains, and ships production models across deep learning, NLP, computer vision, and large language models. I focus on automation, real-time voice agents, and LLM training — deployed as REST APIs on AWS & Azure.

▸ FOCUSAUTOMATION·VOICE AGENTS·LLM TRAINING
1.84B+
URDU TOKENS SCRAPED
7
NLP BENCHMARKS — SOTA
96.4%
TRACKING CONTINUITY
99%+
CLAIM-MODEL ACCURACY
▼ SCROLL TO PLAY ▼
VOICE AI◆COMPUTER VISION◆LLM PRE-TRAINING◆RAG◆MLOps◆YOLO26◆ArcFace◆LiveKit◆PyTorch◆QALB · URDU LLM◆REAL-TIME TRACKING◆LoRA / QLoRA◆VOICE AI◆COMPUTER VISION◆LLM PRE-TRAINING◆RAG◆MLOps◆YOLO26◆ArcFace◆LiveKit◆PyTorch◆QALB · URDU LLM◆REAL-TIME TRACKING◆LoRA / QLoRA◆
01SELECT PLAYER

PLAYER ONE

Mudassar Awan
▸ PLAYER SELECTED
MUDASSAR
@mudassar531
CLASSAI/ML ENGINEER
GUILDFAST NUCES '25 · AI
REGIONPakistan · Remote
QUESTOpen to collaborate

I build production AI end to end — from voice agents and computer-vision pipelines to large language models.

AI/ML engineer who designs, trains, and ships production models across deep learning, NLP, computer vision, and large language models. I focus on automation, real-time voice agents, and LLM training — deployed as REST APIs on AWS & Azure. I contributed to Qalb, the world's largest Urdu LLM, and I'm currently building across three teams at once.

NOW PLAYING:AGENTS LIMITEDEMBERNavAI
02CAREER MAP

LEVELS CLEARED

01

AGENTS LIMITED

Product Owner — CounterVision

Remote

● NOW PLAYINGJan 2026 — Present

In-store retail intelligence — turning anonymous foot traffic into visitor identity, movement, dwell & conversion analytics from existing cameras.

  • ▸Own the product + lead CounterVision v2: real-time person tracking & face recognition, benchmarking 10 model+tracker configs on an RTX 5060.
  • ▸Architected the multi-stage pipeline: YOLO26 detection → ByteTrack / BoT-SORT tracking → SCRFD face detection → ArcFace 512-d embeddings. 96.4% tracking continuity at 18.7 FPS.
  • ▸Shipped quality-gated recognition (~60-70% compute saved), multi-frame embedding fusion, clothing-color HSV histograms for body re-ID, and self-correction logic.
YOLO26ByteTrackBoT-SORTSCRFDArcFaceONNX-GPUVIEW ↗
02

EMBER

Software Engineer

Remote

● NOW PLAYINGMay 2026 — Present

Building on Ember's AI knowledge & operator-brain platform. First mission: a full meeting-connector integration, shipped end to end.

  • ▸Designed and shipped a new meeting-recorder connector across the whole stack — backend connector, admin-panel UI, and the operator-brain auto-trigger pipeline.
  • ▸Indexing → semantic search → cited answers, validated end to end and merged to main with full CI + adversarial review.
PythonFastAPINext.jsCeleryPostgresVespaVIEW ↗
03

NavAI

Voice AI Engineer

Tashkent, Uzbekistan · Remote

● NOW PLAYINGApr 2026 — Present

Building Uzbek voice agents end to end — custom STT/TTS, real-time flows, observability, and a model playground.

  • ▸Built the end-to-end Uzbek voice-agent pipeline: custom STT/ASR + TTS, GPT-4o, LiveKit real-time transport, and SIP telephony (migrated off Yandex).
  • ▸Stood up a full observability platform — Prometheus + Grafana + cAdvisor + Node Exporter on GCP, Langfuse tracing, and Telegram alerting via a webhook relay.
  • ▸Created a model playground + A/B testing harness and a Post-Call QA agent (MongoDB) detecting language switches, transfers, dead-ends & self-resolution.
  • ▸Led the NAV-175 root-cause investigation across 750+ calls — found the non-interruptible greeting driving 22% single-turn calls.
LiveKitSIPGPT-4oLangfusePrometheusGrafanaMongoDBVIEW ↗
04

CARECLOUD MTBC

Junior AI Engineer

Rawalpindi, Pakistan

✓ CLEAREDJul 2025 — May 2026

Healthcare AI — claim-denial prediction, medical-coding automation, and a voice agent for doctors.

  • ▸Built a claim-denial classifier on millions of 837-EDI insurance claims: 99%+ accuracy on accepted, 80%+ on denied.
  • ▸Shipped a CPT/ICD mapping RAG pipeline + clinical-document classification, replacing manual workflows for 500+ staff.
  • ▸Built an AI dashboard over 2007–2024 clinical records and contributed to 'Status AI', a voice agent for doctors.
Scikit-learnPyTorchRAGFastAPIAWSDockerVIEW ↗
05

GENESIS LAB

AI Research Intern

Remote

✓ CLEAREDFeb 2024 — May 2024

Research on LLMs, GANs, and RAG systems for document automation.

  • ▸Researched and implemented LLMs, GANs, and RAG systems for document-automation workflows.
LLMsGANsRAG
06

IMARAT GROUP

ML Intern

Islamabad, Pakistan

✓ CLEAREDSep 2023 — Dec 2023

Classification/regression for real-estate valuation + YOLOv8 detection.

  • ▸Trained classification/regression models for real-estate valuation and optimized YOLOv8 detection pipelines.
TensorFlowYOLOv8Scikit-learnVIEW ↗
VOICE AI◆COMPUTER VISION◆LLM PRE-TRAINING◆RAG◆MLOps◆YOLO26◆ArcFace◆LiveKit◆PyTorch◆QALB · URDU LLM◆REAL-TIME TRACKING◆LoRA / QLoRA◆VOICE AI◆COMPUTER VISION◆LLM PRE-TRAINING◆RAG◆MLOps◆YOLO26◆ArcFace◆LiveKit◆PyTorch◆QALB · URDU LLM◆REAL-TIME TRACKING◆LoRA / QLoRA◆
03SELECT A GAME

CARTRIDGES

WORLD'S LARGEST URDU LLMRESEARCH · arXiv 2601.08141

QALB

Contributor to Qalb — the first dedicated Urdu LLM, closing the gap for 230M+ speakers.

  • +1.97B tokens · LLaMA 3.1 8B continued pre-training with LoRA
  • +Scraped 1.84B+ Urdu tokens, 67.8% retention cleaning pipeline
  • +90.34 weighted avg across 7 benchmarks — +3.24 over prior SOTA
LLaMA 3.1PyTorchcrawl4aiHuggingFaceLoRA
▶ INSERT & PLAY ↗
OPEN SOURCE · SOLOMIT · v0.1.0

HEARSAY

crawl4ai for video & audio — one command turns any YouTube video, podcast, or recording into clean, timestamped, LLM-ready markdown.

  • +Captions-first with automatic faster-whisper fallback
  • +Batch playlists & podcast feeds; JSON sidecars for RAG
  • +Ships an MCP server for Claude & other AI agents
Pythonfaster-whisperyt-dlpffmpegMCP
▶ INSERT & PLAY ↗
VOICE AGENT NEURAL OS

VANOS

A real-time voice AI desktop assistant with sub-second latency running on consumer hardware.

  • +Sub-second end-to-end voice latency on a laptop
  • +5 Voice-to-Action (V2A) tools wired into the OS
  • +Solved critical mic-compatibility & audio challenges
PipecatDeepgram Nova-3GroqKokoro TTSSilero VAD
REAL-TIME RETAIL CVPRODUCT

COUNTERVISION

Person tracking + face recognition that turns store cameras into visitor analytics.

  • +96.4% tracking continuity at 18.7 FPS on an RTX 5060
  • +Quality-gated recognition saving ~60-70% compute
  • +Body re-ID via clothing-color HSV histograms
YOLO26BoT-SORTSCRFDArcFace
▶ INSERT & PLAY ↗
END-TO-END MLOps

CLAIM DENIAL

A full ML pipeline for 837-EDI ingestion, classification, and live claim validation with auto-retraining.

  • +Millions of insurance claims, 99%+ / 80%+ accuracy
  • +Live validation API + automated retraining loop
  • +Dockerized & deployed on AWS
Scikit-learnPyTorchFastAPIAWSDocker
SMART RETAIL (FYP)

DASHGRAB

A computer-vision system that detects items shoppers pick up and auto-generates their bill.

  • +Pick-detection from overhead camera streams
  • +Pose + segmentation fusion for hand-item association
  • +Auto-billing without checkout scanning
YOLOv11Detectron2OpenPoseMediaPipeSAM
AUTONOMOUS AI AGENT

PORTAL AGENT

An autonomous agent that automates university-portal tasks end to end.

  • +Tool-using LangChain agent over a web portal
  • +Plans + executes multi-step portal workflows
  • +FastAPI backend with a JS front-end
LangChainPyTorchFastAPIJavaScript
04INVENTORY

STAT SHEET

COMPUTER VISION

LV 95
YOLO26 / v11RT-DETRInsightFace (SCRFD, ArcFace)ByteTrackBoT-SORTOpenCVSAMOpenPoseMediaPipe

VOICE & SPEECH AI

LV 90
PipecatDeepgram Nova-3Kokoro TTSSilero VADLiveKitSIP TelephonyGroqElevenLabsLangfuse

LLMs & NLP

LV 92
Continued pre-trainingLoRA / PEFT / QLoRARAGLangChainOpenAI / GroqOllamaTokenizationcrawl4ai

ML FRAMEWORKS

LV 93
PyTorchTensorFlowKerasScikit-learnHuggingFaceUltralyticsUnsloth

DEPLOYMENT & MLOps

LV 88
FastAPIFlaskDockerAWSAzureONNX Runtime (GPU)StreamlitA100 / RTX

OBSERVABILITY

LV 85
PrometheusGrafanacAdvisorNode ExporterMongoDB aggregationTelegram Bot API
05 — GAME OVER? NOT YET

CONTINUE?

I'm open to collaborating, contributing, and building ambitious AI. Drop a coin and let's talk.

▮ PRESS ANY KEY TO CONNECT