Available for AI engineering work

Muneeb Ali Anwar

Full-Stack AI Engineer building
|

5 years shipping intelligent systems that solve real problems — from enterprise HCM AI engines and multi-agent pipelines to voice cloning and autonomous web agents.

↓

About

Engineer who ships intelligence

I design, train, and ship machine-learning and deep-learning systems end to end — from CNNs, LSTMs and Transformers (BERT, GPT, T5) to production agentic AI with LangChain, LangGraph and MCP. I build RAG pipelines over vector databases, orchestrate multi-LLM engines with automatic fallback, and back it all with scalable FastAPI / Django / NestJS services, Docker, and CI/CD.

Five years in, currently at IKONIC Dev, engineering 17+ production AI features for an enterprise HCM platform — obsessed with agentic AI, multi-LLM fallback, and clean backends.

Islamabad / Rawalpindi, Pakistan
0+

Years Experience

0

Live Products

0+

AI Features Shipped

0+

Public Repos

PythonFastAPIDjangoNestJSLangChainLangGraphMCPOpenAIClaudeGeminiGroqRAGPineconeChromaDBWhisperXCoqui XTTSOllamaPyTorchTensorFlowDockerNext.jsReactThree.jsPostgreSQLMongoDBn8nPlaywrightAWSPythonFastAPIDjangoNestJSLangChainLangGraphMCPOpenAIClaudeGeminiGroqRAGPineconeChromaDBWhisperXCoqui XTTSOllamaPyTorchTensorFlowDockerNext.jsReactThree.jsPostgreSQLMongoDBn8nPlaywrightAWS

Experience

2+ years shipping AI in production

The timeline fills as you scroll.

IKONIC Dev

Jan 2026 — Present

AI Engineer

  • Engineering production AI features for an enterprise HCM platform — a multi-provider LLM engine (Gemini → Groq → OpenAI) with automatic fallback for high availability and cost control.
  • Building RAG pipelines with vector databases powering policy Q&A assistants, AI helpdesk, recruitment scoring, and performance analytics.
  • Developing scalable backends with FastAPI and Django, containerized with Docker and shipped through CI/CD pipelines.
  • Designing multi-tenant, multi-country architecture with real-time biometric attendance and productivity tracking.

Python Technologies

Oct 2025 — Jan 2026

Senior AI Engineer

  • Led development of AI/ML solutions with Python, TensorFlow, PyTorch; built agentic AI systems with LangChain and LangGraph.
  • Designed LLM-powered apps with RAG architecture, Pinecone / ChromaDB vector stores, prompt engineering, and Transformer fine-tuning (BERT, GPT).
  • Shipped production APIs with FastAPI, Flask, Docker; MLOps pipelines plus automation via n8n, Zapier, and Make.com.
  • Built NLP solutions for document processing, chatbots, and intelligent automation.

GrowUp-TechSols — NASTP

Dec 2024 — Sep 2025

AI Engineer

  • Developed computer-vision models with Python, PyTorch, OpenCV — real-time hand-gesture detection with MediaPipe and facial emotion recognition on FER2013.
  • Deployed ML/DL models via Flask REST APIs with full data preprocessing, EDA, and model evaluation.

NexoraSols

Feb 2024 — Dec 2024

Machine Learning Engineer

  • Built and deployed ML/DL models using Python, PyTorch, and OpenCV with Flask REST APIs.
  • Designed data preprocessing pipelines — feature extraction, cleaning, and augmentation.

Projects

Production systems, live deployments

10 of these are live right now — click through and try them.

hcm.owesome.work
WorkPulse HCM screenshot

WorkPulse HCM

AI-Powered Human Capital Management Platform

LIVE

17+

AI features

3

LLM providers

Multi

Countries

  • Architected a multi-provider AI engine with automatic fallback powering 17+ AI capabilities across the employee lifecycle.
  • Built “Buddy” — a RAG HR assistant with vector search over policies, plus an AI Helpdesk with auto-categorization and sentiment-based priority.
  • Delivered AI Recruitment (LLM candidate scoring), AI Performance (bias detection, flight-risk), and an AI Learning Engine (auto quizzes, skill-gap analysis).
  • Integrated shift-aware biometric attendance with real-time ZKTeco sync, multi-company / multi-country support, and productivity tracking.
FastAPIDjangoMulti-LLM (Gemini/Groq/OpenAI)RAGVector DBDockerCI/CD
websitebuilder-ochre.vercel.app
Site Generator screenshot

Site Generator

Autonomous Agent Pipeline That Researches & Builds Client Websites

LIVE

69%

Cache hit rate

$1.39

Cost / site

25+

Sites shipped

  • An operator console that turns a business name into a researched, designed, deployed website — end to end, with no human writing code.
  • Agents research the business, generate a content brief, build the site, push a repo, and deploy to Vercel; a human review gate holds every link before it goes out.
  • Brief caching cuts repeat research spend — 69% cache-hit rate on 13 tracked jobs, driving cost down to ~$1.39 per generated site.
  • Full job telemetry: per-job cost, stage tracking, status filtering, and one-click preview + repo links for every generated build.
Next.jsLLM AgentsBrief CachingJob QueueCost TelemetryVercel APIGitHub API
hakem.ai
Hakem.ai screenshot

Hakem.ai

AI Insurance Comparison Platform — Saudi Market

LIVE

Seconds

Compare time

🇸🇦 KSA

Market

  • Intelligent insurance comparison with PDF processing, automated quote extraction, and AI policy analysis using RAG.
  • Pinecone vector search for semantic retrieval and context-aware recommendations with LangChain + GPT-4.
  • Side-by-side provider scoring with premium, rate, strengths and a recommended plan, exportable as a PDF report.
  • FastAPI + MongoDB backend with PDF parsing (PyPDF2, pdfplumber), deployed on DigitalOcean with Docker and CI/CD.
LangChainFastAPIMongoDBRAGOpenAI GPT-4PineconeDigitalOcean
householdprofits.com
HouseholdProfits screenshot

HouseholdProfits

Velocity AI — Multi-Agent Brain for Household Financial Governance

LIVE

12+

Endpoints

householdprofits.com

Domain

  • The AI brain behind a live household financial-governance product — assigns income, protects savings, and enforces spending rules before money moves.
  • Multi-agent FastAPI service: a conversational onboarding agent (state-machine intake) plus a Co-Pilot chat agent with router → tools → framing architecture.
  • Tiered LLM routing — a premium model for strategy and financial reveals, a fast model for routing, extraction, and summarization.
  • Every LLM output is structured and Pydantic-validated; deterministic math, rules, risk-scoring, and forecasting engines kept fully separate from the LLM layer.
  • Stateless, provider-agnostic design with optional Chroma vector memory and 12+ endpoints: transaction tagging, governance evaluation, couples synergy, forecasting, risk, milestones.
FastAPILangChainOpenAIPydantic v2ChromaDBMulti-AgentLaravel Integration
Voice Cloning Pipeline screenshot

Voice Cloning Pipeline

Oral-History Voice Preservation — 9-Stage Automated Pipeline

9

Pipeline stages

3–15s

Clip length

22kHz

Sample rate

  • End-to-end voice cloning from raw two-voice interview audio: a fully automated 9-stage pipeline from ingest to a fine-tuned per-person voice model.
  • WhisperX transcription + alignment + speaker diarization, then LLM-based narrator detection (Llama 3.1 via Ollama) to isolate the right speaker automatically.
  • Silence-based chunking to 3–15s clips at 22050 Hz, per-chunk Whisper transcription with text cleaning, and LJSpeech dataset building with eval splits.
  • Per-speaker Coqui XTTS v2 fine-tuning driven by one idempotent CLI orchestrator with resumable status tracking — deployed to an Azure GPU VM via a self-hosted CI runner.
PythonWhisperXWhisperOllama · Llama 3.1Coqui XTTS v2FastAPIAzureCI/CD
verve-alpha-three.vercel.app
Verve AI screenshot

Verve AI

Real-Time AI Interview Copilot — Desktop & Browser

LIVE

Real-time

Latency

4

Platforms

  • Live interview assistant that listens to questions and streams drafted answers fast enough to use mid-conversation.
  • Smart prompt routing keeps time-to-first-token low, with answers framed from the candidate's own résumé context.
  • Ships across browser, macOS, Windows and mobile from one codebase, with a marketing site, pricing, and docs.
Next.jsStreaming LLMsReal-Time AudioPrompt RoutingDesktop App
www.osprix.com
Osprix screenshot

Osprix

Aviation Parts & Supplies E-Commerce — Live on Custom Domain

LIVE

6+

Categories

osprix.com

Domain

  • Full aviation-supply storefront — hangar supplies, engine parts, avionics, sealants, lubricants and adhesives — serving the aviation maintenance community.
  • Catalog search with category faceting, brand browsing, product detail pages, cart and support flows.
  • Shipped on a real custom domain with live support chat and a staged pre-launch ordering gate.
Next.jsE-CommerceSearch & FacetingCart & CheckoutCMSCustom Domain
crm-two-zeta-33.vercel.app
G.O.A.T CRM screenshot

G.O.A.T CRM

Sales CRM SaaS for an Elevator & Escalator Company

LIVE

6

Pipeline stages

Web + iOS + Android

Platforms

  • Full sales-pipeline CRM — leads → site visit → quote → negotiation → won/lost — with dashboard analytics, reminders, and quick status changes.
  • Interactive 3D floating elevator scene on the login page built with react-three-fiber.
  • Domain engineering calculators: hoistway dimensions → rated capacity mapping from manufacturer tables, plus a price calculator.
  • JWT-auth FastAPI backend, Android & iOS builds from the same codebase via Capacitor, deployed on Vercel + VPS with a server-side API proxy.
ReactTypeScriptViteTailwindFramer MotionThree.jsFastAPICapacitor
claudeproject-hazel.vercel.app
Addy — Advisor Console screenshot

Addy — Advisor Console

Multi-Spoke AI Advisor with Human-in-the-Loop Review

LIVE

8

Agent spokes

Never

Auto-send

  • Internal advisor console: pick a client and Addy loads their brief, meeting notes and pipeline straight from Notion — no copy-paste.
  • Eight specialist spokes from marketing to document automation, coordinated by one advisor persona that drafts in a consistent voice.
  • Human-in-the-loop by design — Addy only drafts; every output is reviewed and released by an operator, nothing sends automatically.
Next.jsLLM AgentsNotion APIMulti-Agent RoutingHuman-in-the-Loop
gimme3d.vercel.app
gimme screenshot

gimme

Scroll-Driven 3D Food Ordering Experience

LIVE

Scroll-to-spin

Interaction

WebGL 3D

Render

  • A food-ordering site built as an experience — scroll drives a real-time 3D product spin instead of a static hero image.
  • Narrative scroll journey from craving → kitchen → on the way → delivered, with live order tracking framing.
  • Heavy WebGL work kept smooth with staged loading and a deliberate “enter the experience” gate.
Three.jsreact-three-fiberGSAP / ScrollWebGLNext.js
fyp-iota-nine.vercel.app
AdMotion screenshot

AdMotion

DOOH Advertising Platform — Final Year Project

LIVE

AI route-aware

Targeting

Live GPS

Tracking

  • Out-of-home advertising platform for car roof-mounted screens — advertisers create location-aware campaigns.
  • An AI engine decides which ad runs where, when, and on which vehicle, and generates campaign insights.
  • Separate advertiser and driver portals plus a full admin dashboard for vehicles, ads, campaigns, and analytics.
ReactFastAPIPythonAI TargetingGPS TrackingAnalytics

Activity

I ship every week

Live from my GitHub contribution graph — not a screenshot. Hover any square for that day.

Open Source

Everything public on GitHub

github.com/muneebanwar69 — 16 repositories and counting.

Skills

The stack

Click any module to expand it.

LangChainLangGraphMCPOpenAIClaudeGeminiGroqMulti-LLM OrchestrationPrompt EngineeringFine-TuningTensorFlowPyTorchscikit-learnOpenCVYOLOBERT / GPT / T5

Let's build something intelligent

Open to AI engineering roles, agentic systems work, and ambitious freelance builds. Pick a channel — I reply fast.