AI ENGINEER · AGENT SYSTEMS
I build agent systems that run unattended, and the gates that stop them.
Multi-agent orchestration, retrieval grounding, and human-in-the-loop approval design, for teams whose LLM automation has to clear a security, compliance, or client sign-off before it ships.
OR READ THE CASE STUDIES9
AGENTS, SCOPED WRITES
314
COMMITS BY MY ORCHESTRATOR
12,060
DISCLOSED VULN REPORTS
9.8M
REGISTRY ROWS PER BUILD
$ factory run "cardiologists, US"[stream] NPPES rows ...... 9.8M[match] taxonomy exact 73,394[qa] 7 lenses, 2-of-3 rule[gate] compliance ..... PASS

Muneeb Shafiq
Associate AI Engineer at Symufolk · BS Data Science, Oct 2027
Open to work
01 · WORK · 4 FLAGSHIP · 8 SECONDARY
Case studies, not screenshots.
Each flagship is written up end to end: the problem, the architecture, the trade-offs, and exactly what I did. The rest get one line each.
- MULTI-AGENT SYSTEMSF01
Course Factory
Nine-agent course factory that ships instructor-approved curriculum behind mechanical human approval gates.
- Multi-agent orchestration
- Claude Code subagents
- Python
- AGENTIC DATA PIPELINESF02
Provider Research Factory
Turns a population query into a verified national dataset behind two human approval gates.
- Python 3.12
- Stdlib streaming CSV
- openpyxl, optional
- Claude Code subagents
- RAG / APPLIED SECURITYF03
RAMPART
Static-analysis findings grounded in a corpus of 26,681 real disclosed vulnerabilities before developers see them
- Google Gemini
- ChromaDB
- ONNX Runtime
- Semgrep
- PRODUCTION FULL-STACKF04
Symufolk HRIS
Production HRIS for Pakistani SMEs: encrypted personal data, per-client deployment, 123 endpoints
- FastAPI
- SQLAlchemy 2.0
- Alembic
- PostgreSQL
Secondary · 8
- S01Factor FoundryLLM-driven evolutionary factor search that reached a 0.826 public leaderboard score.LLM x QUANT RESEARCH
- S02Onboarding Pass-Rate ForensicsIsolated a 97%-to-59% onboarding pass-rate collapse to document image quality, clearing biometrics.DATA FORENSICS
- S03Loan SherlockScreens 50,000 loan applications for fraud and default risk using windowed transaction behavior.CLASSICAL ML
- S04Outreach AutopilotAgentic n8n pipeline that finds jobs, locates recruiters, and sends personalized outreach automatically.WORKFLOW AUTOMATION
- S05DocSpeakAsk questions of any PDF or Word file and hear the grounded answer.RAG PRIMITIVE
- S06Puffo Marketing StudioMarketers ship compliant, on-brand assets from plain language, no Git, Markdown, or HTML.AGENT GOVERNANCE
- S07Food Safety Audit DashboardTurns a 178-row audit spreadsheet into department, PRP and closure-performance analytics.BUSINESS INTELLIGENCE
- S08ElitePick AILive agency site where a Vite SPA ships crawler-ready HTML and self-submits to Google.WEB / SEO ENGINEERING
02 · HOW I WORK
Four rules the pipelines are built on.
- 01
Orchestrate deterministically.
Use models only where judgment is actually required. The control flow, the state transitions and the stopping conditions are code.
- 02
Gates are code, not policy documents.
If it can be skipped under deadline pressure, it was never a gate.
- 03
An unexercised guard is not a guard.
I run every guard at least once against a case it should catch.
- 04
Label the gap, let it surface.
When a step can't be completed, I say so rather than dropping it silently.
03 · SKILLS · 36 · EVIDENCE-LINKED
Nothing here that I can't open the file for.
Each skill links to the projects that use it. Counts come from the case studies, not from me.
Lime = core· ×N = projects
01 · ORCHESTRATION & AGENT CONTROL
7 skills02 · RETRIEVAL & GROUNDING
7 skills03 · GATES, PROVENANCE & AUDIT
6 skills04 · DATA & PIPELINES
11 skills05 · INTERFACES & DELIVERY
5 skillsNOT YET
No LLM-output evaluation, no GPU serving, no managed vector index at scale, no distributed tracing. The nearest one is measuring my own verifier. The harness is designed, answer keys anchored to verbatim source strings so a hallucinated location dies to a failed string search, a human adjudicator so an LLM never grades its own output, a held-out split touched once, and building it is what I'm doing next.
04 · EXPERIENCE
Where I've built, and for whom.
Symufolk
Lahore software house: custom AI builds and staff augmentation.
Apr 2026 – Present
Associate AI Engineer
FULL-TIMEJul 2026 – PresentPromoted from AI Engineer Intern after four months.
- Architected a nine-agent course production pipeline, producer, research analyst, curriculum architect, deck writer, assessment designer, lab engineer, QA reviewer, TA agent, video producer, turning a course topic into an instructor-approved package. Every agent has a written brief, a model tier and folder-scoped write permissions.
- Enforced a forward-only state machine with three human approval gates and a 940-line append-only approval log that records a real Gate 2 rejection, not only sign-offs. An 18-check QA reviewer blocks the final gate and is permitted to find defects, never to fix them.
- Wrote the founding architecture of a query-to-dataset factory, deterministic streaming pipeline, 11 declared assertions, a seven-lens adversarial QA panel in which each finding must survive two of three independent refutation attempts, and a compliance gate that can only permit or block, then ran the first end-to-end build by hand, producing a 22,195-record dataset identical to my hand-built reference. The system streams ~9.8M registry rows per national build and has delivered datasets of 409,848, 73,394 and 22,195 records.
- Python 3.12/3.13
- Claude Code subagents
- Claude Opus 5 / Sonnet 5 / Haiku 4.5
- Google Gemini
- RAG grounding packs
- FastAPI
AI Engineer Intern
INTERNSHIPApr 2026 – Jul 2026- Architected a modular crypto futures trading system as a seven-layer pipeline, data ingestion, feature engineering, market-state detection, multi-strategy signals, signal fusion, risk management, execution, with separation of concerns strict enough to hold backtest/live parity.
- Delivered two further production AI systems, Market Pulse and a Reconciliation Agent, on event-driven architecture with structured LLM output and database-level data-integrity constraints.
- Researched token usage, token economics and token-optimization strategies across the leading LLMs to guide cost-efficient model selection and routing.
- Python
- Event-driven architecture
- JSON-schema LLM output
- Backtest–live parity
- LLM token accounting
NETSOL Technologies
NASDAQ-listed Lahore firm; asset-finance and leasing software.
Sep 2026 – Dec 2026
AI/ML Trainee
TRAINEESHIPSep 2026 – Dec 2026Started 1 September 2026. Runs to 31 December 2026, alongside the Symufolk role. Outcomes will be added as the programme produces them.
Fiverr
Global freelance marketplace for digital services.
May 2025 – Present
8 CLIENTS · 5 COUNTRIES · 4.9/5 · VERIFY ON FIVERRAI & Data Analytics Freelancer, Level 1 Seller
FREELANCEMay 2025 – Present- Migrated 300 GB of reporting from Qlik Sense to Power BI, rebuilding the data models and the ETL/cleaning logic rather than porting them.
- Delivered eight publicly reviewed engagements, dashboard development, data analysis and data visualization, for clients in five countries (Pakistan, the UK, Canada, the UAE and the US), one of whom came back for a second order. Order values $50–$200, delivery 1–6 days.
- 4.9/5 across those eight reviews (seven 5-star, one 4-star), at Level 1 Seller, a platform-awarded tier, not a self-reported one.
- Power BI
- Qlik Sense
- SQL
- Excel
- Dimensional modeling
- ETL / data cleaning
Devsinc
Lahore software services firm serving international clients.
Jan 2025 – Present
Campus Ambassador
VOLUNTARYJan 2025 – Present- Led technical workshops on Python, SQL and AI/ML fundamentals for 100+ students.
- Arranged industry tours connecting students to working engineering teams.
- Python
- SQL
- AI/ML fundamentals
Manafa Technologies
Lahore technology arm of Manafa, a Saudi crowdfunding fintech
Jan 2026 – Jun 2026
Campus Ambassador
VOLUNTARYJan 2026 – Jun 2026
05 · WRITING · 3 POSTS
Notes from the build.
· 5 MIN
Two studies, opposite answers, and why both are right
One controlled trial found AI made developers 55.8% faster. Another found experienced developers 19% slower, while they believed they'd been sped up by 20%. The studies don't contradict each other, and the reason they don't is the useful part.
- llm-evaluation
- developer-productivity
- measurement
· 6 MIN
Ten things about Git that actually matter
Most Git tutorials teach you commands. After writing 80 pages of Git documentation, these are the ten ideas I'd keep if I had to throw the rest away.
- git
- engineering-practice
- developer-tools
· 8 MIN
Cheap by construction, unmeasured in practice
I designed two multi-agent systems to be cheap by construction, bounded batches, tiered model routing, grounding over cleverness. Then I went looking for what they actually cost, and found I had never recorded a single number. Here's what I found instead, and why it was the more interesting answer.
- llm-cost
- agents
- evaluation
06 · CONTACT
Open to work. Here's the fastest route.
Email is best, it forwards, it threads, and I answer it. If you're hiring, the role and the timeline are enough to get a useful reply.
I reply within one business day. Weekday mail sent before 4 PM in London arrives inside my working day. If you don't hear back, follow up: it got buried, not ignored.
Availability
Open to work, remote from Lahore (UTC+5, no DST). Core hours 1:00–10:00 PM PKT.
Overlap with your day
London
Full day
09:00 – 17:00
Berlin
7 hrs
10:00 – 17:00
full day Nov–Mar
Dubai
6 hrs
12:00 – 18:00
New York
3 hrs
09:00 – 12:00
4 hrs Mar–Nov
Engagement
- Independent contractor, invoicing from Pakistan. W-8BEN on file for US clients.
- No visa or sponsorship required, all work performed from Lahore.
- Comfortable onboarding through an EOR of your choosing, or your own contractor paper.
- Not seeking relocation before October 2027.
- Happy to take an unscheduled video call, ID in hand.



