MUNIM AHMAD — LAHORE, PAKISTAN
I build AI that ships.
Full-Stack AI Engineer building developer platforms, cloud infrastructure and AI systems that survive contact with production.
ABOUT — 01
I don't ship demos. I ship systems — RAG pipelines answering 500+ queries a day, inference cut by 65%, platforms that deploy themselves. Production is the only benchmark that counts.
SELECTED WORK — 02
Multi-tenant deployment platform that provisions infrastructure into user-owned DigitalOcean accounts via Pulumi — Next.js + NestJS control plane, encrypted credentials, BullMQ queues and live Socket.io deploy logs.
Open-source Python package that adds semantic caching to any OpenAI-compatible client — ONNX embeddings, sub-10ms cache lookups, cutting LLM costs 40–70% on FAQ-style workloads.
Open-source webhook inspector — hosted capture URLs, live request console, CLI tunnel to localhost, replay, and an MCP bridge that lets AI agents capture and replay webhooks directly.
The Artist.THE STACK — 03
One engineer.
Every layer.
IN NUMBERS — 04