- 🔭 Forward Deployed Engineer at Future Path: I embed with customer teams, find where AI actually fits, and ship it with evals and guardrails
- 🌱 Contributing to open-source LLM routing and inference infrastructure: semantic-router, AIBrix, llm-d, llmgateway
- 💬 Ask me about RAG, agents, LLM evaluation and guardrails, and LLM gateways
- 👯 Happy to collaborate on open-source LLM infrastructure
- ⚡ Fun fact: I turn coffee and prompts into production code
- 📫 Reach me at sharmaalok705@gmail.com
Merged pull requests in other projects:
| Project | PR | What it fixed |
|---|---|---|
| vllm-project/semantic-router | #4759 | Accept Ollama's top-level timings object on streamed chat chunks, so a successful stream no longer ends in an error frame |
| vllm-project/semantic-router | #4631 | Keep credential failures (401/500) intact on the looper dispatch path instead of a generic 500 |
| vllm-project/aibrix | #2846 | Enforce the prefix-hash table's maxContexts cap and evict expired pods |
| vllm-project/aibrix | #2832 | Key single-port pods as pod/port in least-request routing |
| theopenco/llmgateway | #3960 | Clamp auto reasoning effort in the gateway |
| llm-d/llm-d | #2599 | Fix broken relative links in the docs |
Languages
AI / LLMs
Backend & Web
Data & Infrastructure

