Shubham

  1. AI Engineer
  2. Voice
  3. RAG
  4. Agents
  5. LLM
  6. Hire for
  7. Work

AI Engineer

Shubham

I ship production AI for contact centres and warehouses. RAG, agents, voice, and the evals that prove they work.

ADPCX · Jun 2026 · Bengaluru · sole owner on the live-call agent

  • RAG
  • AI agents
  • LLM
  • Voice
  • STT
  • Evals
  • MLOps

Voice · STT · TTS

850ms

A live-call assistant. Speech in, grounded reply out. The person on the phone never types a search. Call forms fill from the conversation. Built as sole engineer and shipped on GPU Kubernetes in under 10 weeks.

20× speech engine · first words in 2s · one shared GPU service

  • STT
  • TTS
  • Voice AI
  • live call

RAG · Search

92%

Legal and ops teams ask in English. The system finds the right file, cites it, and keeps each customer’s files private. Ran on 67 legal agreements and later on 500+ document sets.

73% precision · hybrid search · rerank · citations

  • RAG
  • embeddings
  • hybrid search
  • evals

Agents

85%

Agents run warehouse playbooks and truck booking. A person still picks the slot. 32 workflows at 22 sites. About 3 seconds on the warehouse agent.

LangGraph · tool calling · human in the loop · 4.2/5

  • AI agents
  • LangGraph
  • HITL

LLM · Chat

1.5s

Same chat, faster answers. Production LLM serving: time to first word from 5.4s to 1.5s. Chat with memory from about 12s to 2s. Streaming, not a spinner.

structured output · token caps · FastAPI

  • LLM
  • streaming
  • latency
  • Text-to-SQL 82%

Hire for

Stack

What 2026 AI engineer and MLOps posts ask for. Only the terms I can walk through in an interview.

Python · Azure · Docker · Kubernetes · LangSmith

what JDs say

Build

  • Production AI
  • LLM
  • Generative AI
  • Prompting
  • Structured output
  • Streaming

search + RAG

Find

  • RAG
  • Hybrid search
  • Embeddings
  • Rerank
  • Chunking
  • Citations

agents

Act

  • AI agents
  • LangGraph
  • Tool calling
  • Human in the loop
  • Voice
  • STT

MLOps

Run

  • Kubernetes
  • Docker
  • Azure
  • vLLM
  • FastAPI
  • Evals

Python · LangChain · LangSmith · Hugging Face · Triton · Observability · Fine-tuning · Text-to-SQL

Work

ADPCXNewCold

ADPCX now: voice, RAG, LLM serving. NewCold before that: agents, legal RAG, booking, Text-to-SQL. MBA Business Analytics, batch topper. Reliance intern.

Jun 2026 – now · Dec 2024 – May 2026