Generative AI · RAG · LLM applications
AI engineer in India — Deepak Kumar
I build AI features that reach real users, not demos that die in a notebook. Article generation, semantic search and voice automation, all running in production at a national news organisation.
Prod
AI shipped to readers
RAG
On Atlas Vector Search
80%
Faster podcast production
9+ yrs
In production

AI that reaches real users
At India Today Group — the publisher behind Aaj Tak — I build Generative AI into editorial products: article generation, AI summaries and semantic search over the archive, using OpenAI and LangChain with retrieval on MongoDB Atlas Vector Search. The hard part is never the model call — it is chunking tuned to the corpus, evaluation before rollout, and an editorial review step so nothing unchecked reaches a reader.
- AI podcast platform — articles to broadcast-ready audio, cutting production time roughly 80%
- RAG over an editorial archive, with retrieval quality measured rather than assumed
- Semantic search and AI summaries built into tools journalists use daily
- Structured prompting behind typed API contracts, so failures are catchable in code
How I engineer an AI feature
Like any other dependency: with a latency budget, a cost budget, a fallback when it fails, and monitoring that names the failing component. An LLM call is a network call to a probabilistic service — treating it as anything more magical is how teams end up with features they cannot debug or afford.
- Evaluation sets built from real queries before a feature is promoted
- Guardrails on output, plus a human review step wherever output is published
- Cost and latency tracked per call, alerting like any other service dependency
- Graceful degradation — the product still works when the model is slow or down
Full stack, so the AI actually ships
An AI engineer who cannot build the product around the model tends to hand over a prototype. I bring 9+ years of MERN and Next.js behind the AI work, which means the retrieval pipeline, the API, the dashboard and the deployment come from the same person.
Selected work
Shipped for employers and clients.
Seventeen products across news media, healthcare, real estate and adtech. A few with numbers attached.
Live election dashboard
India Today Group | Aaj Tak · 2024–25
Middleware that ingests results feeds from multiple sources, normalises them and publishes to editorial CMS platforms in real time, with a canvas-rendered constituency map on the front end.
Live results delivered to millions of daily users through election night, without a stall
AI podcast generation platform
India Today Group | Aaj Tak · 2025
Turns written news articles into podcast-ready audio — article processing, structured prompting, AI voice synthesis and an editorial dashboard to manage generation.
Cut podcast production time by roughly 80%
Patient relationship management
Clove Dental (via Instant Systems) · 2024–25
Internal PRM handling appointments, follow-ups and clinic-level analytics, built as modular services with REST APIs for the clinical front end.
In use across 500+ clinics and 1,200+ practitioners
VisitVideo and audio meeting platform
Humanize · 2023–24
Zoom-style real-time video and audio calling with screen sharing and meeting management, on a scalable signalling and media layer.
Real-time calling shipped end to end
VisitRecommendations
From people who shipped with me.
“Hard Working, Intelligent, Committed, Sharp and an Excellent team player are just a couple of words that can aptly describe Deepak. I worked with him for almost 3 years on the same project. He is technically very sound, always ready to learn new things, accept new challenges and the best part about him is that he alway…”
“We did a lot together and Deepak is really very talented, he learns new technology quickly and is very hard working. I felt very good after working with him, he is a person of very good personality.”
Straight answers
Hiring a AI engineer in India — the usual questions.
Deepak Kumar is a senior software and AI engineer in New Delhi with 9+ years of production engineering and a postgraduate qualification in Artificial Intelligence and Machine Learning. He builds Generative AI features with OpenAI and LangChain at India Today Group, including RAG on MongoDB Atlas Vector Search and an AI podcast platform that cut production time roughly 80%.
OpenAI and GPT models, LangChain, retrieval-augmented generation, embeddings and semantic search, MongoDB Atlas Vector Search, ElevenLabs voice synthesis, and Python with FastAPI for ingestion pipelines — all integrated into Node.js and Next.js applications.
Yes — that is the most common engagement. RAG over your own content, semantic search, content or voice automation, added to a working product with evaluation, guardrails, cost and latency budgets, and a fallback path when the model is unavailable.
For anything that ships, yes. Deepak brings 9+ years of MERN and Next.js engineering behind the AI work, so the retrieval pipeline, the API, the interface and the deployment are built by one person rather than handed between specialists.
Related
Other ways people search for this.
Hiring, or need something built?
Open to senior full-stack and AI engineering roles, and to focused contract work. Delhi NCR or fully remote. I reply within 24 hours.
What I take on
Full-stack product build
React, Next.js, Node, NestJS, MongoDB and MySQL — from idea to production.
AI features that ship
OpenAI and LangChain, RAG, semantic search, content and voice automation.
Architecture and scale
Microservices, real-time pipelines and caching — proven at national news scale.
Founding engineer work
Idea to MVP to scale, with SEO and analytics built in from the start.