ABOUT

AI Engineer building systems that retrieve, reason, and act reliably.

I'm Anupam Kumar, an AI Engineer and Computer Science graduate from IIITM Gwalior ('25), working at the intersection of applied AI and backend engineering — RAG, agentic workflows, computer vision, and the infrastructure that takes them from demo to production.

I've built production RAG platforms for enterprise clients across legal, real estate, and knowledge-management domains, engineered FastAPI backends on AWS with Docker and Redis, and shipped document-intelligence pipelines that process real documents at high accuracy — systems tested against real usage, not just benchmarks.

I take on problems slightly beyond what I already know, break them, and come back understanding them properly — because that's how reliable systems get built.

Core Areas

Backend EngineeringAgentic AIGraph-RAGMulti-Agent SystemsLangGraphKnowledge GraphsFastAPIComputer VisionMLOpsNeo4jVector SearchAWSDocker

EXPERIENCE

Freelance AI Engineer

Independent Contractor

Remote

Apr 2026 – Present
  • Built multi-domain RAG retrieval systems for legal and enterprise knowledge bases using LangChain and FAISS, raising benchmark retrieval accuracy from 71% to 88% across a 25K+ document corpus.
  • Designed multilingual AI voice agents spanning 4 Indian languages, integrating voice AI, the WhatsApp Business API, and CRM workflows to automate scheduling and client follow-ups end-to-end.

AI Engineer — Applied AI & Backend Systems

Yahweh Innovations — AI Solutions & Consulting

Remote

Apr 2025 – Dec 2025
  • Owned end-to-end delivery of production RAG knowledge platforms for 5 enterprise clients, automating support and onboarding for 10K+ monthly AI requests.
  • Engineered multi-service FastAPI backends on AWS with Docker and Redis, adding async processing and CI/CD automation that doubled deployment throughput and cut release cycles to minutes.
  • Delivered an AI document-intelligence pipeline for lease audits, achieving 94.6%+ accuracy across 150+ documents and trimming turnaround time to 15 minutes.

Featured Projects

Swipe to explore my latest work

AetherCV

Production AI Research Platform for Computer Vision

End-to-end AI platform that combines semantic routing, hybrid retrieval, graph-enhanced reasoning, observability, evaluation pipelines, and production infrastructure to deliver evidence-grounded answers over computer vision research literature.

Key Achievements
Hybrid retrieval system combining FAISS, BM25, RRF Fusion, and Cross-Encoder Reranking
0.94 RAGAS Context Recall, 0.92 context precision and 89.8% grounded-answer benchmark performance
Seven-layer intelligent caching architecture using Redis and semantic cache matching
PythonFastAPIFAISSRedisPostgreSQLPrometheusGrafanaMLflowDockerNginxGroqRAGAS
AetherCV
94%
RAGAS Recall

LeadBoost AI

AI Lead Discovery & Intelligence Platform

Full-stack AI platform that automates lead discovery, enrichment, qualification, and personalized outreach through distributed processing, LLM-powered workflows, and scalable SaaS infrastructure.

Key Achievements
Built multi-tenant SaaS architecture with authentication, subscriptions, and RBAC
Engineered Celery + Redis distributed task queues for large-scale lead processing
Automated lead enrichment and AI-powered prospect qualification workflows
FastAPICeleryRedisPostgreSQLDockerReactLangChainPlaywright
LeadBoost AI
88.5%
Website Resolution success

Ayurgenix AI

Agentic AI Platform for Ayurvedic Knowledge Systems

Production-grade AI platform that transforms classical Ayurvedic literature into an intelligent knowledge system using multi-agent reasoning, semantic retrieval, and grounded LLM generation.

Key Achievements
Built an LLM-powered retrieval platform over 10,000+ medical knowledge documents, validated across 500+ evaluation queries using RAGAS-based evaluation frameworks
Implemented agentic retrieval workflows with query expansion, intent classification, reranking, and confidence-based response generation
Engineered Pinecone-powered semantic search achieving sub-second retrieval latency with improved context relevance and citation quality
FastAPIPineconePostgreSQLDockerPyTorchSentence TransformersLangSmithJWT
Ayurgenix AI
<500
ms Response Time

TalentForge AI

AI-Powered Job Application Orchestration Platform

End-to-end intelligent automation platform that combines LLM-powered career intelligence, browser automation, analytics, and workflow orchestration to streamline large-scale job application management.

Key Achievements
Designed a multi-layer architecture with Controller, Service, Storage, Platform, and Analytics components
Built LLM-based resume compatibility scoring and intelligent application prioritization workflows
Automated application execution using Playwright with anti-detection mechanisms and safety controls
PythonGroqPlaywrightStreamlitSQLitePandasAsyncIODocker
TalentForge AI
90%
Manual Effort Reduced

SentinelAI IDS

Real-Time AI-Powered Network Threat Detection Platform

Production-grade intrusion detection system that combines machine learning, anomaly detection, explainable AI, and real-time traffic analysis to identify cyber threats across enterprise networks.

Key Achievements
Built a hybrid detection pipeline combining Random Forest classification and Autoencoder-based anomaly detection across 39 network flow features
Implemented real-time packet capture, flow aggregation, and threat analysis using Scapy and WebSocket-based streaming architecture
Detected DDoS, Botnet, Brute Force, Infiltration, Port Scan, and Web Attack patterns with explainable AI insights powered by LIME
PythonTensorFlowScikit-LearnFlaskReactNode.jsMongoDBScapyWebSocketsLIME
SentinelAI IDS
95%
Detection Accuracy

View More

Check out more projects on GitHub

Visit GitHub

Skills & Expertise

Designing intelligent systems that combine LLMs, retrieval architectures, backend platforms, observability, and cloud infrastructure for production environments.

LLM Engineering

Building production LLM applications with retrieval, agent workflows, semantic search, model serving, and orchestration frameworks.

LangChainLangGraphCrewAIOpenAITransformersvLLMOllama

Retrieval Systems

Designing knowledge systems powered by hybrid retrieval, reranking, vector search, graph reasoning, and citation-grounded responses.

Neo4jFAISSPineconeBM25Hybrid SearchRRF FusionSemantic RouterRAGAS

AI Platform Engineering

Developing scalable backend systems, APIs, microservices, and workflow automation platforms for production AI applications.

FastAPIPostgreSQLRedisCelerySQLAlchemyREST APIsWebSocketsMicroservices

Computer Vision

Building OCR pipelines, document intelligence systems, image processing workflows, and deep learning applications.

PyTorchOpenCVOCRComputer VisionYOLOCNNsImage ProcessingDeep Learning

Cloud Infrastructure

Deploying containerized applications with CI/CD, cloud-native architectures, automation workflows, and scalable infrastructure.

DockerAWS EC2GitHub ActionsLinuxCI/CDNginxCloud Deployment

MLOps & Observability

Ensuring reliability through monitoring, evaluation, experimentation, telemetry, and continuous improvement pipelines.

MLflowPrometheusGrafanaRAGASModel EvaluationMonitoringTelemetry

Let's Connect

Open to AI engineering roles, Graph-RAG consulting, and technical collaborations.