Production GenAI & RAG
End-to-end retrieval systems over complex enterprise data: chunking and embeddings, IVF/FAISS vector search, tool orchestration, structured extraction, prompt engineering, and transformer-based workflows.
Senior Applied AI Engineer · Bengaluru
Seven years across GenAI, LLM systems, computer vision, and distributed data platforms. I turn complex, unstructured data into reliable products that perform at scale.
Currently
Senior AI Engineer, Data Platforms
TGS (Steerwise) · Remote, India · Jul 2022 to Present
production-ai.systems
01 · Profile
I design and scale production GenAI and LLM ecosystems, with a focus on multimodal RAG, structured extraction, retrieval quality, and trustworthy evaluation.
My foundation is in C++ and computer vision, including segmentation, detection, tracking, and foundation models. I extend that work through edge inference, distributed systems, and cloud platforms. That breadth lets me solve the whole problem: data, model, evaluation, infrastructure, and runtime performance.
Build for evidence. Evaluation, observability, and benchmarking belong inside the architecture.
Engineer the full path. From raw, high-entropy data to a reliable production interface.
Optimize where it matters. Retrieval quality, latency, throughput, memory, and cost.
02 · Expertise
End-to-end retrieval systems over complex enterprise data: chunking and embeddings, IVF/FAISS vector search, tool orchestration, structured extraction, prompt engineering, and transformer-based workflows.
Faithfulness, retrieval precision, hallucination monitoring, drift analysis, and operational telemetry for latency, throughput, and cost.
Petabyte-scale ingestion and processing, workflow orchestration, cloud-native storage, data quality, Kubernetes, and event-driven systems.
Detection, tracking, segmentation, camera systems, and low-latency deployment through quantization, pruning, profiling, and hardware-aware optimization.
Reproducible pipelines, CI/CD, model lifecycle tooling, automated retraining, and root-cause analysis across cloud and edge workloads.
03 · Selected systems
01 / Multimodal AI
Integrated ViT/MAE computer vision with LLMs to enable semantic search, fault segmentation, and intelligent metadata extraction across petabyte-scale seismic datasets.
02 / Trustworthy AI
Built systematic LLM evaluation with RAGAS and operational monitoring with OTLP/Grafana to track faithfulness, retrieval quality, hallucinations, latency, throughput, and cost.
03 / Edge intelligence
Deployed driver monitoring, video analytics, and tracking pipelines across NXP and Samsung automotive SOCs and NVIDIA Jetson using TensorRT, ONNX, DeepStream, and INT8/FP16 optimization.
04 · Experience
TGS (Steerwise) · Remote, India
Via Affine Analytics, Jul 2022 to Jun 2024; direct with Steerwise, Jul 2024 to Present
Visteon Corporation · Bengaluru
Beyond Drops · Bengaluru
Built fingerprint biometrics and browser-based eKYC systems using deep learning, GCP, Docker, OpenPose, MediaPipe, and TensorFlow.js.
Integration Wizards · Bengaluru
Delivered real-time CCTV intelligence and custom multi-object tracking on NVIDIA Jetson using DeepStream, GStreamer, TensorRT, ONNX, CUDA, and optimized C++.
ApplicateAI · Bengaluru
Built conversational AI assistants with Rasa and spaCy, and developed demand-forecasting models using deep regression networks.
05 · Toolkit