I design and build production-ready AI systems that bridge research and real-world applications. My work spans Large Language Models (LLMs), Agentic AI, Retrieval-Augmented Generation (RAG), Vision-Language Models (VLMs), Computer Vision, and scalable AI infrastructure. I enjoy turning cutting-edge AI research into reliable, efficient, and deployable solutions.
- Languages: Python, Java, SQL
- AI/ML: PyTorch, TensorFlow, Hugging Face, OpenCV, YOLO
- Generative AI: Transformers, LangChain, LlamaIndex, RAG, Prompt Engineering
- Backend: FastAPI, REST APIs
- Infrastructure: Docker, Kubernetes, Linux, CUDA, Git
- Model Serving: vLLM, Ollama
- Generative AI & Agentic AI
- AI Research & Development
- Physical AI & Edge AI
- Computer Vision & Multimodal AI
- AI Infrastructure & MLOps
- High-Performance Computing (HPC)
- Distributed Training & Model Serving
- Production LLM inference with vLLM & Nvidia H200
- Agentic AI workflows and AI orchestration
- AI infrastructure for enterprise deployments
- Multimodal AI and Vision-Language Models