Skip to content
View rajveer43's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report rajveer43

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
rajveer43/README.md

Hi, I'm Rajveer Rathod πŸ‘‹

AI/ML Engineer Β· Building agentic platforms, KV-cache & inference systems, and physics-aware ML

Ahmedabad, Gujarat, India Β |Β  πŸ“§ rajveer.rathod1301@gmail.com

LinkedIn GitHub Medium Portfolio


πŸš€ About Me

I'm an AI/ML Engineer working on multi-tenant agentic platforms, internal ML infrastructure, and client-facing ML solutions. My technical interests span LLM fine-tuning, KV-cache optimization, physics-aware machine learning, and production systems engineering.

  • πŸ”­ Currently building CaloSR, a calorimeter super-resolution project for GSoC 2026 under ML4Sci / CMS (CERN open-source), using a GatedINR architecture
  • ⚑ Maintaining VeloxQuant-MLX, a KV-cache quantization library for Apple Silicon with 39+ eviction methods
  • πŸ”¬ Former HEP-ML researcher at Physical Research Laboratory β€” hypergraph neural networks for jet classification
  • 🌱 Active open-source contributor β€” 35+ ML repositories including PyTorch and Hugging Face Transformers
  • πŸŽ“ B.Tech in Computer Engineering, Birla Vishvakarma Mahavidyalaya (GTU), 2023
  • πŸ“œ Linux Foundation PyTorch Certified (LFS116)
  • ✍️ Write about ML systems & engineering on Medium

πŸ› οΈ Featured Projects

🎯 CaloSR

GSoC 2026 Β· ML4Sci / CMS Calorimeter super-resolution using a GatedINR architecture, trained on dual-T4 GPUs. Iterated through v1β†’v2 with split Fourier embeddings and a deeper occupancy head, backed by full MkDocs documentation.

⚑ VeloxQuant-MLX

KV-Cache Quantization for Apple Silicon A library implementing 39+ KV-cache eviction methods for efficient LLM inference on Apple Silicon, with versioned releases and PyPI analytics tracking.

πŸ“Š CallFlow Tracer

Python VS Code Extension Call graph visualization tool with 6,700+ PyPI downloads and 190+ VS Code installs.

πŸ”¬ Open Source Contributions

35+ ML repositories including PyTorch and Hugging Face Transformers, plus regular participation in open-source events like Open Source Summit India and Open Source Day.


πŸ’» Tech Stack

Languages Python C++ JavaScript

ML / AI PyTorch TensorFlow Keras HuggingFace

Backend / APIs FastAPI Flask Docker

Data PostgreSQL MongoDB

Tools Git Linux VS Code


πŸ“ˆ GitHub Stats


trophy


Thanks for visiting me

Pinned Loading

  1. llama llama Public

    Forked from meta-llama/llama

    Inference code for LLaMA models

    Python 1

  2. annotated_deep_learning_paper_implementations annotated_deep_learning_paper_implementations Public

    Forked from labmlai/annotated_deep_learning_paper_implementations

    πŸ§‘β€πŸ« 60 Implementations/tutorials of deep learning papers with side-by-side notes πŸ“; including transformers (original, xl, switch, feedback, vit, ...), optimizers (adam, adabelief, sophia, ...), gan…

    Jupyter Notebook

  3. anthropic-sdk-python anthropic-sdk-python Public

    Forked from anthropics/anthropic-sdk-python

    Python

  4. llm_tutorials llm_tutorials Public

    Tutorials on how to use language models

    Jupyter Notebook 1

  5. tensorflow tensorflow Public

    Forked from tensorflow/tensorflow

    An Open Source Machine Learning Framework for Everyone

    C++

  6. HinglishEase HinglishEase Public

    Work of Text Translation

    Jupyter Notebook 1