Highlights
- Pro
Lists (4)
Sort Name ascending (A-Z)
Starred repositories
High-fidelity, anycloud emulators running in your laptop. For DevOps programming, testing, and simulation.
[ICML 2026] effGen: Enabling Small Language Models as Capable Autonomous Agents
An opinionated, actionable guide for software engineering interviews.
Annotations of the interesting ML papers I read
Fractal Graph-of-Thought. Rhizomatic Mind-Mapping for Ai-Agents, Web-Links, Notes, and Code.
A bibliography and survey of the papers surrounding o1
Provide feedback and suggestions for Fairly AI's global AI regulations map.
Puzzles for learning Triton
Machine Learning Engineering Open Book
Opiniated RAG for integrating GenAI in your apps 🧠 Focus on your product rather than the RAG. Easy integration in existing products with customisation! Any LLM: GPT4, Groq, Llama. Any Vectorstore: …
Implementation of papers in 100 lines of code.
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
An ultimately comprehensive paper list of Vision Transformer/Attention, including papers, codes, and related websites
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama mode…
An attempt to answer the age old interview question "What happens when you type google.com into your browser and press enter?"
One stop shop for running AI/ML on AWS.
A high-throughput and memory-efficient inference and serving engine for LLMs
Large Language Model Text Generation Inference
This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc
✨✨Latest Advances on Multimodal Large Language Models
A curated list of Large Language Model (LLM) Interpretability resources.
A curated list of prompt-based paper in computer vision and vision-language learning.
🤖 Chat with your SQL database 📊. Accurate Text-to-SQL Generation via LLMs using Agentic Retrieval 🔄.
🔥Highlighting the top ML papers every week.
AutoAWQ implements the AWQ algorithm for 4-bit quantization with a 2x speedup during inference. Documentation:
S-LoRA: Serving Thousands of Concurrent LoRA Adapters
A universal scalable machine learning model deployment solution