- London
- @parth_shukla
- in/parthshukla1
Stars
Lightweight coding agent that runs in your terminal
AI Crash Course to help busy builders catch up to the public frontier of AI research in 2 weeks
Copy playlists and liked music from Spotify to YTMusic
Full stack, modern web application template. Using FastAPI, React, SQLModel, PostgreSQL, Docker, GitHub Actions, automatic HTTPS and more.
A curated list of awesome things related to FastAPI
A list of free LLM inference resources accessible via API.
Repository of interview questions for Engineering Leadership roles - Engineering Manager, Director of Engineering, VP Engineering and also senior IC roles
NeMo Retriever Library is a scalable, performance-oriented document content and metadata extraction microservice. NeMo Retriever Library uses specialized NVIDIA NIM microservices to find, contextua…
Attestation and Secret Delivery Components
Model Context Protocol Servers
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
A collection of simple python mini projects to enhance your python skills
LLM Comparator is an interactive data visualization tool for evaluating and analyzing LLM responses side-by-side, developed by the PAIR team.
CUDA Templates and Python DSLs for High-Performance Linear Algebra
The Python Risk Identification Tool for generative AI (PyRIT) is an open source framework built to empower security professionals and engineers to proactively identify risks in generative AI systems.
Cloud-native high-performance edge/middle/service proxy
Meaningful control of data in distributed systems.
Reference code for creating and verifying a GCE firmware signed reference value message.
go-tdx-guest offers a library to wrap the /dev/tdx-guest device in Linux, as well as a library for attestation verification of fundamental components of an attestation quote.
go-sev-guest offers a library to wrap the /dev/sev-guest device in Linux, as well as a library for attestation verification of fundamental components of an attestation report.
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
S-LoRA: Serving Thousands of Concurrent LoRA Adapters
⚡ Build your chatbot within minutes on your favorite device; offer SOTA compression techniques for LLMs; run LLMs efficiently on Intel Platforms⚡
Always know what to expect from your data.
A high-throughput and memory-efficient inference and serving engine for LLMs
An awesome & curated list of best LLMOps tools for developers
A curated list of modern Generative Artificial Intelligence projects and services