Stars
Official code repo for the O'Reilly Book - "Hands-On Large Language Models"
Machine Learning From Scratch. Bare bones NumPy implementations of machine learning models and algorithms with a focus on accessibility. Aims to cover everything from linear regression to deep lear…
This repo is meant to serve as a detailed guide for Machine Learning/AI interviews.
A toolkit to run Ray applications on Kubernetes
Official repository for DistFlashAttn: Distributed Memory-efficient Attention for Long-context LLMs Training
Development repository for the Triton language and compiler
QuantStart.com - QSTrader backtesting simulation engine.
Python wrapper for TA-Lib (http://ta-lib.org/).
AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform
Everything you want to know about Google Cloud TPU
Model parallel transformers in JAX and Haiku
🐙 Guides, papers, lessons, notebooks and resources for prompt engineering, context engineering, RAG, and AI Agents.
An implementation of model parallel GPT-2 and GPT-3-style models using the mesh-tensorflow library.
Easy-to-use data handling for SQL data stores with support for implicit table creation, bulk loading, and transactions.
Model parallel transformers in JAX and Haiku
gcpr / tutorials
Forked from graphcore/tutorialsTraining material for IPU users: tutorials, feature examples, simple applications
Intel-tensorflow / serving
Forked from tensorflow/servingA flexible, high-performance serving system for machine learning models
Example code and applications for machine learning on Graphcore IPUs
Provide all my solutions and explanations in Chinese for all the Leetcode coding problems.
Apache OpenWhisk is an open source serverless cloud platform
Triton Model Analyzer is a CLI tool to help with better understanding of the compute and memory requirements of the Triton Inference Server models.
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
The Triton Inference Server provides an optimized cloud and edge inferencing solution.