Lists (26)
Sort Name ascending (A-Z)
AIME
Benchmark
BKC
CUDA
Data
DeepSpeed
DLStreamer
DPC++
FFmpeg
GPU
Habana
HuggingFace
Level0
LLM
Media
Media Analytics
Nvidia
oneDNN
OpenCL
OpenVINO
Others
PyTorch
Tensorflow
Tools
Training
Workloads
Stars
Scalable toolkit for efficient model alignment
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance…
Production-tested AI infrastructure tools for efficient AGI development and community-driven innovation
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
LiBai(李白): A Toolbox for Large-Scale Distributed Parallel Training
Fast and memory-efficient exact attention
Accessible large language models via k-bit quantization for PyTorch.
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)
Notebooks using the Hugging Face libraries 🤗
The smart city reference pipeline shows how to integrate various media building blocks, with analytics powered by the OpenVINO™ Toolkit, for traffic or stadium sensing, analytics and management tasks.
Awesome-LLM: a curated list of Large Language Model
⚡ Build your chatbot within minutes on your favorite device; offer SOTA compression techniques for LLMs; run LLMs efficiently on Intel Platforms⚡
ChatGLM2-6B: An Open Bilingual Chat LLM | 开源双语对话语言模型
deepspeedai / Megatron-DeepSpeed
Forked from NVIDIA/Megatron-LMOngoing research training transformer language models at scale, including: BERT & GPT-2
Ongoing research training transformer models at scale
Simple OpenCL examples for exploiting GPU computing
A GPU benchmark tool for evaluating GPUs and CPUs on mixed operational intensity kernels (CUDA, OpenCL, HIP, SYCL, OpenMP)
Intel® Extension for TensorFlow*
C, C++ and Python Code for Exercises and Solutions