Stars
DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms
Tensors and Dynamic neural networks in Python with strong GPU acceleration
SWE-bench: Can Language Models Resolve Real-world Github Issues?
The fastest path to AI-powered full stack observability, even for lean teams.
个人制作的模型优化skills,但是因为对模型优化学习不深,效果并不好
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
分享 GitHub 上有趣、入门级的开源项目。Share interesting, entry-level open source projects on GitHub.
一款简单易用和高性能的AI部署框架 | An Easy-to-Use and High-Performance AI Deployment Framework
Wan: Open and Advanced Large-Scale Video Generative Models
A lightweight, production-ready C++ library for LLM tokenization, fully compatible with HuggingFace tokenizer.json.
A lightweight, single-header C++11 Jinja2 template engine for LLM chat templates.
A collection of practical, end-to-end AI application examples accelerated by MemryX hardware and software solutions. This repository offers examples for real-time video inference, object detection…
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…
The repository provides code for running inference with the Meta Segment Anything Model 3 (SAM 3).
Domain-specific language designed to streamline the development of high-performance GPU/CPU/Accelerators kernels
Ultralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
Multi-stream video inference with Ultralytics YOLO - Display multiple video streams in a grid layout with real-time object detection.
🤗 Optimum ONNX: Export your model to ONNX and run inference with ONNX Runtime
A high-performance tool for video upscaling, interpolation, depth estimation, and more. Available as a CLI and Adobe Extension.