Stars
A lightweight, powerful framework for multi-agent workflows
Open-source, low-cost 10.5 GHz PLFM phased array RADAR system
Baichuan-M3 Modeling Clinical Inquiry for Reliable Medical Decision-Making
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
Lets make video diffusion practical!
Leaderboard Comparing LLM Performance at Producing Hallucinations when Summarizing Short Documents
FlashMLA: Efficient Multi-head Latent Attention Kernels
A visuailzation tool to make deep understaning and easier debugging for RLHF training.
A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.
Small Models, Big Insights: Leveraging Slim Proxy Models To Decide When and What to Retrieve for LLMs (ACL 2024)
Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.
The paper list of the 86-page SCIS cover paper "The Rise and Potential of Large Language Model Based Agents: A Survey" by Zhiheng Xi et al.
f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
A framework to evaluate the generalization capability of safety alignment for LLMs
FaceChain is a deep-learning toolchain for generating your Digital-Twin.
A series of large language models developed by Baichuan Intelligent Technology
✨✨Latest Advances on Multimodal Large Language Models
A trend starts from "Chain of Thought Prompting Elicits Reasoning in Large Language Models".
fastllm是后端无依赖的高性能大模型推理库。同时支持张量并行推理稠密模型和混合模式推理MOE模型,任意10G以上显卡即可推理满血DeepSeek。双路9004/9005服务器+单显卡部署DeepSeek满血满精度原版模型,单并发20tps;INT4量化模型单并发30tps,多并发可达60+。