Lists (8)
Sort Name ascending (A-Z)
Stars
A Collection on Large Language Models for Optimization
Technical guide to making money and investing(最全赚钱投资指南)
谷歌新书Agent设计模式(agentic design patterns)最佳中文版,持续优化。附:在线阅读、pdf和epub电子书下载。
小隐寺投资百科官方公开索引:美股、期权与加密货币知识框架
Sutskever 30 implementations inspired by https://papercode.vercel.app/ | For Agents, use https://github.com/pageman/Sutskever-Agent | Polyglot / Multi-Backed version at https://github.com/pageman/s…
Evolutionary algorithm that uses Large Language Models (LLMs) to automatically improve programs through iterative mutation and selection
Open-source implementation of AlphaEvolve
Companion webpage for the book "Bayesian Optimization" by Roman Garnett
Bayesian optimization in PyTorch
Repo for the Deep Reinforcement Learning Nanodegree program
Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch
Official implementation of GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization
High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)
基于多智能体LLM的中文金融交易框架 - TradingAgents中文增强版
[CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.
FlashMLA: Efficient Multi-head Latent Attention Kernels
The #1 AI Harness for Building Resumes, PDFs, Cover Letters & more, locally with 100+ LLMs support.
这是一个简单的技术科普教程项目,主要聚焦于解释一些有趣的,前沿的技术概念和原理。每篇文章都力求在 5 分钟内阅读完成。
Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
The open-source materials for paper "Sparsing Law: Towards Large Language Models with Greater Activation Sparsity".
A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.