Skip to content
View yumath's full-sized avatar
:octocat:
:octocat:

Block or report yumath

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A Collection on Large Language Models for Optimization

390 41 Updated Mar 31, 2026

Technical guide to making money and investing(最全赚钱投资指南)

2,552 470 Updated Aug 6, 2026

谷歌新书Agent设计模式(agentic design patterns)最佳中文版,持续优化。附:在线阅读、pdf和epub电子书下载。

HTML 7,665 1,101 Updated Jul 22, 2026

小隐寺投资百科官方公开索引:美股、期权与加密货币知识框架

JavaScript 3,262 212 Updated Aug 10, 2026

Sutskever 30 implementations inspired by https://papercode.vercel.app/ | For Agents, use https://github.com/pageman/Sutskever-Agent | Polyglot / Multi-Backed version at https://github.com/pageman/s…

Jupyter Notebook 4,279 560 Updated Mar 15, 2026
Jupyter Notebook 1,105 184 Updated Feb 5, 2024

Evolutionary algorithm that uses Large Language Models (LLMs) to automatically improve programs through iterative mutation and selection

Python 125 27 Updated May 26, 2026

Open-source implementation of AlphaEvolve

Python 7,190 1,132 Updated Jul 18, 2026

Python implementation of CMA-ES

Jupyter Notebook 1,344 199 Updated Aug 5, 2026

Companion webpage for the book "Bayesian Optimization" by Roman Garnett

HTML 950 54 Updated May 15, 2024

Bayesian optimization in PyTorch

Jupyter Notebook 3,584 494 Updated Aug 15, 2026

An Optimizer for Nvidia Compilers.

Python 128 12 Updated Aug 14, 2026

A collective list of free APIs

Python 460,837 50,915 Updated Aug 13, 2026

Repo for the Deep Reinforcement Learning Nanodegree program

Jupyter Notebook 5,178 2,373 Updated Jul 8, 2026

Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch

Python 2,373 424 Updated Jul 9, 2024

Official implementation of GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization

Python 499 35 Updated May 20, 2026

High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)

Python 10,274 1,153 Updated Apr 20, 2026

基于多智能体LLM的中文金融交易框架 - TradingAgents中文增强版

Python 31,162 6,558 Updated Jul 24, 2026

[CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.

Python 3,343 386 Updated Sep 7, 2025

FlashMLA: Efficient Multi-head Latent Attention Kernels

C++ 12,847 1,125 Updated Jul 28, 2026

The #1 AI Harness for Building Resumes, PDFs, Cover Letters & more, locally with 100+ LLMs support.

TypeScript 28,154 4,982 Updated Aug 11, 2026

这是一个简单的技术科普教程项目,主要聚焦于解释一些有趣的,前沿的技术概念和原理。每篇文章都力求在 5 分钟内阅读完成。

Python 6,979 606 Updated Mar 8, 2026

Nano vLLM

Python 15,018 2,464 Updated Apr 26, 2026

Mamba SSM architecture

Python 18,743 1,795 Updated Jul 22, 2026

Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.

Python 27,509 2,035 Updated Jan 9, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,919 1,003 Updated Aug 13, 2026

The open-source materials for paper "Sparsing Law: Towards Large Language Models with Greater Activation Sparsity".

Python 32 2 Updated Nov 12, 2024

A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.

6,899 369 Updated Dec 17, 2025
Next