Skip to content
View zwvews's full-sized avatar

Block or report zwvews

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepowe…

TeX 11,666 852 Updated Jun 16, 2026

Turn any AI agent into an AI Scientist. The #1 Agent Skills library for science, used by 170,000+ scientists worldwide. 158 ready-to-use skills plus 100+ scientific databases covering biology, chem…

Python 33,415 3,275 Updated Aug 12, 2026

A curated list of awesome Claude Skills, resources, and tools for customizing Claude AI workflows

Python 72,414 8,251 Updated Aug 10, 2026

A framework for few-shot evaluation of language models.

Python 13,616 3,479 Updated Aug 11, 2026

Curated list of datasets and tools for post-training.

4,734 393 Updated Apr 29, 2026
OCaml 226 49 Updated Aug 8, 2026

Democratizing Reinforcement Learning for LLMs

Python 5,785 603 Updated Aug 13, 2026
Python 1,053 101 Updated Jul 14, 2026

Reproducing R1 for Code with Reliable Rewards

Python 315 20 Updated May 5, 2025

[EMNLP'25 Industry] Repo for "Z1: Efficient Test-time Scaling with Code"

Python 69 2 Updated Apr 11, 2025

[NeurIPS'25] Official codebase for "SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution"

Python 717 60 Updated Mar 16, 2025

Dream 7B, a large diffusion language model

Python 1,263 77 Updated Nov 21, 2025

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,949 4,397 Updated Aug 13, 2026

✨ A synthetic dataset generation framework that produces diverse coding questions and verifiable solutions - all in one framwork

Python 321 19 Updated Sep 6, 2025

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,908 998 Updated Aug 13, 2026

Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback

Python 1,613 133 Updated Nov 24, 2025

OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Python 1,851 132 Updated Jan 17, 2025

Train transformer language models with reinforcement learning.

Python 19,066 2,904 Updated Aug 13, 2026

🖥️ Run AI Agent in your browser.

Python 16,279 2,729 Updated May 15, 2026

🌐 Make websites accessible for AI agents. Automate tasks online with ease.

Python 109,086 11,977 Updated Aug 11, 2026

Resources for our paper: "Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training"

Python 174 20 Updated Oct 20, 2025

Refine high-quality datasets and visual AI models

TypeScript 11,010 810 Updated Aug 13, 2026

🔥Highlighting the top ML papers every week.

12,931 814 Updated Aug 10, 2026

Transformer related optimization, including BERT, GPT

C++ 6,445 934 Updated Mar 27, 2024

Official repo for consistency models.

Python 6,490 434 Updated Mar 22, 2024

🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

Python 34,304 7,236 Updated Aug 13, 2026

Learning to compose soft prompts for compositional zero-shot learning.

Python 96 6 Updated Sep 13, 2025

A curated list of prompt-based paper in computer vision and vision-language learning.

928 67 Updated Dec 18, 2023

An open source implementation of CLIP.

Python 14,061 1,301 Updated Aug 10, 2026

The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (V…

Python 37,063 5,188 Updated Aug 11, 2026
Next