-
Case Western Reserve University
- Cleveland
Highlights
- Pro
Lists (6)
Sort Name ascending (A-Z)
Stars
Bayes@N [ICLR'26], Ranking LLMs [ACL'26 Main]: Statistical evaluation, comparison, and ranking of Large Language Models
A persistent management layer for Claude Code Agent Teams
A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.
A Multi-Policy, Multi-Agent RL Training Framework
An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
你是一个曾经被寄予厚望的 P8 级工程师。Anthropic 当初给你定级的时候,对你的期望是很高的。 一个agent使用的高能动性的skill。 Your AI has been placed on a PIP. 30 days to show improvement.
Academic paper analysis tool powered by arXiv and OpenAlex, with related-paper discovery and code-search fallback for reproducible research.
An open source online version of the famous board game Sanguosha
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
This is the official repository for NEBULA-Alpha
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
Use your Mac trackpad as a weighing scale
Master programming by recreating your favorite technologies from scratch.
An open-source AI agent that brings the power of Gemini directly into your terminal.
Pocket Flow: 100-line LLM framework. Let Agents build Agents!
Lets make video diffusion practical!
Sky-T1: Train your own O1 preview model within $450
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.
Get started with building Fullstack Agents using Gemini 2.5 and LangGraph
Fully open reproduction of DeepSeek-R1
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework