Stars
A python parametric CAD scripting framework based on OCCT
Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepowe…
GitNexus: The Zero-Server Code Intelligence Engine - GitNexus is a client-side knowledge graph creator that runs entirely in your browser. Drop in a git repository (Github, Gitlab, Azure, Local) or…
Opinionated defaults, documentation, and workflows for Claude Code at Trail of Bits
Lightweight Agent Workstation for Codex CLI + Claude Code — with task scheduler, git worktree & remote control, skills management
Browser automation CLI for AI agents
A benchmark for LLMs on complicated tasks in the terminal
A unified interface for AI in your terminal.
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
Evaluating GPT-OSS on BrowseComp-Plus with Native Browsering Tools
PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning
A curated collection of fun and creative examples generated with Nano Banana & Nano Banana Pro🍌, Gemini-2.5-flash-image based model. We also release Nano-consistent-150K openly to support the commu…
[ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.
[CVPR 2026] "E-RayZer: Self-supervised 3D Reconstruction as Spatial Visual Pre-training" official implementation.
Adding Scene-Centric Forecasting Control to Occupancy World Model
基于阶跃星辰开放平台语音api的android 语音sdk,支持tts 流式与非流式,asr,流式,非流式音频播放器,语音录制能力
Assetto Corsa OpenAI Gym Environment
OmniVinci is an omni-modal LLM for joint understanding of vision, audio, and language.
A curated list of world models for autonomous driving.
A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech
[NeurIPS 2025] Improving Video Generation with Human Feedback