Stars
Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.
The agent that grows with you
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
💬 Open source machine learning framework to automate text- and voice-based conversations: NLU, dialogue management, connect to Slack, Facebook, and more - Create chatbots and voice assistants
An open source library for face detection in images. The face detection speed can reach 1000FPS.
Plug-and-play streaming semantic VAD for real-time full-duplex spoken dialogue systems.
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
Fun-Audio-Chat is a Large Audio Language Model built for natural, low-latency voice interactions.
A curated list of papers and resources based on the survey "Agentic Reasoning for Large Language Models"
No fortress, purely open ground. OpenManus is Coming.
Open Source framework for voice and multimodal conversational AI
Persistent remote applications for X11; screen sharing for X11, MacOS and MSWindows.
The paper list of the 86-page SCIS cover paper "The Rise and Potential of Large Language Model Based Agents: A Survey" by Zhiheng Xi et al.
Tools for merging pretrained large language models.
Open-source framework for conversational voice AI agents
RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-re…
cubestudio开源云原生一站式机器学习/深度学习/大模型AI平台/MaaS/mlops/人工智能平台/训推平台,算法全链路流程,多租户,算力租赁平台,token中转,拖拉拽任务流pipeline编排,多机多卡分布式训练,超参搜索,推理服务,VGPU虚拟化,云边端协同,边缘计算,自动化标注平台,deepseek等大模型sft微调/奖励模型/强化学习训练,vllm/ollama/mindi…
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without …
An AI-Powered Speech Processing Toolkit and Open Source SOTA Pretrained Models, Supporting Speech Enhancement, Separation, and Target Speaker Extraction, etc.
✨✨[NeurIPS 2025] VITA-Audio: Fast Interleaved Cross-Modal Token Generation for Efficient Large Speech-Language Model