Lists (3)
Sort Name ascending (A-Z)
Stars
Sample application that showcases Data Cloud, Agents and Prompts.
The AI that really does things. Any OS. Any Platform. The lobster way. 🦞
Awesome_Multimodel is a curated GitHub repository that provides a comprehensive collection of resources for Multimodal Large Language Models (MLLM). It covers datasets, tuning techniques, in-contex…
Central repo to connect and document components/repos needed for IOS stf support
from vibe coding to agentic engineering - practice makes claude perfect
Curated list of awesome Cursor Rules .mdc files
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
✨✨Latest Advances on Multimodal Large Language Models
Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
Agentic Design Patterns: A Hands-On Guide to Building Intelligent Systems by Antonio Gulli
谷歌新书Agent设计模式(agentic design patterns)最佳中文版,持续优化。附:在线阅读、pdf和epub电子书下载。
A Step-by-Step Implementation of Qwen 3 MoE Architecture from Scratch
[ACL 2024] MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
Codes for our paper "ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate"
Evaluate your LLM's response with Prometheus and GPT4 💯
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
The code and data for "MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark" [NeurIPS 2024]
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
Z-Bench 1.0 by 真格基金:一个麻瓜的大语言模型中文测试集。Z-Bench is a LLM prompt dataset for non-technical users, developed by an enthusiastic AI-focused team in Zhenfund.