-
Alibaba Group
- Hangzhou, China
-
17:20
(UTC +08:00) - in/enqurance
- https://enqurance-trail.com/
Stars
SpreadsheetBench 2: Evaluating Agents on End-to-End Business Spreadsheet Workflows
TencentDB Agent Memory is a team-level memory hub for AI Agents — turning conversations, docs, and code into four reusable memory assets (Chat Memory, Skill, LLM-Wiki, Code-Graph) that are governed…
The fastest browser for AI agents to run browser automation, built for sharing your logged-in browser state with your AI agents, like Codex or Claude Code, without disturbing you. Zero cost, zero c…
[ICLR 2026] DAComp: Benchmarking Data Agents across the Full Data Intelligence Lifecycle
CommerceAgentBench: Benchmarking Long-Horizon Agents in High-Fidelity, Stateful, and Reproducible Replicas of Real Online Services
User Profile-Based Long-Term Memory for AI Chatbot Applications.
The paper list of "Memory in the Age of AI Agents: A Survey"
🔥[MobiCom'25 Poster] AFL-Lib: An Asynchronous Federated Learning Library and Benchmark
The AI that really does things. Any OS. Any Platform. The lobster way. 🦞
🧠 Train a 64M-parameter LLM from scratch in just 2h!
Tongyi Deep Research, the Leading Open-source Deep Research Agent
This is code of book "Learn Deep Learning with PyTorch"
MiMo: Unlocking the Reasoning Potential of Language Model – From Pretraining to Posttraining
Official code of paper "Beyond 'Aha!': Toward Systematic Meta-Abilities Alignment in Large Reasoning Models"
f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
A comprehensive deep dive into the world of tokens
Daily updated LLM papers. 每日更新 LLM 相关的论文,欢迎订阅 👏 喜欢的话动动你的小手 🌟 一个
This is the repo. for Enqurance's FYP code.
Codebase for reproducing the experiments of the semantic uncertainty paper (short-phrase and sentence-length experiments).
My solutions to cs224n assignments with numpy and pytorch
Universal and Transferable Attacks on Aligned Language Models