Highlights
- Pro
Lists (1)
Sort Name ascending (A-Z)
Stars
SIGIR'26 short paper "RankEvolve: Automating the Discovery of Retrieval Algorithms via LLM-Driven Evolution"
[ACM CAIS 2026] Improving Coherence and Persistence in Agentic AI for System Optimization
AI agents running research on single-GPU nanochat training automatically
ShinkaEvolve: Towards Open-Ended and Sample-Efficient Program Evolution 🧬
AdalFlow: The library to build & auto-optimize LLM applications.
🏆 Top-1 on 5+ benchmarks | Web UI | Supports MiroThinker, Claude, Kimi, OpenAI
Official repository of paper "Does Reasoning Introduce Bias? A Study of Social Bias Evaluation and Mitigation in LLM Reasoning"
🥇 Amazon Nova AI Challenge Winner - ASTRA emerged victorious as the top attacking team in Amazon's global AI safety competition, defeating elite defending teams from universities worldwide in live …
Get started with building Fullstack Agents using Gemini 2.5 and LangGraph
Open-source implementation of AlphaEvolve
RankLLM is a Python toolkit for reproducible information retrieval research using rerankers, with a focus on listwise reranking.
Minimal reproduction of DeepSeek R1-Zero
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
The code of SIGIR'25 accepted paper "ELOQ: Resources for Enhancing LLM Detection of Out-of-Scope"
Fully open reproduction of DeepSeek-R1
主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题
This is the repository that contains the source code for the Self-Evaluation Guided MCTS for online DPO.
[ICLR 2025] BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval
Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama mode…
Multimodal rumor detection model using an evidence-based dataset. Current version uses CLIP embeddings for both text and image inputs.
🤖 Multi-platform IM AI Agent for Telegram, WhatsApp, Lark, and WeChat. Connects ChatGPT / Claude / Kimi / DeepSeek / Ollama / Pi for auto-replies, community analysis, contact management, and inacti…
Taxonomy tree that will allow you to create models tuned with your data
DSPy: The framework for programming—not prompting—language models
🐙 Guides, papers, lessons, notebooks and resources for prompt engineering, context engineering, RAG, and AI Agents.
Awesome LLM Self-Consistency: a curated list of Self-consistency in Large Language Models