-
University of California San Diego
- La Jolla
-
02:14
(UTC -07:00)
Lists (1)
Sort Name ascending (A-Z)
Starred repositories
𝐍𝐈𝐖 𝐀𝐩𝐩𝐫𝐨𝐯𝐚𝐥 𝐢𝐧 𝐎𝐧𝐞 𝐖𝐞𝐞𝐤 [Latex Template]
Generate editable scientific SVG figures from method text with local SAM3 and dual-provider routing.
Automatically crawl arXiv papers daily and summarize them using AI. Illustrating them using GitHub Pages.
Codebase of 'From Denoising to Refining: A Corrective Framework for Vision-Language Diffusion Model'
[NeurIPS 2025] Implementation for the paper "The Surprising Effectiveness of Negative Reinforcement in LLM Reasoning"
[EMNLP 2024 Findings] Benchmarking Language Model Agents for Data-Driven Science
Large Language Models Can Self-Improve in Long-context Reasoning
[2025-TMLR] A Survey on the Honesty of Large Language Models
[ACL 2023] Reasoning with Language Model Prompting: A Survey
[NeurIPS 2024] CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs
[EMNLP 2024] A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models.
[NeurIPS 2024] On the Worst Prompt Performance of LLMs
[ICLR 2025] ChartMimic: Evaluating LMM’s Cross-Modal Reasoning Capability via Chart-to-Code Generation
python数据可视化之美
Implementation of LREC-COLING 2024 paper A Frustratingly Simple Decoding Method for Neural Text Generation
Paper collections of the continuous effort start from World Models.
Repo for BenCao [original name: HuaTuo (华驼)], Instruction-tuning Large Language Models with Chinese Medical Knowledge. 本草(原名:华驼)模型仓库,基于中文医学知识的大语言模型指令微调
Official Repository for "Eureka: Human-Level Reward Design via Coding Large Language Models" (ICLR 2024)
Official Repo for ICML 2024 paper "Executable Code Actions Elicit Better LLM Agents" by Xingyao Wang, Yangyi Chen, Lifan Yuan, Yizhe Zhang, Yunzhu Li, Hao Peng, Heng Ji.
[ACL'24] Chain of Thought (CoT) is significant in improving the reasoning abilities of large language models (LLMs). However, the correlation between the effectiveness of CoT and the length of reas…
An Analytical Evaluation Board of Multi-turn LLM Agents [NeurIPS 2024 Oral]
Implementation of the training framework proposed in Self-Rewarding Language Model, from MetaAI
[EMNLP 2023] Question Answering as Programming for Solving Time-Sensitive Questions
Codes for "Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models".
[ICLR 2024] Lemur: Open Foundation Models for Language Agents