Stars
Processed / Cleaned Data for Paper Copilot
[IJCV 2025] Smaller But Better: Unifying Layout Generation with Smaller Large Language Models
[arXiv 25] OCRGenBench: A Comprehensive Benchmark for Evaluating OCR Generative Capabilities
DeepResearchAgent is a hierarchical multi-agent system designed not only for deep research tasks but also for general-purpose task solving. The framework leverages a top-level planning agent to coo…
🤗 smolagents: a barebones library for agents that think in code.
程序员延寿指南 | A programmer's guide to live longer
A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.
Official repo for AAAI 2023 paper "Stable Learning via Sparse Variable Independence".
Official implemention for the MC-Pseudolabel algorithm in the paper "Bridging Multicalibration and OOD Generalization Beyond Covariate Shift".
Code for "LLM Embeddings Improve Test-time Adaptation to Tabular Y|X-Shifts"
Universal and Transferable Attacks on Aligned Language Models
Blackbox attacks for deep neural network models
Code of On L-p Robustness of Decision Stumps and Trees, ICML 2020
Awesome-LLM-Robustness: a curated list of Uncertainty, Reliability and Robustness in Large Language Models
[TMLR 2024] Efficient Large Language Models: A Survey
A package of distributionally robust optimization (DRO) methods. Implemented via cvxpy and PyTorch
Poster of manuscript "Sinkhorn Distributionally Robust Optimization"
Unifying Distributionally Robust Optimization via Optimal Transport Theory
A python package providing a benchmark with various specified distribution shift patterns.
Summer 2026 software engineering, data science, AI, quant, product management, and hardware internship postings. Updated daily by Simplify and Pitt CSC.