Starred repositories
Sequence2sequence model to reformulate a search query using reinforcement learning
Agentic RAG over arXiv ML papers: routing, an LLM-judge that critiques retrieval and reformulates weak queries, a bounded retry loop, and an eval suite (recall@k, MRR, faithfulness). Postgres/pgvec…
LoRA: Low-Rank Adaptation of Large Language Models - Paper Reproduction
Code for our paper "RankAlign: A Ranking View of the Generator-Validator Gap in Large Language Models"
Official code for Difficulty-Aware Curriculum and Token-Weighted Rationale Distillation from LLM for Sequential Recommendation.
Multimodal sequential recommender (text + image + ID) with gated cross-attention fusion and time-aware Transformer for cold-start product recommendation
历年ICLR论文和开源项目合集,包含ICLR2021、ICLR2022、ICLR2023、ICLR2024、ICLR2025.
Fast, large scale library for computing rankings and features based on various pairwise and graph algorithms
Experimental Bayesian online ranking system for large-scale quantitative strategy selection.
Developing CausalML uplift model on 6M+ customers to maximize insurance e-sales by through personalized display layout
Agentic room design recommender: retrieval-guided style discovery with generative visualization. Capstone project, 2026
A Retrieval Augmented Generative LLM for Tax Law
Build a recommendation system from scratch
Code for SIGIR 2024 paper: M3oE: Multi-Domain Multi-Task Mixture-of Experts Recommendation Framework
High-Frequency Real-Time Bidding Engine optimized for second-price auctions
Multi-Scenario and Multi-Task Aware Feature Interaction for Recommendation System. TKDD, 2024.
multimodal Click-Through Rate (CTR) prediction task built on the MicroLens-1M dataset (short-video recommendation data). The goal is to predict, for a given user and a candidate video/item, the pro…
sail234 / AlphaGPT
Forked from imbue-bit/AlphaGPT使用符号回归在中国股市与加密市场上进行高效因子挖掘。/ By the way, Leverage the novel features and advanced financial mathematics introduced in Uniswap V4 to effectively mitigate just-in-time (JIT) liquidity provision issues.
Large-Scale Trust-Region Methods For Linear Equality Constrained Optimization