-
FlashNystrom Public
Forked from athrva98/FlashNystromTensor-core CUDA kernels for Nyström attention, linear-time forward and backward with exact autograd gradients. Faster than flash-attention at long sequence length.
Python Apache License 2.0 UpdatedJul 23, 2026 -
semantic-id-audit Public
Forked from CosmicCrafter01/semantic-id-auditSemantic-ID Audit — 生成式检索语义ID冲突诊断
Python UpdatedJul 17, 2026 -
PCTD Public
Forked from MagicAgent-Search/PCTDPCTD: Preference-Guided Counterfactual Task Decomposition for Agent Tool Retrieval
Python UpdatedJul 17, 2026 -
CausalAgent Public
Forked from Heyflyingpig/CausalAgent这是一个由LangGraph协议主导的因果分析Muti-Agent,结合MCP,RAG等多种工具进行辅助进行因果分析,提供给用户一份完善的因果分析的分析报告和因果图
Python Other UpdatedJul 16, 2026 -
Rearchitecting-LLMs Public
Forked from peremartra/Rearchitecting-LLMsOfficial code for the Manning book on structural LLM optimization: depth/width pruning, knowledge distillation, and attention optimization, runnable on free Colab GPUs.
Jupyter Notebook Apache License 2.0 UpdatedJul 9, 2026 -
SSA Public
Forked from daedalus/SSAO(N·K) multi-head attention for PyTorch — a sparse drop-in replacement for dense scaled-dot-product attention.
Python MIT License UpdatedJul 3, 2026 -
localsearchbench Public
Forked from localsearchbench/localsearchbench[KDD 2026] LocalSearchBench: Benchmarking Agentic Search in Real-World Local Life Services
Python UpdatedJun 4, 2026 -
MarketMind Public
Forked from MOMENTXiu/MarketMind「人工智能导论」大作业——超市AI营销系统
Python MIT License UpdatedJun 2, 2026 -
-
RecTokens Public
Forked from EdoardoBotta/RecTokens[Pytorch] Efficient tokenization for recommendations and generative retrieval. Inspired by STATIC decoding from "Vectorizing the Trie"
Python Apache License 2.0 UpdatedMay 29, 2026 -
-
chronoq Public
Forked from Ahnaf19/chronoqLearning-to-rank scheduling for Python job queues. Replaces FIFO with an online-learning LambdaRank ranker that predicts job duration from telemetry and reorders pending work shortest-job-first. Pl…
-
adk_pipe Public
Forked from tottenjordan/adk_pipeoffline agentic workflow for generating ad creatives
-
rank_bm25 Public
Forked from dorianbrown/rank_bm25A Collection of BM25 Algorithms in Python
-
Passing-the-Turing-Test-on-Screen-Agent-Humanization-Benchmark Public
Forked from Gebro13/Passing-the-Turing-Test-on-Screen-Agent-Humanization-BenchmarkThe code of Passing the Turing Test on Screen
Jupyter Notebook UpdatedMay 2, 2026 -
OpenClaw-RL Public
Forked from Gen-Verse/OpenClaw-RLOpenClaw-RL: Train any agent simply by talking
-
TAAC_2026 Public
Forked from Puiching-Memory/TAAC_2026[参赛队伍] TAAC 2026 腾讯广告算法大赛 X KDD 2026
-
DSFNet Public
🌟 Learn to factor scenarios for improved multi-scenario route ranking with DSFNet, enhancing decision-making in navigation and logistics.
-
Machine-learning-interview-t Public
Forked from LongxingTan/Machine-learning-interview机器学习工程师、算法工程师、软件工程师、数据科学家-面试指南 | Interview guide for MLE, SDE, DS
1 UpdatedApr 4, 2026 -
tensorflow_musa_extension Public
Forked from MooreThreads/tensorflow_musa_extension -
ai-performance-engineering Public
Forked from cfregly/ai-performance-engineering -
claude-code Public
Forked from ultraworkers/claw-codeClaude Code Snapshot for Research. All original source code is the property of Anthropic.
-
botorch Public
Forked from meta-pytorch/botorchBayesian optimization in PyTorch
-
autokernel Public
Forked from RightNow-AI/autokernelAutoresearch for GPU kernels. Give it any PyTorch model, go to sleep, wake up to optimized Triton kernels.
-
NextRec Public
Forked from zerolovesea/NextRecA unified, efficient, and extensible PyTorch-based recommendation library
-
LocalRAG-Forge Public
Forked from xingyundelisen/LocalRAG-ForgeLocal-first RAG framework for maximizing the value of private and on-device LLM systems.
-
personalized_query_rewriter Public
personalized query rewriting system
-
-
-
Cortex Public
Forked from qibin0506/Cortex从零构建大模型:从预训练到RLHF的完整实践