-
Fudan University
- Shanghai City
- https://guox18.github.io/
- https://orcid.org/0000-0002-6510-8579
Stars
AI turns documents or topics into real, native PowerPoint decks—with native shapes, transitions and animations, data-backed charts and tables on demand, audio narration from speaker notes, and supp…
Native PJLab/CPMS print client for Apple Silicon macOS with PDF/image preview and batch printing.
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
Curated academic CV templates and guidelines for PhD students, researchers, and faculty job applicants.
2027 AI/ML internship & new graduate job list updated daily
Self-hosted AI agent harness in TypeScript. Serves multiple users over Feishu/Lark with sandboxed runtimes (local / Docker / GPU cluster), a durable task ledger, long-horizon memory, skills, and fi…
Writing AI Conference Papers: A Handbook for Beginners
飞书文档写作整合包 | Feishu Writing Bundle for OpenClaw — 从零创建到链接交付的完整飞书文档写作能力包
Overlapping Communication & Reconfiguration for Collective Communication in Optical Networks
Elevate your AI research writing, no more tedious polishing ✨
Enable seamless collaboration between Claude Code and Codex, transforming from a single agent to multiple agents for significantly enhanced productivity!
Search Google using scrapling and return structured results (title, link, snippet). Invoke when user asks to search Google or find information online.
Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports
IronClaw is an Agent OS focused on privacy, security and extensibility
A Model Context Protocol server for searching and analyzing arXiv papers
The hub for EleutherAI's work on interpretability and learning dynamics
Semantic Search & Call Graphs for AI Agents (100% Local)
Meta Lingua: a lean, efficient, and easy-to-hack codebase to research LLMs.
DSPydantic: Auto-Optimize Your Prompts and Pydantic Models with DSPy
Multi-harness agentic plugin marketplace for Claude Code, Codex CLI, Cursor, OpenCode, GitHub Copilot, and Gemini CLI
[CVPR 2026] This repository is the official implementation of MVGGT: Multimodal Visual Geometry Grounded Transformer for Multiview 3D Referring Expression Segmentation
Collection of handy online tools for developers, with great UX.
The official repository of the paper "From Detection to Association: Learning Discriminative Object Embeddings for Multi-Object Tracking"