-
Preferred Networks
- Remote Work in Japan
-
01:53
(UTC +09:00) - shyyhs.github.io
- https://orcid.org/0000-0003-1159-0918
- https://scholar.google.com/citations?user=IP5UyqcAAAAJ&hl=en
- in/haiyue-song-844a74186
- @shyoyhs
Highlights
Stars
Codex skill for full academic paper lifecycle analysis and revision
Fully autonomous & self-evolving research from idea to paper. Chat an Idea. Get a Paper. 🦞
Code for "From Context to Skills: Can Language Models Learn from Context Skillfully? "
My two cents on how interview questions should look like in AI era
OmX - Oh My codeX: Your codex is not alone. Add hooks, agent teams, HUDs, and so much more.
Synthetic pretraining data by rephrasing the web
MiroThinker is a deep research agent optimized for complex research and prediction tasks. Our latest models, MiroThinker-1.7, achieves 74.0 and 75.3 on the BrowseComp and BrowseComp Zh, respectively.
🏆 Top-1 on 5+ benchmarks | Web UI | Supports MiroThinker, Claude, Kimi, OpenAI
Gym-Anything: Turn any Software into an Agent Environment
A Model Context Protocol server for searching and analyzing arXiv papers
Faster way to switch between clusters and namespaces in kubectl
[NeurIPS 2025 Spotlight] Reasoning Environments for Reinforcement Learning with Verifiable Rewards
An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
foundation model plugin for Julius decoder
[ICLR 2025] Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing. Your efficient and high-quality synthetic data generation pipeline!
ACL 2026 - "Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded Updates"
Matrix (Multi-Agent daTa geneRation Infra and eXperimentation framework) is a versatile engine for multi-agent conversational data generation.
Flexible library for merging large language models (LLMs) via evolutionary optimization (ACL 2025 Demo).
[NeurIPS2025] "AI-Researcher: Autonomous Scientific Innovation" -- A production-ready version: https://novix.science/chat
ShinkaEvolve: Towards Open-Ended and Sample-Efficient Program Evolution 🧬
Pralekha: Cross-Lingual Document Alignment for Indic Languages
Official repository for paper "Structured Document Translation via Format Reinforcement Learning"
A collaboratively written review paper on deep learning, genomics, and precision medicine