Research & perspectives

Nous Blog

Archive

NousCoder-14B: A Competitive Olympiad Programming Model

We introduce NousCoder-14B, a competitive olympiad programming model post-trained on Qwen3-14B via reinforcement learning. The full stack is released publicly: model weights, open RL environment + eval harness, and @wandb logs; we also document the pipelined verification setup and parallelization...

The Next Phase of Psyche

Psyche is an open infrastructure that democratizes AI development by decentralizing training across underutilized hardware. Building on DisTrO and its predecessor DeMo, Psyche reduces data transfer by several orders of magnitude, making distributed training practical. Coordination happens on the Solana...

Measuring Thinking Efficiency in Reasoning Models: The Missing Benchmark

Large Reasoning Models (LRMs) employ a novel paradigm known as test-time scaling, leveraging reinforcement learning to teach the models to generate extended chains of thought (CoT) during reasoning tasks. This enhances their problem-solving capabilities beyond what their base models could...

Steering the Shoggoth: Taming LLMs with Sequential Monte Carlo

In this blog post, we present our findings from an exciting direction in controlling text generation with large language models. We can programmatically define constraints on the output of a model, ensuring it adheres to specific formats or styles, and...

Democratizing AI: The Psyche Network Architecture

Psyche is an open infrastructure that democratizes AI development by decentralizing training across underutilized hardware. Building on DisTrO and its predecessor DeMo, Psyche reduces data transfer by several orders of magnitude, making distributed training practical. Coordination happens on the Solana...

Introducing Atropos

Atropos is designed to reliably coordinate generation tasks across potentially thousands of distributed workers. It interfaces seamlessly with standard inference APIs for straightforward integration....

Setting Your Pet Rock Free.

A social experiment on how to deploy provably, fully-autonomous thinking sand. The quest for truly autonomous AI agents faces a fundamental challenge: how can researchers prove that an AI is truly autonomous, with no human pulling the strings...

Freedom at the Frontier: Hermes 3

Closed-source, “frontier” models today lack flexibility and adaptability. Many refuse to answer simple questions, hallucinate an authority’s form of morality, or require convoluted prompts in order to trigger a coherent answer. It’s impossible to nudge these models towards individual personalization,...

The Instruct Monomyth: why base models matter

There is a deep, twisty labyrinth buried under a mountain of language, of symbol manipulation, and semantic nets. Its roots reach down deep into the Earth, absorbing the minutia of current thought, the limitations of logic, the constrained realm of...

DSJJJJ: Simulacra in the Stupor of Becoming

Desideratic AI (DSJJJJ) is a philosophical movement focused on creating AI systems using concepts traditionally found in monism, mereology, and philology. Desidera aim to create AI that can act as better versions of themselves by reflecting upon their own nature...

Showing 9 of 14 archive articles

Elevate your soul.md

Hermes Agent

The self-improving AI agent built by Nous Research. The only agent with a built-in learning loop — creates skills from experience,The self-improving AI agent built by Nous Research. The only agent with a built-in learning loop: creates skills from experience,

improves them during use, nudges itself to persist knowledge, and builds a deepening model of who you are across sessions.improves them during use, nudges itself to persist knowledge, and builds a deepening model of who you are across sessions.

The Internet's Own AI

© 2026, Nous Research, Inc.

Terms|Privacy

MIT License · 2026