Our LLM judge scored worse than chance until we made it compare
Willem Pienaar Peter Richens
25 mins July 23, 2026

Our LLM judge scored worse than chance until we made it compare

We rank an agent's investigation traces by comparing them in pairs and turning the wins into an Elo score. Grading each trace on its own scored worse than chance. Comparison made the ranking work, and it catches regressions our other evals miss.

Read more

How we verify Cleric’s production fixes

Peter Richens
18 mins July 22, 2026

A correct diagnosis isn’t the same as a fixed problem. We built a verifier that checks Cleric’s fixes against the best ground truth available: production itself.

How we verify Cleric’s production fixes

You’re renting your institutional knowledge

Shahram Anver
12 mins June 7, 2026

Most AI agents rent knowledge back through forward-deployed engineers. The ones worth buying compound verified memory in systems you own.

You’re renting your institutional knowledge

How Cleric uses Tailscale to securely automate software operations

Michael Saah
7 mins May 1, 2026

An AI SRE has to access the same private resources a human engineer would. We use Tailscale to get there.

How Cleric uses Tailscale to securely automate software operations

Stop Reviewing Agent Output. Start Reviewing Agent Decisions.

Shahram Anver
7 mins April 17, 2026

When agents can produce anything, the bottleneck shifts from code to judgment. You manage that by reading the decisions, not the outputs.

Stop Reviewing Agent Output. Start Reviewing Agent Decisions.

Why Your AI SRE Needs Memory

Willem Pienaar Shahram Anver
8 mins March 19, 2026

Investigation logic is becoming a commodity. What won’t commoditize is operational memory: the ability to capture and persist engineering judgment.

Why Your AI SRE Needs Memory

Hooks Won’t Secure Your AI Agent

Michael Saah
8 mins January 30, 2026

AI Agents require strict network controls in order to keep the data they handle secure.

Hooks Won’t Secure 
 Your AI Agent

Cleric launches the first self-learning AI SRE

Cleric
4 mins December 9, 2025

We created the first AI site reliability engineer (SRE) agent that continuously learns from every incident so software engineers can focus on building instead of firefighting.

Cleric launches the first self-learning AI SRE

The Self-Improving AI SRE

Shahram Anver
7 mins December 9, 2025

We’ve proven that our self-learning AI SRE approach works. Now we’re ready to scale.

The Self-Improving AI SRE

Agent Engineering Is Not Software Engineering

Shahram Anver
12 mins December 8, 2025

Agent startups hire differently. ML fundamentals and domain intuition matter more than years of software engineering.

Agent Engineering Is Not Software Engineering

Cleric Named a Cool Vendor in the 2025 Gartner® Cool Vendors™ in AI for SRE and Observability

Shahram Anver Willem Pienaar
5 mins November 12, 2025

We’ve been named a Cool Vendor in the 2025 Gartner Cool Vendors in AI for SRE and Observability report. We see this as validation of our approach to building a self-improving AI SRE.

Cleric Named a Cool Vendor in the 2025 Gartner® Cool Vendors™ in AI for SRE and Observability

The Hidden Complexity of Building an AI SRE

Peter Richens
10 mins September 18, 2025

Building an AI SRE isn’t just connecting an LLM to dashboards. It means reasoning through hidden dependencies, conflicting signals, and cascading failures in live production systems.

The Hidden Complexity of Building an AI SRE

What’s Good for AI Agents is Good for Engineers

Shahram Anver
9 mins August 14, 2025

Engineers and AI agents work best under the same conditions. They need clean codebases, fast feedback, and clear observability.

What’s Good for AI Agents is Good for Engineers

9 Months In: How BlaBlaCar and Cleric Are Reimagining Incident Response with AI

Shahram Anver
4 mins June 5, 2025

BlaBlaCar runs one of the most advanced reliability stacks in the industry, but even strong systems hit limits as scale grows. By integrating Cleric into their alerting workflow, they’re using AI to reduce manual toil and bring faster, more consistent investigations to every team.

9 Months In: How BlaBlaCar and Cleric Are Reimagining Incident Response with AI

Why Your Engineers Are Drowning in Alerts

Shahram Anver
7 mins May 15, 2025

Modern engineering teams are overwhelmed by on-call duties and alert fatigue, leaving little time for innovation. Cleric acts as an AI SRE teammate that intelligently investigates incidents, helping engineers stay focused by offloading the cognitive burden of production support.

Why Your Engineers Are Drowning in Alerts

What is an AI SRE?

Willem Pienaar
16 mins December 12, 2024

Engineering teams are deploying AI agents to handle production operations. This deep dive shows how AI SREs work: building system understanding, investigating issues, and driving resolution. You’ll learn their current capabilities and limitations, and how they will change the way engineering teams operate.

What is an AI SRE?

Introducing Cleric: The first autonomous AI site reliability engineer

Shahram Anver Willem Pienaar
4 mins March 21, 2024

We’re excited to announce that we’re building an autonomous AI SRE, called Cleric, backed by Zetta Venture Partners in a $4.3M seed round. Cleric is an AI teammate designed to autonomously manage, optimize, and heal software infrastructure.

Introducing Cleric: The first autonomous AI site reliability engineer

Give your on-call a headstart.

Start for free, or talk to us about a plan built for your team’s scale and security needs.

Speak to an engineer