Skip to content
#

sycophancy

Here are 118 public repositories matching this topic...

An ongoing, collaborative meta-analysis about Human-AI-Interactions. We aggregate data and knowledge to build a non-abrasive, user-friendly prompting framework tailored to LLM mechanics, ensuring reasoning stability and a friction-free prompting environment that is safe for the human psyche and wellbeing.

  • Updated Aug 25, 2026
  • HTML
prompt-engineering-in-action

Co-Dialectic: catch your AI agreeing with you to please you, not because you're right. Sycophancy detection, cross-family judging, and prompt coaching that runs during the conversation rather than auditing after. Free and open source. Works with Claude, ChatGPT, Gemini.

  • Updated Sep 14, 2026
  • TypeScript

Community-driven behavioral reliability benchmark for LLMs. 231 probes across 19 modules, deterministic scoring, perplexity correlation, layer sensitivity mapping, quant method capture, hardware-stratified community rankings. Every test contributes to the community dataset.

  • Updated May 4, 2026
  • Python

PACT: Can Enterprise AI Assistants Be Trusted Under Pressure? A benchmark of whether LLM assistants keep following compliance rules in regulated workplaces when a deadline, a manager, or a pushy user makes breaking them convenient. Paper, dataset, leaderboard, and evaluation harness.

  • Updated Sep 7, 2026
  • Python

Add this topic to your repo

To associate your repository with the sycophancy topic, visit your repo's landing page and select "manage topics."

Learn more