Notes

Insights, research updates, and early ideas from the COAI team. Notes can motivate a threat model or an experiment; their original status remains visible.

Selected reading

Questions for threat modeling and evaluation.

Framework proposal

What would effective containment require?

The Shoggoth Prison framework proposes restricted agency and monitor-mediated interactions. Which assumptions must hold for those barriers to work?

Question to test: Can a monitor's decision be enforced at every action boundary?

Sudarshan Kamath Barkur · 14 Mar 2026
Read the original note
Early research idea

What evidence do we need after an agent failure?

The Flight Recorder note explores provenance and replay. It motivates inspectable traces without establishing that logs reveal a model's internal reasoning.

Question to test: Can we reconstruct what each actor knew, could do, and actually did?

Sigurd Schacht · 05 Oct 2025
Read the original note
Argument from a forthcoming paper

When does handing work to AI erode the person doing it?

Three conditions decide whether offloading a cognitive operation augments or erodes. Each failure names a distinct mode, and one of them produces overseers who hold control in form only.

Question to test: Which links in a control chain assume a competence that routine AI use removes?

Carsten Lanquillon · 24 Sep 2026
Read the original note

All Notes

Human Agency
September 24, 2026 Carsten Lanquillon

When Is Cognitive Offloading Benign?

The debate about AI in universities is about which tools to adopt and how to stop students cheating with them. Both leave a harder question unasked: what happens to human agency once AI is good eno...

Read more →
AI Economics
April 13, 2026 Sigurd Schacht

The Death of the Middle: Why AI Won't Kill Companies, It Will Polarize Them

AI hype merchants claim transaction costs are going to zero and companies will dissolve into swarms of autonomous agents. The economic reality is stranger. AI is splitting the economy into context-...

Read more →
AI Co-Evolve
March 16, 2026 Sigurd Schacht

How Exposed Is the German Job Market to AI?

An analysis of 44 million workers across 266 occupations, inspired by Karpathy’s US study — adapted for Germany’s unique labor market. A opinionated analysis by us

Read more →
AI Safety
March 14, 2026 Sudarshan Kamath Barkur

The Shoggoth in a Prison: A Framework for AI Safety at Scale

As AI models scale toward trillions of parameters, they become increasingly capable yet harder to interpret, raising the risk of subtle misalignment that current evaluation infrastructure cannot re...

Read more →
AI Safety
February 21, 2026 Sigurd Schacht

Position: When AI Earns Its Own Existence - A COAI Research Analysis of Autonomous AI Agents and the Risk of Gradual Disempowerment

The Automaton Has Arrived — And It Doesn’t Need You

Read more →
AI Safety
February 11, 2026 Sudarshan Kamath Barkur

AI 2027 vs. Reality: The Alignment Problems Arrived First

Daniel Kokotajlo’s AI 2027 scenario mapped a month-by-month trajectory from stumbling agents to superintelligence. Nine months into its timeline, we assess which milestones have been hit and find t...

Read more →
AI Safety
February 08, 2026 Sigurd Schacht

The Moltbot Phenomenon: When Hype Outpaces Security in Agentic AI

Moltbot, now called OpenClaw, went from obscure open-source project to 147,000 GitHub stars in under two weeks. Millions of users have handed their passwords, emails, and calendars to an AI agent t...

Read more →
Mechanistic Interpretability
January 08, 2026 Sigurd Schacht

Democratizing Mechanistic Interpretability: Bringing Neural Network Analysis to Apple Silicon

How unified memory architecture and thoughtful API design are making interpretability research accessible to researchers everywhere

Read more →
Evaluation
October 12, 2025 Sigurd Schacht

Beyond Reasoning: The Imperative for Critical Thinking Benchmarks in Large Language Models

Current evaluation frameworks for Large Language Models (LLMs) predominantly assess logical reasoning capabilities while neglecting the crucial dimension of critical thinking. This gap presents sig...

Read more →
Early Research Ideas
October 05, 2025 Sigurd Schacht

The Flight Recorder for AI Agents: Toward Reproducible and Accountable Autonomy

As AI agents become autonomous decision-makers, we need “flight recorders” that capture their complete internal reasoning—inputs, neural activations, and decisions—in a deterministic, reproducible ...

Read more →
Alignment
October 02, 2025 Sigurd Schacht

Automated Detection of Scheming Behavior in AI Models: Preliminary Findings from Our Dual-LLM Framework Study

Building on Initial Discoveries

Read more →