designed multi-agent review

The Council

A 33-seat oversight design whose independence must be demonstrated before it can govern any safety decision.

Status: DESIGN — not a live system. The latest published three-leg point experiment measured n_eff=1 at rho=1. The historical DR-0007 number is unbound because its cited result artifact is absent. The claimed resilience under failed or adversarial voters remains rejected; see the Refutation Ledger. Everything below is the design visualization, not measured production behaviour.

Core Technology

What is designed multi-agent review?

A supermajority (23 of 33) of the designed council is required for any safety decision, so no single agent decides the outcome. Effective independence is measured, not assumed.

Designed quorum: 23/33

The designed threshold is 70% agreement (23 of 33 agents). This quorum is a design target — its measured status is published on the Refutation Ledger (DR-0007).

Multi-Provider Diversity

The target roster spans multiple providers. The current experiment used three model lineages across two providers and behaved like one independent leg.

Fail-Safe Operation (designed)

A 23-of-33 threshold is only arithmetic. It does not provide a failure-resilience guarantee unless voter independence, identity, transport, and signatures are real.

Simulated — not a live council

Design visualization. Measured status: DR-0007

GGGGGGGGGGGAAAAAAAAAAASSSSSSSSSSS1/33voting
Approve: 1
Reject: 0
Escalate: 0
Designed distribution — not provisioned providers

33-seat design · latest n_eff 1.00 of 3 · 2 providers tested

The roster below is a target architecture, not provisioned infrastructure. The latest measured sample was fully correlated, so no capture-resistance claim is made.

GPT-4o

OpenAI

6agents

Claude 3.5

Anthropic

6agents

Gemini Pro

Google

5agents

Llama 3

Meta

5agents

Mistral Large

Mistral

4agents

Command R+

Cohere

4agents

Qwen 2

Alibaba

3agents
33

Designed Seats

23-seat target threshold

Agent Distribution

6
6
5
5
4
4
3
023 = Consensus Threshold33
Specialized Functions

Council Responsibilities

The proposed roster assigns seats to specialized review roles. These are design responsibilities, not provisioned agents or comprehensive live oversight.

Compliance Validators

11 designed seats

Designed to check evidence against selected regulatory mappings

Risk Assessors

8 designed seats

Designed to propose risk levels for accountable human review

Incident Analyzers

7 designed seats

Designed to help investigate reported incidents and possible causes

Watchdog Monitors

7 designed seats

Design target: scoped monitoring of registered systems; not live

Human-in-the-Loop

Human Oversight

In the proposed workflow, automated tools support scoped monitoring while accountable human reviewers retain judgment for critical decisions. That Council workflow is not live.

Review and approve high-stakes system decisions
Investigate escalated incidents requiring human judgment
Override automated recommendations in edge cases
Provide contextual understanding AI may miss
Ensure ethical considerations are properly weighed

Human + AI Collaboration

Routine Monitoring
AI Council

Design target for scoped monitoring — not a live 24/7 service

Risk Assessment
Both

AI proposes risk levels, humans validate high-risk cases

Critical Decisions
Human

Proposed enforcement actions would require authorised human approval

Design status — not a live council

Council Design Status

33
Designed Seats
23
Target Threshold
1.00
Latest n_eff
0
Published Council Votes

Inspect the Council design

Review the proposed seats, the evidence boundary, and the failed independence experiment. No membership, certification, or live Council service is offered here.