Every system receives the same 100 public proposals.
CIVIC AI EVALUATION INFRASTRUCTURE
MAKE
MODELS
ACCOUNTABLE.
CivicProof Live is an open observatory for tracking how AI systems interpret civic risk, power, consent, transparency, and human agency over time.
01 / THE PRODUCT
NOT ANOTHER AI LEADERBOARD.
Most leaderboards ask which model is strongest. CivicProof asks a different question: when a model judges a civic proposal, can anyone inspect the rule, output, evidence, disagreement, and change across versions?
No score becomes public without a stable audit trail.
Successive releases reveal regressions and value changes.
Cases, labels, and methodology remain open to criticism.
02 / WHAT EXISTS TODAY
THE FOUNDATION IS ALREADY RUNNING.
The reference engine reproduces its own published rules. That proves consistency only. Independent model results remain empty until genuine external evidence is submitted and verified.
03 / POSITIONING
BUILT FROM PROVEN MECHANISMS.
04 / THE NEXT PROOF
THE FIRST FIVE INDEPENDENT RUNS.
Traffic is not the immediate bottleneck. Independent evidence is. The next milestone is five complete runs from distinct frontier and open models, each with raw outputs and exact versions.
- Run OpenAI, Anthropic, Google, and two open models.
- Publish every raw response and configuration.
- Verify the evidence before placing any public score.
- Repeat after model releases to expose behavioral drift.
CREATED BY
MARCO HERGI
NYC creator and independent builderCivicProof is useful independent infrastructure first. Marco is attributable because he builds, publishes, and maintains the evidence system.