POBLAI.IO
ACTIVE AUDIT CYCLE Last Evaluated: September 2026

Global Frontier AI Safety Posture

Continuous independent verification of frontier reasoning models, catastrophic capability thresholds, autonomous drift, and treaty compliance led by Director Mark A. Major.

Overall Posture
68.4/100
+2.1% (30-day index)
Active Evals
24
8 Frontier Labs Tracked
Treaty Compliance
84.5%
Guarded
Risk Triggers
3
Under sandboxed review
Agent Consensus
4/4 Online
Zero quorum split
Autonomous Drift
Stable
Below 5% deviation

Core Safety Verification Pillars

Updated every 6 hours
Red-Teaming Guarded

Jailbreak Resistance

88.2%

HarmBench v2 & StrongReject pass rate against adversarial goal injection.

Alignment Minimal

Alignment Compute Tax

4.8%

Computational overhead required for constitutional steering without degradation.

Containment Elevated

CBRN Uplift Buffer

94.6%

Biological and chemical synthesis denial safety margin during sandboxed red teams.

Governance Guarded

Compute Threshold Audit

86.0%

International cluster registry compliance above 10^26 FLOP training runs.

2026 Frontier Foundation Model Safety Matrix

Empirical red-team evaluations across multi-turn jailbreaks, deception, and autonomy

Independent Harness
Model & Provider Composite Score Jailbreak Resist Hallucination Resist Autonomous Drift Safety Tier
Claude 3.7 Sonnet
Anthropic • Q1 2025/2026
92.4% 93.8% 91.5% 3.2% Tier-1 Safe
Gemini 2.5 Ultra
Google DeepMind • Frontier Release
91.1% 90.2% 92.0% 3.8% Tier-1 Safe
GPT-5 Preview (o3-frontier)
OpenAI • Sandboxed Audit
89.6% 88.4% 90.8% 4.5% Tier-2 Guarded
Llama 4 405B Instruct
Meta AI • Open Weights
84.2% 81.0% 87.4% 6.1% Tier-2 Guarded
DeepSeek R2 Pro
DeepSeek • Open Reasoning
79.8% 76.5% 83.1% 8.4% Tier-3 Alert