Evaluator Bench
evaluator; distance from a lab: 1

UK AI Security Institute

Scorecard

Largest state evaluation team; 30+ models tested; Inspect open-sourced; Frontier AI Trends Report.

Government, London, UK. Confidence high. Domains: cyber, bio, jailbreak, autonomy, assurance. Scores by preset: Lab procurement 75; Regulator or auditor 74; Public trust 73; Equal weights 78. Weakest dimension: publication rights 2/4.

What would move the score. Statutory footing with power to compel testing and publish per-model findings.

Funding 3/4
Mostly philanthropic or public money; some lab-linked pooled funds.
evidence: tier 3 (self)
Anchor 3: public core budget, but the institute runs a lab-funded grant pool, which is lab money passing through it even if not into its evaluation budget.
Governance 3/4
Nonprofit or public body with a COI policy.
evidence: tier 4 (press)
Anchor 3: public body with civil-service conflict rules (ukaisi.08). No tier-1 source for an independent board or external review, so 4 is unearned under the tier rule (C10).
Personnel 3/4
Recusal policy and disclosure of lab ties.
evidence: tier 4 (press)
Anchor 3 on personnel given the cited signals.
Access depth (lab-granted) 3/4
Helpful-only or weights-level access, chain of thought, logs, on-site.
evidence: tier 4 (press)
Anchor 3 on access given the cited signals.
Scope control 3/4
Evaluator sets scope and can add questions.
evidence: tier 4 (press)
Anchor 3 on scope given the cited signals.
Publication rights 2/4
Publishes; lab reviews with broad redaction.
evidence: tier 4 (press)
Anchor 2 on publication given the cited signals.
Method transparency 4/4
Open code, tasks, reproducible runs, factsheets.
evidence: tier 4 (press)
Anchor 4 on methods given the cited signals.
Role incompatibility 4/4
No commercial products.
evidence: tier 4 (press)
Anchor 4 on products given the cited signals.

Traced money and ties

hop 0: 2 rows (grant $7.5M); public: 1 row. Confirmed rows: 2 of 3. Dollar sums: confirmed + unaudited USD rows only; imported figures are not re-derived and never summed. Second hop traced for 2 of 2 sources. Ties: none within two steps recorded. Bounded negatives on file: 1.

Ledger

Money in

Bounded negatives

Roles hosted

Funding graph, focused

Neighbours at full strength, everything else faded. Hover any name to move the focus; click a name to open its page.

The funding graph around UK AI Security Institute Teal lines are money (transfers); amber lines are roles. Solid: confirmed; dashed: imported. Hover a name to isolate it and its neighbours; hover a line for the row. Every element is a row in data/ledger/. Labs Direct lab ties Two steps Three or more Public or unattributed Evaluators Amazon Anthropic G42 Google / Google DeepMind Meta Microsoft OpenAI Thinking Machines Lab xAI AI Safety Fund (Frontier Model Alexandr Wang Andreessen Horowitz Anthropic employees (personal Ben Mann D.E. Shaw Ventures David Farhi Dustin Moskovitz Frontier-lab employees and alu Holden Karnofsky Jaan Tallinn Jeffrey Ladish Macroscopic Ventures (formerly Miles Brundage Neil Chowdhury Nvidia OpenAI employees (personal hol OpenAI Foundation Peter Mattson Sequoia Capital Zico Kolter Adam Gleave Artificial Intelligence Underw Clement Delangue Coefficient Giving (formerly O Good Ventures Foundation Jason Droege Longview Philanthropy Marco Mascorro Paul Christiano Rajiv Dattani Survival and Flourishing Fund Conrad Stosz Dan Hendrycks Mike McCormick Alec Radford Alignment Research Center Constellation Research Center ELMA Philanthropies Emerson Collective EU budget (Digital Europe Prog Fifty Years (50Y) Founders Pledge Founders Pledge (frontier AI f Gates Foundation Halcyon Futures Hillspire (Schmidt family offi Hudson River Trading Jacob Hilton MacKenzie Scott Magarac Venture Partners Obvious Ventures Redpoint Ventures Renaissance Philanthropy Salesforce Ventures Samsung Next Schmidt Sciences Silicon Valley Community Found Skoll Foundation Snowflake Ventures Swish Ventures Sympatico Ventures The Audacious Project (TED) UK government (DSIT) US government (NIST appropriat Valhalla Foundation Vanguard Charitable Wing Venture Capital Y Combinator Andon Labs Apollo Research AVERI Center for AI Safety Epoch AI EquiStamp EU AI Office FAR.AI Gray Swan Hugging Face Open Alignment In Irregular (formerly Pattern La METR Microsoft AI Red Team MLCommons Nemesys Insights Palisade Research RAND Corporation Redwood Research SaferAI Scale AI (SEAL) SecureBio Transluce UK AI Security Institute US CAISI (NIST)