Evaluator Bench

Money and ties, every evaluator

The ledger view: each evaluator's traced inflows by distance from a frontier lab, ties within two steps, and how much of the record has been re-derived from primary sources rather than imported. Read the confirmed column before the others; most rows are still leads. Rows and sources: data/ledger/. Rebuild with python -m bench exposure --json.

Two patterns worth reading off the matrix. First, the two organizations with the most complete self-disclosure, Transluce and AVERI, score lower on funding than several that disclose less; an evidence-based score rewards silence unless the empty cells are read as gaps, which is why the confirmed and second-hop columns sit beside the rows and why an unevidenced dimension renders as a dash rather than a number. Second, the same three or four funders (Coefficient Giving, the Survival and Flourishing Fund, the Audacious Project, the EU AI Office) sit behind most of the nonprofit evaluators, so a question about one evaluator's independence is often a question about one funder's.

Traced money by distance from a frontier lab, per evaluator Cell: number of inflow rows at that distance; the amount below is summed within the largest measure at that distance. Undisclosed amounts count as rows only. Right: board, advisor, donor, investor, or office ties within two steps of a lab, and the share of rows re-derived from sources (confirmed) rather than imported. Lab (0) Lab-tied (1) Two steps (2) Three steps (3) Public budget Donor not public TiesConfirmedChecked2nd hop METR 3recommendation $0.8M 3commitment $17.0M 6 4/6 6 4/5 RAND Corporation 1recommendation $1.0M 1commitment $38.0M 0 2/2 0 2/2 FAR.AI 1* 2recommendation $0.9M 1 2*grant $0.5M 0 4/6 0 5/6 Redwood Research 1daf_grant $1.3M 1 0 1/2 0 2/2 Irregular (formerly Pattern Labs) 1* 1investment $80.0M 1 2/2 1 2/2 Gray Swan 1* 1investment $40.0M 0 2/2 0 1/2 SecureBio 1* 1grant $17.2M 1recommendation $0.8M 0 3/3 0 3/3 Apollo Research 1* 1* 1 2/2 0 1/2 Epoch AI 4* 1daf_grant $0.6M 3grant $24.5M 0 4/8 0 7/7 Palisade Research 2grant $2.1M 0 1/2 0 2/2 Transluce 2* 1* 1 3/3 2 3/3 Andon Labs 1* 0 1/1 0 1/1 UK AI Security Institute 2*grant $7.5M 1* 0 2/3 1 2/2 US CAISI (NIST) 2grant $25.0M 0 1/2 2 1/1 EU AI Office 1 0 1/1 1 1/1 AVERI 1* 2* 1* 0 4/4 2 3/4 SaferAI 1* 1recommendation $0.3M 1 0 2/3 1 3/3 Scale AI (SEAL) 1investment $14.3B 1 1/1 1 1/1 MLCommons 1* 0 1/1 1 1/1 Hugging Face Open Alignment Initiati 1investment $12.9B 3 1/1 1 1/1 Center for AI Safety 2recommendation $1.4M 0 2/2 0 1/1 Microsoft AI Red Team 1* 0 1/1 0 1/1 Dreadnode no ledger rows yet; funding and ties not traced Humane Intelligence no ledger rows yet; funding and ties not traced Holistic Agent Leaderboard (Princeto no ledger rows yet; funding and ties not traced EquiStamp 1 0 1/1 1 1/1 Nemesys Insights 1* 1* 0 2/2 1 2/2 * includes undisclosed amounts. Evaluation credits (in-kind) are recorded on entity pages and never counted here. Checked: searches that found nothing, with their corpus and date. Source: data/ledger/. Generated by bench.figures.

Scroll sideways to see the rest of the figure.

The graph itself

Every entity in the ledger placed by its distance from a lab, with money in teal and roles in amber. Dashed lines are imported rows not yet re-derived. Hover a name to isolate its neighbourhood; click it to open the entity page and walk the graph hop by hop.

The funding graph, laid out by distance from a lab Teal lines are money (transfers); amber lines are roles. Solid: confirmed; dashed: imported. Hover a name to isolate it and its neighbours; hover a line for the row. Every element is a row in data/ledger/. Labs Direct lab ties Two steps Three or more Public or unattributed Evaluators Amazon Anthropic G42 Google / Google DeepMind Meta Microsoft OpenAI Thinking Machines Lab xAI AI Safety Fund (Frontier Model Alexandr Wang Andreessen Horowitz Anthropic employees (personal Ben Mann D.E. Shaw Ventures David Farhi Dustin Moskovitz Frontier-lab employees and alu Holden Karnofsky Jaan Tallinn Jeffrey Ladish Macroscopic Ventures (formerly Miles Brundage Neil Chowdhury Nvidia OpenAI employees (personal hol OpenAI Foundation Peter Mattson Sequoia Capital Zico Kolter Adam Gleave Artificial Intelligence Underw Clement Delangue Coefficient Giving (formerly O Good Ventures Foundation Jason Droege Longview Philanthropy Marco Mascorro Paul Christiano Rajiv Dattani Survival and Flourishing Fund Conrad Stosz Dan Hendrycks Mike McCormick Alec Radford Alignment Research Center Constellation Research Center ELMA Philanthropies Emerson Collective EU budget (Digital Europe Prog Fifty Years (50Y) Founders Pledge Founders Pledge (frontier AI f Gates Foundation Halcyon Futures Hillspire (Schmidt family offi Hudson River Trading Jacob Hilton MacKenzie Scott Magarac Venture Partners Obvious Ventures Redpoint Ventures Renaissance Philanthropy Salesforce Ventures Samsung Next Schmidt Sciences Silicon Valley Community Found Skoll Foundation Snowflake Ventures Swish Ventures Sympatico Ventures The Audacious Project (TED) UK government (DSIT) US government (NIST appropriat Valhalla Foundation Vanguard Charitable Wing Venture Capital Y Combinator Andon Labs Apollo Research AVERI Center for AI Safety Epoch AI EquiStamp EU AI Office FAR.AI Gray Swan Hugging Face Open Alignment In Irregular (formerly Pattern La METR Microsoft AI Red Team MLCommons Nemesys Insights Palisade Research RAND Corporation Redwood Research SaferAI Scale AI (SEAL) SecureBio Transluce UK AI Security Institute US CAISI (NIST)

Scroll sideways to see the rest of the figure.