Redwood Research
Scorecard
AI control research group; contributed a contractor to the METR incident investigation.
Nonprofit, Berkeley, US. Confidence med. Domains: scheming, incident, assurance. Scores by preset: Lab procurement 74; Regulator or auditor 73; Public trust 73; Equal weights 75. Weakest dimension: personnel 2/4.
What would move the score. A diversified funder list and a stated policy on co-authoring with labs it evaluates.
- No lab revenue; historically about $25M from Open Philanthropy.
Who Funds the AI Safety Watchdogs (Inside The Black Box (Substack), 2026-06-03) tier 4 pressconfirmed2026-09-14 - Funding concentration: revenue nearly disappeared in one year when a few large grants ended.
Who Funds the AI Safety Watchdogs (Inside The Black Box (Substack), 2026-06-03) tier 4 pressconfirmed2026-09-14 - Coefficient grants of about $63M across five grants dominate income; SFF about $2.4M and a $1.33M Tallinn gift; both funders sit one to two steps from an Anthropic investor.
evaluators.csv (24 evaluators: Coefficient and SFF totals, lab money, government contracts, leadership) (Kevin Bass (GitHub), 2026-09-14) tier 2 ledgerunverifiable Coefficient Giving grants index (Coefficient Giving, 2026-09-11) tier 1 indeximported2026-09-11 - Largest single grant: $36,566,000 from Coefficient Giving in November 2025 for AI control and alignment-faking work, after $10M in 2021 and $10M in 2022.
Redwood Research General Support (Open Philanthropy / Coefficient Giving, 2025-11) tier 1 indexconfirmed The nonprofit that investigated OpenAI's rogue agents runs on a $36m grant (TNW, 2026-09) tier 4 pressconfirmed2025-11
- Co-authored alignment research with Anthropic; close talent flow with lab safety teams.
Who Funds the AI Safety Watchdogs (Inside The Black Box (Substack), 2026-06-03) tier 4 pressconfirmed2026-09-14
- Contracted a staff member to METR for the on-site OpenAI investigation.
Brief independent investigation of agents' behavior in the OpenAI / Hugging Face hacking incident (METR, 2026-08-26) tier 3 selfconfirmed2026-09-14
- Sets its own research agenda; incident work under METR terms.
Brief independent investigation of agents' behavior in the OpenAI / Hugging Face hacking incident (METR, 2026-08-26) tier 3 selfconfirmed2026-09-14
- Co-authored the incident report with a redaction statement.
Brief independent investigation of agents' behavior in the OpenAI / Hugging Face hacking incident (METR, 2026-08-26) tier 3 selfconfirmed2026-09-14
- Publishes control methods and code.
Who Funds the AI Safety Watchdogs (Inside The Black Box (Substack), 2026-06-03) tier 4 pressconfirmed2026-09-14
- No commercial products.
Who Funds the AI Safety Watchdogs (Inside The Black Box (Substack), 2026-06-03) tier 4 pressconfirmed2026-09-14 - The OpenAI incident investigation report states no payment was accepted other than API credits used in the investigation.
METR and Redwood Offer Postmortem Of The HuggingFace Hack (Zvi Mowshowitz, 2026-08) tier 4 pressconfirmed2026-08
Traced money and ties
hop 1: 1 row (daf_grant $1.3M); hop 2: 1 row. Confirmed rows: 1 of 2 (2 quarantined: T04, T39). Dollar sums: confirmed + unaudited USD rows only; imported figures are not re-derived and never summed. Second hop traced for 2 of 2 sources. Ties: none within two steps recorded. Bounded negatives on file: 0.
Ledger
Money in
- T04 from Coefficient Giving (formerly Open Philanthropy): grant $63.09M; quarantined — excluded from sums, 2021-2025. cumulative incl. $36.6M general support Nov 2025 unverifiabletier 1 indexhttps://thenextweb.com/news/coefficient-giving-ai-safety-funding-ipo-correlation
- T38 from Jaan Tallinn: daf_grant $1.33M, 2021-11. direct via Founders Pledge US confirmedtier 2 ledgerhttps://jaan.info/philanthropy/donations.csv
- T39 from Survival and Flourishing Fund: recommendation $2.37M; quarantined — excluded from sums, 2022-2023. SFF recommendations unverifiabletier 3 selfhttps://survivalandflourishing.fund/recommendations
- T73 from Coefficient Giving (formerly Open Philanthropy): grant $36.57M; detail of T04 — not additive, 2025-11. AI control and alignment-faking work unauditedtier 1 indexhttps://www.openphilanthropy.org/grants/redwood-research-general-support/
Funding graph, focused
Neighbours at full strength, everything else faded. Hover any name to move the focus; click a name to open its page.