English

Organisations

Redwood Research

Type
Civil society organisation
Place
United States — United States (Berkeley)
Last checked
2026-09-22
Next check due
2027-03-21

Ways to reach them

  • General enquiries · Email address
    Verified · 2026-09-22

    Anyone

    Note the unusual domain — rdwrs.com, not redwoodresearch.org.

  • Research engagement via public writing · Written submission
    Verified · 2026-09-22

    Researchers

    Redwood staff engage actively and publicly on the Alignment Forum and LessWrong — in practice the highest-bandwidth way to reach them with a technical argument.

  • Homepage · Homepage
    Unchecked

What it does

Non-profit research organisation that effectively founded the 'AI control' agenda: designing and red-teaming protocols that keep deployments safe even if the model is actively trying to subvert them. Known for control evaluations, monitoring protocols for malign LLM agents, the alignment-faking work with Anthropic, and helping developers build safety cases.

Honest assessment

Small and technically serious; will engage substantively with a good technical argument, and its researchers are notably responsive in public research forums. Will not engage with policy complaints or non-technical concerns.

Concerns it covers

Intentional misalignment and scheming — models that purposefully act against their developers' interests. Distinctively assumes alignment may fail and asks what safety measures survive that assumption. Loss-of-control, not misuse or near-term harms.

Government channel

Indirect — AI control has been adopted into frontier-lab safety frameworks and cited in safety-institute work; Redwood advises labs rather than governments.

Sources

Something wrong here?