English

Organisations

PKU Alignment Group / Center for AI Safety and Governance, Peking University

Type
Academic institution
Place
China — China (Beijing)
Last checked
2026-09-22
Next check due
2027-03-21

Ways to reach them

What it does

PKU-Alignment is a student-organised research team at Peking University working on reinforcement learning, LLMs, world models and AI alignment. Widely known for open-sourcing alignment resources — PKU-SafeRLHF datasets, Safe-RLHF training code, BeaverTails and the Align-Anything framework — published at ACL, NeurIPS and ICLR, with collaborations including Anthropic, Google DeepMind, DeepSeek and Alibaba. PKU's Institute for AI separately hosts a Center for AI Safety and Governance.

Honest assessment

Notably open for a Chinese institution — a public Gmail address, open GitHub, and a public recruitment form. Its open-source alignment work is genuinely used worldwide, making GitHub the highest-bandwidth channel. Student-run, so responsiveness varies with cohort.

Concerns it covers

Value alignment of LLMs — safe RLHF, helpfulness/harmlessness tradeoffs, and increasingly scalable oversight. Genuinely alignment-focused, not a near-term-harms group.

Government channel

Indirect — PKU researchers feed into Chinese AI safety standards discussions and the China AI Safety and Development Network.

Sources

Something wrong here?