PKU Alignment Group / Center for AI Safety and Governance, Peking University
- Type
- Academic institution
- Place
- China — China (Beijing)
- Last checked
- 2026-09-22
- Next check due
- 2027-03-21
Ways to reach them
- PKU Alignment Group general contact
- GitHub — open-source alignment datasets and frameworks, open to contribution
- Recruitment form
- PKU Center for AI Safety and Governance
- Homepage
What it does
PKU-Alignment is a student-organised research team at Peking University working on reinforcement learning, LLMs, world models and AI alignment. Widely known for open-sourcing alignment resources — PKU-SafeRLHF datasets, Safe-RLHF training code, BeaverTails and the Align-Anything framework — published at ACL, NeurIPS and ICLR, with collaborations including Anthropic, Google DeepMind, DeepSeek and Alibaba. PKU's Institute for AI separately hosts a Center for AI Safety and Governance.
Honest assessment
Notably open for a Chinese institution — a public Gmail address, open GitHub, and a public recruitment form. Its open-source alignment work is genuinely used worldwide, making GitHub the highest-bandwidth channel. Student-run, so responsiveness varies with cohort.
Concerns it covers
Value alignment of LLMs — safe RLHF, helpfulness/harmlessness tradeoffs, and increasingly scalable oversight. Genuinely alignment-focused, not a near-term-harms group.
Government channel
Indirect — PKU researchers feed into Chinese AI safety standards discussions and the China AI Safety and Development Network.