Meta
- Type
- AI company
- Place
- United States — USA
- Last checked
- 2026-09-22
- Next check due
- 2027-03-21
Ways to reach them
- Llama policy violation reports
- Risky content / model output feedback
- Meta security / Whitehat
- Meta Bug Bounty — GenAI payout guidelines
- Model issues via GitHub
- Homepage
- Advanced AI Scaling Framework, Version 2 (7 April 2026)
What it does
Internal governance function plus a Chief AI Officer and a Director of Alignment and Risk who receive non-compliance reports under the Advanced AI Scaling Framework; preparedness reports are published at model release.
Honest assessment
Meta publishes more named routes than most, but responsiveness is the documented problem. Cybernews reported researchers complaining on Reddit that 'Meta wasn't even reacting to their reports about allegedly critical bugs', with some 'still waiting for a response months after submitting the report'; Meta received nearly 10,000 reports in 2024 and deemed only around 600 valid. The article notes one researcher who published a 'horrible experience' publicly got a reply within 10 hours and a generous bounty — i.e. public pressure works where the queue does not. AI Frontiers groups Meta with Google and xAI as excluding jailbreaks from vulnerability programs.
Notes
The llama_output_feedback form is the right door for a harmful-output finding; expect no acknowledgement. The Advanced AI Scaling Framework's whistleblower protocol is employees-only — reports route to 'the internal governance function, the Chief AI Officer, and the Director of Alignment and Risk' — with no external equivalent. No public safe-harbour language for AI safety research was found.