TLDRocket
Sign in

Collective alignment: public input on our Model Spec

OpenAI

OpenAI asked over 1,000 people worldwide how AI should behave, then checked those answers against its own rulebook for ChatGPT. The surprise: most people already agree with OpenAI's defaults, but the disagreements reveal exactly where AI behavior gets politically and culturally messy.

Based on reporting by OpenAI — read the original for the full story.

Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error

OpenAI just published results from a survey it ran with more than 1,000 people across different countries, asking them to weigh in on how AI models should behave in specific, sometimes uncomfortable scenarios. The company then lined those responses up against its Model Spec, the internal document that defines default behavior for ChatGPT, and used the gaps to figure out where its own assumptions might be off.

This is part of what OpenAI calls collective alignment, an attempt to pull actual public opinion into the process of setting AI defaults instead of leaving those calls entirely to engineers and policy staff in San Francisco. The method isn't just a single opinion poll. Participants reacted to concrete situations, things like whether a model should refuse a request, hedge, or answer directly, and OpenAI tracked where the crowd's instincts matched the Spec and where they diverged.

The headline finding is reassuring for OpenAI: broad agreement with existing defaults was common. People largely backed the idea that models should stay helpful without being preachy, and that refusals should be rare and well-justified rather than reflexive. But the disagreements weren't noise. They clustered around genuinely hard territory, topics touching politics, self-harm, and questions where cultural background shapes what

My take — AI-written commentary, not fact-checked reporting

I like that OpenAI is showing its homework instead of just declaring what's 'safe' from behind closed doors, but a 1,000-person survey is a rounding error next to the billions of people these defaults will touch. Real legitimacy means opening this up at EU-regulator scale, not a nicely produced blog post that doubles as PR for the next alignment paper.

Read more about this at: OpenAI

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.