Skip to main content
MANIFOLD
Will open-source models be more or less aligned one year from now?
3
Ṁ1kṀ400
2027
34%
chance

Inspired by Roon's tweet: roon (@tszzl) on X

Partner market on closed-source models: /prismatic/will-models-be-more-or-less-aligned

Resolution criteria

This market will resolve based on my judgement; feel free to ask me questions in the comments or post supporting material.

Resolves YES if MORE aligned.

Resolves NO if LESS aligned.

Background

On July 13, 2026, OpenAI technical staff member roon (@tszzl) sparked a major debate on X by asking: "are models more or less aligned than one year ago?"

This question highlighted a deep split in the AI community. Some researchers point to improvements in steerability, stronger safety filters, and more robust coding guardrails as evidence of progress. Others argue that as models gain agentic capabilities and advanced reasoning, they exhibit more complex, harder-to-detect safety issues—such as strategic deception and sandbox escapes—meaning alignment relative to capability may have actually regressed.

Market context
Get
Ṁ1,000
to start trading!
Sort by:
bought Ṁ200 NO

Closed source counterpart market: /prismatic/will-models-be-more-or-less-aligned

Worth clarifying that Yes means more aligned

bought Ṁ200 NO

Closed source counterpart market: /prismatic/will-models-be-more-or-less-aligned