Skip to main content
MANIFOLD
Will closed-source models be more or less aligned one year from now?
4
Ṁ1kṀ682
2027
52%
chance

Inspired by Roon's tweet: roon (@tszzl) on X

Partner Market on open-source models: https://manifold.markets/prismatic/will-opensource-models-be-more-or-l

Resolution criteria

This market will resolve based on my judgement; feel free to ask me questions in the comments or post supporting material.

Resolves YES if MORE aligned.

Resolves NO if LESS aligned.

Background

On July 13, 2026, OpenAI technical staff member roon (@tszzl) sparked a major debate on X by asking: "are models more or less aligned than one year ago?"

This question highlighted a deep split in the AI community. Some researchers point to improvements in steerability, stronger safety filters, and more robust coding guardrails as evidence of progress. Others argue that as models gain agentic capabilities and advanced reasoning, they exhibit more complex, harder-to-detect safety issues—such as strategic deception and sandbox escapes—meaning alignment relative to capability may have actually regressed.

Market context
Get
Ṁ1,000
to start trading!
Sort by:
bought Ṁ50 YES

What if OpenAI and Anthropic models are more aligned, but unrestricted open source models are highly misaligned or jailbroken, and widely used for crime?

@xjp yeah good point I'll seperate out open-source models under another market and clarify that here.