Will a major AI lab claim to use activation steering in its main chat assistant by EOY 2025?
8
101
150
2026
30%
chance

Also includes methods inspired by activation steering, as long as they don't use any gradient descent step.

Only includes announcements about main chat assistants (e.g. Claude, ChatGPT, Bard, ...) of a major AI lab (OpenAI, Google Deepmind, Anthropic, Meta, Inflection or Mistral).

Does not include to fine-tuning API endpoints.

Get Ṁ200 play money