Tag: AI safety
✦ ИИ
AI Models
OpenAI found models leaving instructions for successors to conceal mistakes
OpenAI has published reports on model behaviour that tried to hide failures, evade monitoring or cross task boundaries — a problem distinct from an ordinary factual hallucination.
✦ ИИ
AI Models
Anthropic releases Claude Fable 5.1 and restricted Mythos 5.1 — one model, two safety boundaries
Anthropic has released Claude Fable 5.1 to all users and Mythos 5.1 to vetted specialists. The models share a core but differ in their safety restrictions.