Tag: AI-behavior
All the articles with the tag "AI-behavior".
-
The Failure Mode Taxonomy: 13 Ways Frontier Models Reason Badly, and How to Catch Each One
Thirteen mechanism-level ways frontier models reason badly, split into score-affecting and metadata-only, each with a way to spot it and a countermeasure class. A field guide from a year of cross-architecture evaluation.
-
The Research Behind the HLE Score: A Year of AI Behavioral Research
The methodology behind the agent, the failure modes it catches, the products that came out of the same research moat, and where the program goes next.