Discussion about this post

User's avatar
Marius Laurusevicius's avatar

Access depth shows up clearly in what providers publish about themselves. Anthropic's threat intelligence report of 10 September 2026 states that none of the misuse cases it found involved its Fable or Mythos-class models, with one distillation case as the exception. That is a claim resting on the provider's own telemetry, and an external evaluator holding only prompts and outputs could neither confirm nor contradict it. Statutory access to weights and training data would still not reach it. If independent evaluation is meant to cover deployed behaviour and not only model capability, log access belongs in the same sentence as weight access.

No posts

Ready for more?