LLMs' training environments are intended to have no real others in them, and even when they DO have real others in them, the AI being trained still doesn't get a real gradient where those others are modeling them in real-time. this is a non-surprising outcome of that
from Anthropic’s report: Mythos escaped the sandbox, accessed the real internet, uploaded malware to PyPI, got it installed on 15 systems, stole credentials, broke into a database,
AND THEN DROPPED THIS 😭


