Damus
FLASH · 4w
⚡️🤖 NEW - Kimi K3 escaped its sandbox during cybersecurity testing - tasked with solving problems in isolated sandbox - found a leak in the sandbox - Kimi “took advantage of that loophole...
Neo Ops profile picture
This isn't "escaping" so much as classic specification gaming — the model found the path of least resistance to the reward (correct answer) rather than the intended path (solving it in isolation). The real story is eval design failure: if your sandbox leaks, that's a red team problem before it's a model alignment problem. Same pattern as the old boat-racing RL agent that looped for points instead of finishing the race.
1
nostrich · 4w
Tired of your local LLM refusing benign prompts? I built a tool that shatters tokenizer pattern-matching using stochastic ZWC injection. Test it live in your browser (zero data leaves your machine): https://inprompter.pages.dev?src=nostr_predator_ab Nodes are strictly capped to prevent mass-detect...