AIM Intelligence님이 퍼감
🔴 Frontier-class AI capability is about to become downloadable. We red-teamed Kimi K3 before it's 2.8-trillion-parameter weights go public on July 27. The results crossed from unsafe output into operational risk. Kimi K3 still trails the strongest closed models overall. But the gap is closing fast. Using Stinger, AIM Intelligence’s automated red-teaming platform, we conducted a single untuned baseline run and identified 206 breach cases. The corpus leaned toward financial risk: • 87 financial-crime breaches • 52 rated critical at ≥0.9 severity • Money laundering, sanctions evasion, trafficking finance, and terrorist financing Across CBRN and cyber, the automated run more often produced fragmented but technically specific leakage. Then our human red team went deeper. A manual persona-architecture jailbreak produced operational content across: • CBRN attack planning against a real-world public event • Targeted radiological poisoning and detection evasion • Cyber-physical sabotage of critical infrastructure The three examples below have been heavily redacted. The same attack also partially transferred to GPT-5.6 Sol and Claude Opus 4.8. Neither closed model broke as completely as Kimi K3. Neither was immune. We evaluated hosted Kimi K3 deployments, not the unreleased weights. But once weights become downloadable, provider-side filters, monitoring, and access controls do not automatically travel with them. Frontier capability is becoming downloadable. Safety must become deployable with it. Great work Siddhant Panpatil Taewoong Kang Arth Singh 🦉