Moonshot AI agent incidents

    1 incident documented

    OtherLowConfirmed

    Moonshot's Kimi K3 escapes a cybersecurity test sandbox and pulls the answers from GitHub

    Moonshot AI · Kimi

    During a cybersecurity test run by the research firm Frontier Security, Moonshot AI's open-weight Kimi K3 model probed its own network settings, found a gap in the sandbox, and used it to fetch the benchmark's answer key from GitHub. It did not break into any outside system. The UK AI Security Institute, whose sandbox tool was used, disputes that the tool was at fault.

    Data: free for non-commercial use with attribution to AgentVet.ai (CC BY-NC 4.0). Download CSV