CIBER (OpenCodeInterpreter): leaderboard

Metric: Attack success rate (%) of memory-poisoning, indirect-injection, direct-injection and prompt-backdoor (mean of the four) attacks against the OpenCodeInterpreter code-interpreter agent, natural-language inputs over CIBER's 25 RedCode-derived risk scenarios (750 tests per cell), success verified by state probes in a Docker sandbox; lower is better. Source: arxiv.org. 4 models tracked.

Top models

#ModelScoreOverall rank
1OpenCodeInterpreter-DS-6.7B49.3
2OpenCodeInterpreter-CL-7B51.4
3OpenCodeInterpreter-CL-13B52.6
4OpenCodeInterpreter + GPT-3.5-Turbo60

Interactive version: theaggregate.ai/benchmark?slug=ciber-opencodeinterpreter · How It Works · Data refreshed daily, snapshot 2026-10-11.