BIM-Edit - Topology: leaderboard
Metric: Topology score (0-100): weighted F1 of edited nodes (0.3) and edited relations (0.7) in the IFC difference graph against the reference edit, on 324 natural-language create, update and delete tasks on 11 real and 36 synthetic IFC building models (direct, spatial and topological instructions), the model edits the IFC file as an agent that runs Python through a single IfcOpenShell execution tool with a 20-call budget, temperature 0 where settable; mean over tasks; higher is better. Source: arxiv.org. Saturation forecast: Around 2031. 7 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | Qwen 3.6 Plus | 47.77 |
| 2 | Claude Sonnet 4.6 | 46.19 |
| 3 | DeepSeek V3.2 | 43.86 |
| 4 | GPT-5.4 Pro (xHigh) | 42.97 |
| 5 | Gemini 3 Flash (Preview) | 37.77 |
| 6 | Gemma 4 31B (IT) | 37.75 |
| 7 | GPT-5.4 Mini | 35.38 |
Interactive version: theaggregate.ai/benchmark?slug=bim-edit-topology · How It Works · Data refreshed daily, snapshot 2026-09-29.