Rapidata SVG Benchmark just landed on ModelScope, comparing 30 frontier LLMs on static SVG generation from text prompts, with 1.35M+ human votes across preference, coherence, and prompt alignment. 🚀
🤖
📊 Scale: 188,754 head-to-head comparisons, 500 prompts, 14,872 rasterized SVG images, and 1,355,161 human responses
🎨 Evaluation target: raw SVG markup generated by LLMs, rendered to 768x768 PNGs, then ranked by humans instead of automated metrics
🏆 Overall ranking: Claude Fable 5 Thinking leads with 1232.9 ELO, followed by Claude Fable 5 and Gemini 3.1 Pro Preview
License: CC-BY-4.0 for the benchmark prompts, with generated outputs governed by each model provider's terms.
As “draw an SVG pelican riding a bicycle” nears saturation we must expand to harder variations—like an SVG pelican on a bicycle chasing you through the Backrooms.
Here’s Claude Fable 5 Max:
I gave Fable 5.1 the SVG of my blog logo and told it to make a Pixar x Grok Bot animation mashup and it did not disappoint.
It even added sound effects with Web Audio. Pointless, but fun!
GPT-6 Astra vs GPT-5.6 on the same SVG task (pelican on a bicycle):
Higher effort = cleaner drawings.
But look at the cost:
Astra Max → 63¢
Sol Max → 32¢
Terra Max → 26¢
Luna Max → 1.6¢
Same prompt.
Very different bills.