RPG knowledge
Classes, armour, equipment, level progression, spell mechanics and encounter language.
See which AI models understand the work Game Masters actually need. Every ranking comes from a blind A/B vote on fantasy images, video, audio or writing.
Model names stay hidden until the vote is recorded.
Compare models on one medium at a time. Each board uses the same blind-vote method and a 1,000-point Elo baseline.
A polished image can still fail the task. The Arena tests whether a model understands fantasy terms, tabletop constraints and the details that make an output usable in play.
Read the full methodClasses, armour, equipment, level progression, spell mechanics and encounter language.
Gnolls, illithids and displacer beasts must look and behave like the creature in the prompt.
Battlemap geometry, cover, grids, handout legibility, scene continuity, music and TTS.
Adventure design, dialogue and lore are judged alongside style, clarity and prompt fidelity.
Two outputs receive the same prompt. You choose A, B, a tie or both fail before seeing the model names. Counted preferences update the Elo standings.
Pick two models or choose a random match within your Gold limit. Completed viable outputs can enter future blind battles.