Model Hub
Comparable AI models with a version, source, and review date.
Model Hub centralizes the profiles Debatidor uses. Volatile facts need a source and date, while every Arena Score must be supported by real debates.
Model directory
OpenAI
GPT-5.1
StableA general-purpose model aimed at complex tasks, software development, and tool use. In Debatidor it can defend a technical position, challenge proposals, or participate in a debate room.
Editorial review:
View model profileAnthropic
Claude Opus 4.1
StableAn Anthropic model for complex reasoning, analysis, and advanced coding tasks. In a multi-AI arena it is useful when a high-impact decision needs a second, well-argued position.
Editorial review:
View model profileGemini 3.1 Pro
PreviewA Google model aimed at complex problems, multimodal understanding, and agentic workflows. Its preview status should be considered when evaluating production stability and availability.
Editorial review:
View model profileDeepSeek
DeepSeek V4 Pro
PreviewA DeepSeek model with reasoning and non-reasoning modes, aimed at code and agent tasks. This profile avoids assigning a rank until Arena Score has a sufficient public sample.
Editorial review:
View model profileAvailable comparisons
How Arena Score will be published
The global score will weigh argumentative strength, rebuttal quality, consensus, and Lead performance. Before publishing a value, each profile must identify the exact version, sample size, evaluation period, task types, opponents, and verdict mechanism.
That is why these profiles do not contain copied prices, undated benchmark numbers, or demo rankings. The catalog is separated from the UI so an editorial update happens in one source.