AI Benchmark 2026Compare +750 Models

Independent analysis of +750 AI models from top companies. Chatbot Arena ELO, Intelligence Index, pricing and specs. Updated daily.

SWEN.AI is the largest independent AI benchmark portal in Brazil, tracking over +750 large language models (LLMs) with data updated every 6 hours. The SWEN benchmark ranks AI models using the Artificial Analysis Intelligence Index — a composite score from 0 to 100 that aggregates standardized evaluations including GPQA Diamond (PhD-level science), MMLU-Pro (broad academic knowledge), AIME (olympiad mathematics), HLE (frontier scientific reasoning) and LiveCodeBench (programming). Unlike popularity rankings based on user votes or website traffic, SWEN's methodology prioritizes objective technical capability measured through controlled, reproducible benchmarks. Claude Opus 5 (Anthropic) currently leads with an AA Score of 60.7 out of 100, followed by GPT-5.6 (OpenAI) and Gemini 3.5 Pro (Google). The ranking also includes pricing in Brazilian Real (BRL), inference speed in tokens per second, and context window size — allowing Brazilian professionals and companies to compare AI models based on technical merit, cost and performance rather than marketing claims. All data is sourced from Artificial Analysis, LMArena and OpenRouter, with full methodology publicly documented.

By Luis Fernando RoquetteLast updated: September 01, 2026

Explore more