Azinove AIEarly access

Model comparison

Compare open-weight AI models

This page compares 59 open-weight AI models, from text and code to vision, speech and embeddings. Benchmark results come from Epoch AI and LMArena, model details from Hugging Face, and every figure names its source and the day it was read.

Go to the table

Data as of Sep 24, 2026

Filter and sort

59 of 59 models shown

Open-weight AI models with public benchmark results and model details. Select up to four rows to compare them side by side. Scroll sideways to see every column.

Open-weight AI models with public benchmark results and model details. Select up to four rows to compare them side by side.
CompareModalityLicenceOrigin
Kimi K3moonshotai
  • Text (LLM)
  • Vision
kimi-k3Publisher’s terms2.8T1,48593.1%97.2%No value published by the source for this model50.6%No value published by the source for this model1,660No value published by the source for this model1M1.8MChina
GLM-5.3-Flashzai-org
  • Text (LLM)
  • Vision
mitPermissive321.3B1,47590.2%93.9%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,6121,2991M4MChina
GLM-5.2zai-org
  • Text (LLM)
mitPermissive753.3B1,47291.9%86.4%78.7%34.2%No value published by the source for this model1,600No value published by the source for this model1M927.7KChina
MiMo-V2.5-ProXiaomiMiMo
  • Text (LLM)
mitPermissive1T1,467No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,476No value published by the source for this model1M22.5KChina
GLM-5.1zai-org
  • Text (LLM)
mitPermissive753.9B1,46689.9%93.3%74.2%34.0%No value published by the source for this model1,508No value published by the source for this model202.8K74KChina
Kimi K2.6moonshotai
  • Text (LLM)
  • Vision
modified-mitPublisher’s terms1T1,46090.8%96.1%76.7%34.9%No value published by the source for this model1,5091,280262.1K435KChina
DeepSeek-V4-Prodeepseek-ai
  • Text (LLM)
mitPermissive1.6T1,45790.9%96.7%77.6%47.0%No value published by the source for this model1,445No value published by the source for this model1M514.4KChina
Hy3tencent
  • Text (LLM)
apache-2.0Permissive298.8B1,456No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,509No value published by the source for this model262.1K12.8KNo value published by the source for this model
Gemma 4 31Bgoogle
  • Text (LLM)
  • Vision
apache-2.0Permissive31.3B1,45175.8%73.3%No value published by the source for this model10.4%No value published by the source for this model1,3641,276262.1K9.3MUnited States
Qwen3.5-397B-A17BQwen
  • Text (LLM)
  • Vision
apache-2.0Permissive403.4B1,44286.4%88.9%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,3991,265262.1K199.7KChina
GLM-4.7zai-org
  • Text (LLM)
mitPermissive358.3B1,44283.3%83.3%No value published by the source for this model32.2%No value published by the source for this model1,435No value published by the source for this model202.8K101.1KChina
MiniMax-M3MiniMaxAI
  • Text (LLM)
  • Vision
minimax-communityPublisher’s terms427B1,44190.9%71.1%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,4851,2541M167.8KChina
Inklingthinkingmachines
  • Text (LLM)
  • Vision
apache-2.0Permissive952.4B1,44088.3%88.9%No value published by the source for this model40.3%No value published by the source for this model1,412No value published by the source for this modelNo value published by the source for this model369.2KUnited States
Gemma 4 26B A4Bgoogle
  • Text (LLM)
  • Vision
apache-2.0Permissive25.8B1,43873.2%82.2%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,3621,259262.1K10.7MUnited States
Qwen3.8-27BQwen
  • Text (LLM)
  • Vision
apache-2.0Permissive27.8B1,437No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,5911,272262.1K6.8MChina
DeepSeek-V4-Flashdeepseek-ai
  • Text (LLM)
mitPermissive290.9B1,436No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1M1.4MChina
Mistral Medium 3.5mistralai
  • Text (LLM)
otherPublisher’s terms127.7B1,426No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,2651,222262.1K99.2KNo value published by the source for this model
DeepSeek-V3.2deepseek-ai
  • Text (LLM)
mitPermissive685.4B1,42583.4%87.8%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,325No value published by the source for this model163.8K2.8MChina
DeepSeek-R1-0528deepseek-ai
  • Text (LLM)
mitPermissive684.5B1,42176.3%66.4%No value published by the source for this modelNo value published by the source for this model96.6%No value published by the source for this modelNo value published by the source for this model163.8K157.4KChina
Qwen3.5-122B-A10BQwen
  • Text (LLM)
  • Vision
apache-2.0Permissive125.1B1,417No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,3581,245262.1K323.7KChina
Qwen3-VL-235B-A22BQwen
  • Vision
apache-2.0Permissive235.7B1,414No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,247262.1K329.6KNo value published by the source for this model
Mistral Large 3mistralai
  • Text (LLM)
apache-2.0PermissiveNo value published by the source for this model1,413No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,2291,223No value published by the source for this model2.2KNo value published by the source for this model
Inkling-Smallthinkingmachines
  • Text (LLM)
  • Vision
apache-2.0Permissive266B1,40588.5%90.0%No value published by the source for this model19.1%No value published by the source for this model1,4071,236No value published by the source for this model449.6KUnited States
Qwen3-235B-A22B-Thinking-2507Qwen
  • Text (LLM)
apache-2.0Permissive235.1B1,40080.1%86.7%No value published by the source for this model40.4%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model262.1K13.7KChina
Qwen3-Coder-480B-A35BQwen
  • Code
apache-2.0Permissive480.2B1,388No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,274No value published by the source for this model262.1K35.1KChina
Qwen3-30B-A3B-Instruct-2507Qwen
  • Text (LLM)
apache-2.0Permissive30.5B1,38255.6%62.2%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model262.1K789KChina
GLM-4.6Vzai-org
  • Vision
mitPermissive107.7B1,378No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,163131.1K3.7KNo value published by the source for this model
Gemma 3 27BgoogleGated
  • Text (LLM)
  • Vision
gemmaPublisher’s terms27.4B1,36547.7%22.5%No value published by the source for this modelNo value published by the source for this model74.0%No value published by the source for this model1,164No value published by the source for this model389.6KUnited States
Mistral Small 3.2mistralai
  • Text (LLM)
apache-2.0Permissive24B1,35749.1%30.3%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1,141131.1K199.7KFrance
Command ACohereLabsGated
  • Text (LLM)
cc-by-nc-4.0Non-commercial111.1B1,354No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model2.5KNo value published by the source for this model
gpt-oss-120bopenai
  • Text (LLM)
apache-2.0Permissive116.8B1,35275.8%88.9%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model131.1K4.8MUnited States
Qwen3-32BQwen
  • Text (LLM)
apache-2.0Permissive32.8B1,34765.7%66.9%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model41K5.1MChina
Olmo 3.1 32B Instructallenai
  • Text (LLM)
apache-2.0Permissive32.2B1,330No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model65.5K15.6KNo value published by the source for this model
Llama 4 Maverickmeta-llamaGated
  • Text (LLM)
  • Vision
llama4Publisher’s terms401.6B1,32767.0%20.6%No value published by the source for this modelNo value published by the source for this model73.0%No value published by the source for this model1,142No value published by the source for this model10KUnited States
Llama 4 Scoutmeta-llamaGated
  • Text (LLM)
  • Vision
llama4Publisher’s terms108.6B1,32151.8%7.8%No value published by the source for this modelNo value published by the source for this model62.3%No value published by the source for this model1,118No value published by the source for this model161.8KUnited States
Llama 3.3 70Bmeta-llamaGated
  • Text (LLM)
llama3.3Publisher’s terms70.6B1,31847.4%5.1%No value published by the source for this modelNo value published by the source for this model41.6%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model916.8KUnited States
gpt-oss-20bopenai
  • Text (LLM)
apache-2.0Permissive20.9B1,31760.8%65.3%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model131.1K6.7MUnited States
Command R+ (08-2024)CohereLabsGated
  • Text (LLM)
cc-by-nc-4.0Non-commercial103.8B1,276No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model142No value published by the source for this model
Phi-4microsoft
  • Text (LLM)
mitPermissive14.7B1,25656.1%13.8%No value published by the source for this modelNo value published by the source for this model64.9%No value published by the source for this modelNo value published by the source for this model16.4K621.5KUnited States
Mixtral 8x22Bmistralai
  • Text (LLM)
apache-2.0Permissive140.6B1,22934.1%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model24.2%No value published by the source for this modelNo value published by the source for this model65.5K30.3KFrance
Llama 3.1 8Bmeta-llamaGated
  • Text (LLM)
llama3.1Publisher’s terms8B1,21127.0%1.7%No value published by the source for this modelNo value published by the source for this model22.9%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model6.2MUnited States
BGE-M3BAAI
  • Embeddings
mitPermissiveNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model8.2K37.2MNo value published by the source for this model
ChatterboxResembleAI
  • Text to speech
mitPermissiveNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model1.8MNo value published by the source for this model
Devstral Small 2mistralai
  • Code
apache-2.0Permissive24BNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model393.2K305.2KFrance
Granite 4.1 30Bibm-granite
  • Text (LLM)
apache-2.0Permissive28.9BNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model131.1K550.2KUnited States
Kimi K2.7 Codemoonshotai
  • Code
  • Vision
modified-mitPublisher’s terms1TNo value published by the source for this model87.9%95.6%No value published by the source for this model36.5%No value published by the source for this model1,472No value published by the source for this model262.1K104.7KChina
Kokoro-82Mhexgrad
  • Text to speech
apache-2.0PermissiveNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model11.9MNo value published by the source for this model
Magistral Small 1.2mistralai
  • Text (LLM)
apache-2.0Permissive24BNo value published by the source for this model47.6%28.1%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model131.1K13.6KFrance
Mistral 7B v0.3mistralai
  • Text (LLM)
apache-2.0Permissive7.2BNo value published by the source for this model15.2%0.3%No value published by the source for this modelNo value published by the source for this model3.7%No value published by the source for this modelNo value published by the source for this model32.8K2.3MFrance
multilingual-e5-largeintfloat
  • Embeddings
mitPermissive560MNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model5147.3MNo value published by the source for this model
Nemotron 3 Ultranvidia
  • Text (LLM)
openmdw-1.1Publisher’s terms560.5BNo value published by the source for this model85.4%86.7%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model262.1K171.3KUnited States
nomic-embed-text-v1.5nomic-ai
  • Embeddings
apache-2.0Permissive137MNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model2K14.2MNo value published by the source for this model
Parakeet TDT 0.6B v3nvidia
  • Speech to text
cc-by-4.0Permissive627MNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model571.6KNo value published by the source for this model
Qwen3-Embedding-8BQwen
  • Embeddings
apache-2.0Permissive7.6BNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model41K2.6MNo value published by the source for this model
Qwen3.6-27BQwen
  • Text (LLM)
  • Vision
apache-2.0Permissive27.8BNo value published by the source for this model85.9%91.1%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model262.1K3MChina
Qwen3.6-35B-A3BQwen
  • Text (LLM)
  • Vision
apache-2.0Permissive36BNo value published by the source for this model84.8%86.7%No value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model262.1K3.1MChina
Voxtral Smallmistralai
  • Speech to text
apache-2.0Permissive24.3BNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model131.1K24.4KNo value published by the source for this model
Whisper large-v3openai
  • Speech to text
apache-2.0Permissive1.5BNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model4.7MNo value published by the source for this model
Whisper large-v3-turboopenai
  • Speech to text
mitPermissive809MNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this modelNo value published by the source for this model6.5MNo value published by the source for this model

Side by side

Compare selected models

Select two to four models in the table to compare them here.

What the figures measure

The benchmarks, in plain words

Each benchmark tests one narrow skill under fixed conditions. Read them together, check the date, and test on your own tasks before choosing a model.

LMArena Text

Human preference on general chat. People compare two anonymous answers to the same prompt and vote; LMArena turns the votes into a rating (Bradley–Terry, on an Elo-like scale) with style control on, so longer or more formatted answers are not rewarded for form alone. Compare ratings together with their interval, not ranks alone.

Higher is better. Source: LMArena · read on Sep 22, 2026

GPQA Diamond

Graduate-level multiple-choice questions in biology, physics and chemistry, written by domain experts to be hard to look up. Epoch AI’s own runs; the score is the share answered correctly, where random guessing scores a quarter.

Higher is better. Source: Epoch AI · read on Sep 24, 2026

OTIS Mock AIME 2024–2025

Competition mathematics: problems from mock AIME exams written for the OTIS programme, with a single numeric answer each. Epoch AI’s own runs; share of problems solved.

Higher is better. Source: Epoch AI · read on Sep 24, 2026

SWE-bench Verified

Real software engineering: fixing actual GitHub issues in open-source Python repositories, checked by the projects’ tests, on a subset verified by human engineers. Epoch AI’s own runs; share of issues resolved.

Higher is better. Source: Epoch AI · read on Sep 24, 2026

SimpleQA Verified

Short factual questions with one verifiable answer, a test of what a model knows without search and of how often it makes things up. Epoch AI’s own runs; share answered correctly.

Higher is better. Source: Epoch AI · read on Sep 24, 2026

MATH Level 5

The hardest level of the MATH dataset of competition problems. Mostly run on older models, so newer ones often show “—”. Epoch AI’s own runs; share solved.

Higher is better. Source: Epoch AI · read on Sep 24, 2026

LMArena WebDev

Human preference on building web apps. Voters compare two models’ working web pages for the same request; the rating uses the same method as the text arena.

Higher is better. Source: LMArena · read on Sep 22, 2026

LMArena Vision

Human preference on prompts that include an image, such as charts, photos or documents; same rating method as the text arena.

Higher is better. Source: LMArena · read on Sep 22, 2026

Model details

  • Parameters: the total count from the model’s safetensors index on Hugging Face. Mixture-of-experts models use only part of them for each token, so size is not the running cost.
  • Context (config): max_position_embeddings in the model’s config.json. The context window a provider serves can be shorter, and gated repositories do not publish the file, hence “—”.
  • Licence: the identifier on the model’s Hugging Face page. “Publisher’s own terms” covers custom licences such as Llama, Gemma or modified MIT: read the terms before any commercial use.
  • Downloads and likes: as counted by Hugging Face, for this repository only (quantised copies elsewhere are not included).
  • Origin: organisation and country as recorded by Epoch AI. Left empty when no source states it; Azinove does not infer it.

Where Epoch AI ran a model at several reasoning settings, the table shows its best result and names that run in the cell’s tooltip. Rows marked “—” have no published result from that source for these exact weights.

Sources

Sources, licences and freshness

Azinove reads these public sources on its own server, at most once a day, and keeps a saved copy so the page keeps working when a source is unavailable.

  • Hugging Face Hub

    Last read on Sep 24, 2026

    Used for
    Parameters, licence, context, downloads, likes and last update of each repository
    Citation
    Model metadata from the Hugging Face Hub API (huggingface.co), each row linked to its repository.
  • Epoch AI

    Last read on Sep 24, 2026

    Used for
    GPQA Diamond, OTIS Mock AIME 2024–2025, SWE-bench Verified, SimpleQA Verified, MATH Level 5, plus organisation and country
    Licence
    CC BY 4.0
    Citation
    Epoch AI, ‘Capabilities & Benchmarking’. Published online at epoch.ai. Retrieved from ‘https://epoch.ai/benchmarks’ [online resource].
  • LMArena

    Last read on Sep 22, 2026

    Used for
    LMArena Text, LMArena WebDev, LMArena Vision, from the public leaderboard dataset
    Licence
    CC BY 4.0
    Citation
    LMArena, Arena Leaderboard Dataset (lmarena-ai/leaderboard-dataset), Hugging Face, CC BY 4.0.

Azinove AI early access

Tell us which models you need

Azinove AI runs open-weight models on GPUs Azinove operates. During early access, the models, hosting location and terms are agreed with each company in writing, so the models your team needs shape what is offered first.