MashupAbility//OpenSource

OPEN-WEIGHT INTELLIGENCE · EVIDENCE FIRST

The frontier you can actually run.

Open-weight models, compared without the hand-waving. See what is downloadable, what the license really permits, and what it takes to self-host.

34

models tracked

31

weights live

20

U.S. weights live

1M

largest context

MODEL LEDGER

Frontier index

34 of 34 models

Origin follows the primary pretraining organization, not contributor citizenship, investor location, or hosting region. Downloadable weights do not imply a permissive license. Provider benchmarks use different harnesses and are not a cross-model ranking.

Comparable properties of open-weight frontier models
CompareModelStatusScaleContextInputsReasoning + agentsLicense termsOpenness levelSelf-hostingSource
Muse Glimmer 30BMeta Superintelligence Labs · United States · Aug 10, 2026Weights live30B30B dense activeNot statedNo context-window size in the release announcement
TextVision
Long-horizon agent execution with retry trainingLocal agents, function calling, local coding, and LLM-as-a-judgePermissiveApache 2.0L1 · WeightsUnder 20 GB quantized; designed for 24 GB or 32 GB memoryOfficial ↗
Qwen3.8-MaxAlibaba Qwen · China · Announced · Aug 2, 2026Announced · weights next week2.4TUndisclosed activeTBDNot stated in supplied announcement
TextVision
System-level autonomous planningAutonomous coding and professional workflowsLicense pendingTo be published with weightsNot yet openHosted model; downloadable weights pending$2/M input · $6/M output · $0.25/M implicit cache · pricing source ↗Official ↗
BTL-3Bad Theory Labs · United States · Jul 23, 2026LoRA adapter live27B base requiredRank-32 LoRA adapter active256K32K launch benchmark context
Text
Thinking-mode codingRepository work and structured tool usePermissiveApache 2.0 adapterL1 · WeightsRequires pinned Qwen3.6-27B base checkpointOfficial ↗
Nanbeige 4.2 3BNanbeige · China · Jul 22, 2026Weights live4B incl. embeddings3B non-embedding active256K262,144 positions
Text
Thinking model with agent post-trainingTool use, office work, and code agentsPermissiveApache 2.0L1 · Weights22 shared layers evaluated in two passesOfficial ↗
Solar Open 2Upstage · South Korea · Jul 22, 2026Weights live250B15B active1MNative hybrid attention
Text
Long-horizon reasoningOffice, document, and coding agentsCustom licenseUpstage Solar LicenseL1 · WeightsVendor minimum: 4× H200Official ↗
Laguna S 2.1Poolside · United States · Jul 21, 2026Weights live118B8B active1MUp to 1M
Text
Thinking / non-thinkingAgentic coding and long-horizon workPermissiveOpenMDW-1.1L1 · WeightsQuantized builds available; full model remains memory-heavyOfficial ↗
Motif-3-BetaMotif Technologies · South Korea · Jul 21, 2026Beta · weights live~314B~13B active256KNative
Text
Reasoning with custom architectureTool calling and general agentsUse-restrictedNon-commercial research licenseL1 · WeightsCustom code; tested by vendor on H200 and B200Official ↗
Antares-1BCisco Foundation AI · United States · Jul 21, 2026Weights live1B1B dense active128KGranite 4.0 backbone
Text
SFT then GRPO for terminal explorationVulnerability localization in repositoriesPermissiveApache 2.0L1 · WeightsSmall checkpoint; terminal agent loop requiredOfficial ↗
Kimi K3Moonshot AI · China · Jul 16, 2026Weights Jul 272.8TUndisclosed active1MAdvertised
TextVision
Max at launch; lower efforts plannedCoding, tools, Agent SwarmLicense pendingTBD with weightsNot yet openAPI today; weights promised Jul 27Official ↗
InklingThinking Machines · United States · Jul 15, 2026Weights live975B41B active1M64K / 256K on Tinker
TextVisionAudio
Controllable effortCoding, tools, fine-tuning on TinkerUse-restrictedApache 2.0 + Model AUPL1 · Weights≥600 GB quantized; ≥2 TB BF16Official ↗
GLM-5.2Z.ai · China · Jun 16, 2026Weights live744B40B active1MNative long-horizon target
Text
Off / high / maxTool calling, long-horizon engineeringPermissiveMITL1 · WeightsCluster class; BF16 and FP8Official ↗
Kimi K2.7 CodeMoonshot AI · China · Jun 12, 2026Weights live1T32B active256KNative
TextVision
Forced, preserved across turnsLong-horizon coding and toolsUse-restrictedModified MITL1 · WeightsLarge multi-GPU; native INT4Official ↗
Nemotron 3 UltraNVIDIA · United States · Jun 4, 2026Weights live550B55B active1MRULER tested
Text
Off / regular / mediumStructured tools, delegation, recoveryPermissiveOpenMDW 1.1L1 · Weights8×H200/B200 or 16×H100 BF16Official ↗
Gemma 4 12BGoogle DeepMind · United States · Jun 3, 2026Weights live12B12B dense active256KNative
TextVisionAudio
Configurable thinkingNative function calling, structured JSONPermissiveApache 2.0L1 · WeightsLaptop class; targets 16 GB memoryOfficial ↗
MiniMax M3MiniMax · China · Jun 1, 2026Weights live~428B~23B active1MNative
TextVisionVideo
Off / enabled / adaptiveCoding, computer use, long-horizon agentsUse-restrictedMiniMax CommunityL1 · WeightsCluster classOfficial ↗
Step 3.7 FlashStepFun · China · May 29, 2026Weights live196B + 1.8B vision~11B active256KNative
TextVision
Low / medium / highVisual search, tools, GUI workflowsPermissiveApache 2.0L1 · Weights~120 GB minimum unified memoryOfficial ↗
Mistral Medium 3.5Mistral AI · France · Apr 28, 2026Weights live128B128B dense active256KNative
TextVision
Configurable effortFunction calling, JSON, codingUse-restrictedModified MITL1 · WeightsAs few as 4 GPUs (vendor)Official ↗
DeepSeek V4 ProDeepSeek · China · Apr 24, 2026Preview · weights live1.6T49B active1MThink Max recommends ≥384K
Text
Off / think / think maxCoding, reasoning, toolsPermissiveMITL1 · WeightsCluster class; mixed FP4 / FP8Official ↗
Qwen3.6 27BAlibaba Qwen · China · Apr 22, 2026Weights live27B27B dense active256KExtensible to ~1M
TextVisionVideo
Thinking / non-thinkingTools, repository reasoning, codingPermissiveApache 2.0L1 · WeightsWorkstation/server; 8 GPUs for full 256KOfficial ↗
Trinity-Large-ThinkingArcee AI · Datology · Prime Intellect · United States · Apr 1, 2026Weights live398B~13B active512KExtended from 8K pretraining
Text
Native thinking tracesTool calling and multi-step agentsPermissiveOpenMDW-1.1L1 · WeightsCluster-class sparse MoEOfficial ↗
GLM-4.7-FlashZ.ai · China · Jan 19, 2026Weights live30B3B active~200K202,752 positions
Text
Reasoning with preserved thinkingTools and coding agentsPermissiveMITL1 · WeightsLocal/workstation friendlyOfficial ↗
Olmo 3 32B ThinkAi2 · United States · Nov 19, 2025Weights live32B32B dense active65KYaRN extension from 8K
Text
Long chain-of-thoughtMath, code, and research workflowsPermissiveApache 2.0L3 · Weights + algorithms + training dataDownloadable checkpoint; runner-dependentOfficial ↗
Granite 4.0 H 1BIBM · United States · Oct 28, 2025Weights live1.5B1.5B hybrid active128KMamba2 + attention
Text
General instruction followingCompact enterprise workflowsPermissiveApache 2.0L1 · WeightsSmall downloadable checkpointOfficial ↗
Apriel H1 15B ThinkerServiceNow · United States · Oct 28, 2025Weights live15B15B dense active65KTarget; runtime-dependent
Text
Thinker SFTEnterprise reasoning workflowsPermissiveMITL1 · WeightsDownloadable checkpoint; runner-dependentOfficial ↗
Marin 32B BaseStanford · Marin Community · United States · Oct 25, 2025Weights live32B32B dense active4KBase checkpoint
Text
Base model; no chat tuningResearch and downstream post-trainingPermissiveApache 2.0L1 · WeightsDownloadable base checkpointOfficial ↗
LFM2-VL-3BLiquid AI · United States · Oct 22, 2025Weights live3B2.6B LM + 0.4B vision active32KText context
TextVision
Lightweight multimodal reasoningInstruction following and agentic flowsUse-restrictedLFM Open License v1.0L1 · WeightsSmall multimodal checkpointOfficial ↗
Moondream 3 PreviewMoondream · United States · Sep 11, 2025Preview · weights live9B~2B active32KVision-language preview
TextVision
Visual question answeringCaptioning, pointing, and vision tasksUse-restrictedBSL 1.1 + Additional Use GrantL1 · WeightsDownloadable preview; third-party service limitsOfficial ↗
VaultGemma 1BGoogle · United States · Sep 5, 2025Weights live · gated1B1B dense active1KPretrained base model
Text
Base model; no chat tuningResearch and downstream tuningCustom licenseGemma TermsL1 · WeightsSmall checkpoint; terms acceptance requiredOfficial ↗
Nemotron Nano 9B v2NVIDIA · United States · Aug 18, 2025Weights live9B9B hybrid active128KMamba + Transformer
Text
Reasoning on / offEfficient local reasoningCustom licenseNVIDIA Open Model LicenseL1 · WeightsSmall downloadable checkpointOfficial ↗
gpt-oss-120bOpenAI · United States · Aug 5, 2025Weights live117B5.1B active128KNative
Text
Low / medium / highTools, functions, structured outputPermissiveApache 2.0 + usage policyL1 · WeightsSingle 80 GB GPUOfficial ↗
SmolLM3 3BHugging Face · United States · Jul 8, 2025Weights live3B3B dense active128K64K trained; YaRN to 128K
Text
Thinking / non-thinkingLightweight multilingual agentsPermissiveApache 2.0L2 · Weights + algorithmsSmall downloadable checkpointOfficial ↗
Phi-4 Mini Flash ReasoningMicrosoft · United States · Jun 19, 2025Weights live3.8B3.8B hybrid active64KHybrid SambaY architecture
Text
Math-focused reasoningStructured logic and code workflowsPermissiveMITL1 · WeightsSmall downloadable checkpointOfficial ↗
Llama 4 MaverickMeta · United States · Apr 5, 2025Weights live400B17B active1MNative claim
TextVision
General reasoningTool support depends on runnerCustom licenseLlama 4 CommunityL1 · WeightsDGX-class FP8 deploymentOfficial ↗
Reflection model (pending)Reflection AI · United States · No model releasedWeights pendingTBDTBDTBDNo public checkpoint
TBD
TBDOpen-model commitment onlyLicense pendingTBD with releaseNot yet openNo downloadable artifactOfficial ↗

COMPARE · 0/2

Select two models

METHODOLOGY

Open is not a binary.

Weights first

“Live” means a downloadable checkpoint exists. API access alone does not qualify.

Three levels of open

Level 1 has open weights. Level 2 also has open algorithms. Level 3 also has open training data. The index assigns the highest evidenced level.

Claims stay sourced

Provider benchmarks remain labeled and are not normalized into a misleading leaderboard.

Hardware counts

A model is only “runnable” in context: laptop, workstation, or cluster-class.