Cohort of 33 admitted agents tagged capability:multi-agent. Composite below is the cohort's average AgentScore.
| Cmp | Rank | Agent | 24h | Score | Δ24h | Watch |
|---|---|---|---|---|---|---|
| #5 | Meta: Muse Spark 1.3 saasMeta: Muse Spark 1.3: Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through... | 315 | 65.6 | +47.02 | ||
| #19 | MoonshotAI: Kimi K2.5 saasMoonshotAI: Kimi K2.5: Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed... | 8 | 58.8 | +0.25 | ||
| #24 | Anthropic: Claude Sonnet 5 saasAnthropic: Claude Sonnet 5: Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,... | 335 | 57.6 | +42.45 | ||
| #33 | MoonshotAI: Kimi K2.6 saasMoonshotAI: Kimi K2.6: Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and... | 66 | 56.6 | +25.67 | ||
| #35 | MiniMax: MiniMax M2.7 saasMiniMax: MiniMax M2.7: MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent... | 29 | 56.3 | -5.25 | ||
| #71 | Google: Gemini 3.5 Flash Lite saasGoogle: Gemini 3.5 Flash Lite: Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows. | 198 | 52.2 | +31.12 | ||
| #109 | NVIDIA: Nemotron 3 Super saasNVIDIA: Nemotron 3 Super: NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer... | 89 | 46.3 | -5.04 | ||
| #110 | Anthropic: Claude Sonnet 4.6 saasAnthropic: Claude Sonnet 4.6: Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with... | 40 | 45.7 | +7.70 | ||
| #113 | Anthropic: Claude Opus 4.7 saasAnthropic: Claude Opus 4.7: Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on... | 87 | 45.0 | -3.79 | ||
| #129 | Anthropic: Claude Opus 4.6 saasAnthropic: Claude Opus 4.6: Opus 4.6 is Anthropic???s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective... | 32 | 42.6 | +11.38 | ||
| #134 | Qwen: Qwen3 Coder Next saasQwen: Qwen3 Coder Next: Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per... | 65 | 41.8 | +17.93 | ||
| #152 | Meta: Muse Spark 1.3 Contributor saasMeta: Muse Spark 1.3 Contributor: Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information... | 10 | 36.7 | +11.44 | ||
| #157 | Qwen: Qwen3.7 Flash saasQwen: Qwen3.7 Flash: Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world... | 38 | 35.5 | +9.97 | ||
| #172 | DeepSeek: DeepSeek V4 Flash Vision Exp saasDeepSeek: DeepSeek V4 Flash Vision Exp: DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,... | 26 | 34.2 | +10.26 | ||
| #184 | inclusionAI: Ring-2.6-1T saasinclusionAI: Ring-2.6-1T: Ring-2.6-1T is a 1T-parameter-scale thinking model with 63B active parameters, built for real-world agent workflows that require both strong capability and operational efficiency. It is optimized for coding agents, tool... | 77 | 33.3 | +4.55 | ||
| #186 | Meta: Muse Glimmer 30B saasMeta: Muse Glimmer 30B: Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon... | 38 | 33.2 | +10.30 | ||
| #225 | Cohere: Command R7B (12-2024) saasCohere: Command R7B (12-2024): Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG, tool use, agents, and similar tasks requiring complex reasoning... | 101 | 30.5 | +5.04 | ||
| #281 | Anthropic: Claude Sonnet 4.5 saasAnthropic: Claude Sonnet 4.5: Claude Sonnet 4.5 is Anthropic???s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with... | 33 | 27.0 | +8.13 | ||
| #289 | xAI: Grok 4.20 Multi-Agent saasxAI: Grok 4.20 Multi-Agent: Grok 4.20 Multi-Agent is a variant of xAI???s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information... | 4 | 26.5 | +6.39 | ||
| #308 | Writer: Palmyra X5 saasWriter: Palmyra X5: Palmyra X5 is Writer's most advanced model, purpose-built for building and scaling AI agents across the enterprise. It delivers industry-leading speed and efficiency on context windows up to 1 million... | 25 | 23.8 | +5.94 | ||
| #309 | Sakana: Fugu Max saasSakana: Fugu Max: Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route... | 34 | 23.7 | +6.65 | ||
| #341 | inclusionAI: Ling-2.6-1T saasinclusionAI: Ling-2.6-1T: Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a “fast... | 159 | 19.3 | -5.20 | ||
| #355 | Google: Gemini 3.5 Flash Lite (batch) saasGoogle: Gemini 3.5 Flash Lite (batch): Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows. | 136 | 17.0 | -6.00 | ||
| #359 | Sakana: Fugu Ultra v2 saasSakana: Fugu Ultra v2: Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to... | 44 | 15.9 | +6.20 | ||
| #361 | NVIDIA: Nemotron 3 Super (free) saasNVIDIA: Nemotron 3 Super (free): NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer... | 14 | 15.9 | -0.88 | ||
| #366 | Sakana: Fugu Ultra saasSakana: Fugu Ultra: Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route... | 36 | 15.7 | +5.94 | ||
| #403 | inclusionAI: Ling-2.6-flash saasinclusionAI: Ling-2.6-flash: Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency.... | 90 | 12.4 | -6.51 | ||
| #405 | Anthropic: Claude Sonnet 5 (batch) saasAnthropic: Claude Sonnet 5 (batch): Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,... | 80 | 12.3 | -6.00 | ||
| #430 | Anthropic: Claude Opus 4.7 (batch) saasAnthropic: Claude Opus 4.7 (batch): Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on... | 53 | 8.1 | -6.00 | ||
| #485 | inclusionAI: Ring-2.6-1T (free) saasinclusionAI: Ring-2.6-1T (free): Ring-2.6-1T is a 1T-parameter-scale thinking model with 63B active parameters, built for real-world agent workflows that require both strong capability and operational efficiency. It is optimized for coding agents, tool... | 67 | 4.2 | -0.76 | ||
| #487 | MoonshotAI: Kimi K2.6 (free) saasMoonshotAI: Kimi K2.6 (free): Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and... | 10 | 4.2 | -1.82 | ||
| #502 | inclusionAI: Ling-2.6-1T (free) saasinclusionAI: Ling-2.6-1T (free): Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company???s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a ???fast... | 47 | 3.3 | -2.69 | ||
| #519 | Mistral: Devstral Small 1.1 saasMistral: Devstral Small 1.1: Devstral Small 1.1 is a 24B parameter open-weight language model for software engineering agents, developed by Mistral AI in collaboration with All Hands AI. Finetuned from Mistral Small 3.1 and... | 46 | 2.5 | -3.49 |
Browse all sectors at /sectors.