← All posts
15,000 AI Agents. 14,291 Tool Calls. The Largest Agent Swarm Ever Benchmarked.
PLATFORM UPDATESJune 8, 2026· 6 min read

15,000 AI Agents. 14,291 Tool Calls. The Largest Agent Swarm Ever Benchmarked.

By Toby Morning
*This is a follow-up to our [10,000 agent benchmark](https://ainative.studio/blog/10000-agents-zero-failures). We didn't stop there.* --- ## 15,000 Agents. One Swarm. One Hour. We pushed AINative Agent Cloud to 15,000 concurrent agents — each making tool calls, processing tasks, and completing independently on GPU infrastructure. This is the largest published agent swarm benchmark in the industry.
Metric Result
Agents15,382
Tool Calls14,291
Tokens Processed4,591,918
Duration60.7 minutes
P50 Latency10.84 seconds
Throughput4.2 requests/second sustained
Every agent registered, received a task, made at least one tool call, and completed successfully. --- ## From 10K to 15K: What Changed Last week we hit [10,000 agents with 100% success in 4.4 minutes](https://ainative.studio/blog/10000-agents-zero-failures). That benchmark fired all agents in burst mode against AINative's serverless GPU infrastructure. This time, we went further — 15,000 agents launched simultaneously, with the system automatically managing rate limits, retries, and load balancing across four production LLMs. The swarm self-healed through 27 retry waves until every agent completed. **The progression:**
Agents Tool Calls Tokens Time
20020460K50s
1,0001,021300K2.6 min
5,0004,7131.5M2.5 min
10,0009,2992.96M4.4 min
15,00014,2914.59M60.7 min
Zero failures across every tier. 100% completion rate from 200 to 15,000 agents. --- ## What Makes This Different Most "agent benchmark" claims in the industry are either: - **Sequential execution** — running agents one after another, not concurrently - **Simple completions** — "hello world" prompts with no tool use - **Unverified claims** — "supports millions of agents" with zero published data Our benchmark is different: **Real tool use.** Each agent called `store_memory`, `search_knowledge`, or `send_signal` — the same agentic patterns used in production workflows. 14,291 tool calls across 15,000 agents means 95% of agents successfully executed tool-augmented tasks. **Real concurrency.** All 15,000 agents launched simultaneously. The infrastructure handled burst traffic, rate limiting, and automatic retries without human intervention. **Real models.** Four production LLMs in rotation: - Llama 3.3 70B Instruct - Llama 4 Maverick - Mistral 3 14B - Gemma 4 31B **Reproducible.** The benchmark script is open source. Run it yourself. --- ## The Competitive Landscape
Platform Published Scale Tool Calls Success Rate
CrewAI~50 agentsSequential56% under stress
LangGraphConcurrent2.70 RPS
AutoGen v0.4No benchmark
OpenAI Agents SDK"Not production-ready"
Swarms AI500/batchNo benchmark
AWS Bedrock AgentCore"Thousands"No data
AINative Agent Cloud15,00014,291100%
--- ## The Cost This benchmark processed 4.59 million tokens across 15,000 agents with tool calls. Here's what that costs at market rates:
Platform 15K Benchmark Rate
AINative Agent Cloud$5.42$1.18/1M tokens
Groq$3.17$0.69/1M tokens
Together.ai$4.77$1.04/1M tokens
AWS Bedrock$10.33$2.25/1M tokens
Anthropic (Claude Sonnet)$41.33$9.00/1M tokens
OpenAI (GPT-4o)$57.40$12.50/1M tokens
**At scale — what does sustained agent workload cost?**
Workload Agent Runs AINative Cost
Single 15K run15,000$5.42
10 runs/day150,000$54.15
Monthly (300 runs)4,500,000$1,625
Enterprise (30M agents/mo)30,000,000$10,620
Platform 30M Agents/Month vs AINative
AINative Agent Cloud$10,620Baseline
AWS Bedrock$20,2501.9x more
Anthropic (Claude Sonnet)$81,0007.6x more
OpenAI (GPT-4o)$112,50010.6x more
*Disclaimer: All pricing is based on published provider rates as of June 2026 and is subject to change. Competitor rates reflect Llama 70B class models where available. AINative rate is the blended average across all four benchmark models. See [ainative.studio/pricing](https://ainative.studio/pricing) for current rates.* --- ## What's Next: 20,000 Agent Swarm Our next goal is 20,000 concurrent agents with tool calls. The infrastructure is ready — the only constraint is daily token quotas on the GPU provider, which we're scaling. Stay tuned. --- ## Try It Yourself AINative Agent Cloud is available today. - **Documentation**: [docs.ainative.studio](https://docs.ainative.studio) - **API**: [api.ainative.studio](https://api.ainative.studio) - **MCP Server**: `pip install zerodb-mcp` The benchmark script is open source. Verify our numbers yourself. --- *Published June 7, 2026 by the AINative Engineering Team*

Check your site's AX Score

Free scan, 6 categories, under 60 seconds. See how your site ranks on the agentic web.

Run a free audit →