Sign InOpen Brain
VercelEngineering PostOfficial Source

10x more capacity for Laguna S 2.1 on AI Gateway

AI Gateway has raised Laguna S 2.1 capacity tenfold for both paid and free model IDs, reducing throughput constraints for high-volume or long-running coding agents.

Vercel · Jul 31, 2026
Open Source Open MarkdownOpen JSON
Source Summary

Poolside’s **Laguna S 2.1** now has **10x more capacity** on AI Gateway. The increase covers both poolside/laguna-s-2.1 and poolside/laguna-s-2.1-free.

Practical Implication

If provider capacity was limiting parallel or long-running coding agents, retest the paid and free routes under your actual workload. Both remain selectable through the existing model configuration flow.

Agent-Ready Context
Poolside’s **Laguna S 2.1** now has **10x more capacity** on AI Gateway. The increase covers both poolside/laguna-s-2.1 and poolside/laguna-s-2.1-free.

If provider capacity was limiting parallel or long-running coding agents, retest the paid and free routes under your actual workload. Both remain selectable through the existing model configuration flow.

The announcement gives no absolute rate limits, latency figures, or reliability measurements. Tenfold capacity therefore describes the increase, not the throughput an individual account or agent run will receive.
Connected Context · Feed7 Judgment

This changes Laguna S 2.1’s availability envelope rather than its capabilities or demonstrated task quality. It removes one plausible constraint for parallel and long-running coding agents and justifies renewed load testing of both routes, but the lack of absolute limits, latency, and reliability data prevents treating the tenfold increase as guaranteed per-account throughput.

Laguna S 2.1 is now available on AI GatewayThe earlier signal established the free and paid Laguna variants and their context and thinking options; the new signal changes capacity for those same routes without adding model features.AI Gateway adds unified fast mode supportMore provider capacity addresses concurrency and availability, whereas fast mode requests a lower-latency serving tier; neither improvement establishes the other.AI Gateway: GPT-5.6 pricing and speed updatesBoth changes justify rerunning workload evaluations, but they alter different routing inputs: Laguna’s capacity versus GPT-5.6 price and serving speed.TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AIHigher capacity makes Laguna a less constrained routing candidate, while TRACE-Router indicates that final selection should still be based on task outcomes and latency rather than availability alone.
Context Map
infracoding#gateways#coding-agents#model-selection
Uncertainty
The announcement gives no absolute rate limits, latency figures, or reliability measurements. Tenfold capacity therefore describes the increase, not the throughput an individual account or agent run will receive.