Skip to Content
IDEsWindsurfQuota & billing

Windsurf quota and billing

How Windsurf bills a request depends on plan: self-serve quota (daily + weekly refresh) or enterprise ACUs. This page answers quota meter questions and the GSC fragment how-windsurf-bills-a-request.

Hub: Tokenminning in Windsurf.

Self-serve quota plans

Quota docs : allowance refreshes automatically. Consumption scales with tokens processed; cost per token varies by model.

Model tierQuota impact
SWE-1.5 / SWE-1.6Free on current quota plans
Claude / GPT frontierScales with tokens + context size
Fast variantsHigher per-token cost for speed

What does not burn Cascade quota:

  • Command (Cmd/Ctrl+I) inline edits
  • Tab autocomplete
  • Auto-generated Memories creation/retrieval (but memories still add tokens when attached)

Prompt caching

Follow-up messages in the same Cascade conversation with the same model reuse cached context at reduced cost. Switching models mid-thread loses the cache — start a new Cascade chat when changing tiers.

Where to measure

SurfaceLocation
In-editor meterDaily/weekly remaining quota
Plan pagewindsurf.com/subscription/manage-plan 
Per messageCascade Stats for Nerds on chat rows
Teams / EnterpriseAnalytics  or Cascade Analytics API

Enterprise ACUs

Enterprise may bill Agent Compute Units  instead of quota. Legacy credit plans charge per Cascade message to premium models. Confirm your contract before optimizing.

Guardrails

  • Enable extra-usage caps before on-demand billing
  • Team admins: usage configuration API for per-user caps
  • Glance at meter after heavy Cascade weeks
Last updated on