Skip to main content
Limits vary by access type, plan, endpoint, and model availability.

Text generation (/v1/chat/completions)

Access type1 second1 minute1 hour
Anonymous11060
Free token240100
Pro ($12/mo)251,50015,000

Token usage availability

Token usage is counted as input tokens plus output tokens.
Access typeToken availability
Anonymous500,000 tokens per 24 hours
Free token1,000,000 tokens per 24 hours
Pro ($12/mo)Higher dynamically managed token limits than the free tier
For Pro accounts, token availability is calculated across the monthly billing period based on your usage relative to the percentage of the billing period that has elapsed. You can check current allowance in the dashboard. After the Pro allowance is reached, requests can continue from a topped-up balance and are billed from the live model pricing. Pro token availability remains subject to model availability, service capacity, fair-use calculations, and abuse-prevention controls. Note: Maximum input sizes also depend on your access tier and the specific model context window.
Use the Models API to inspect live model tiers, pricing, context windows, and capability flags. Get a free token at dash.llm7.io or upgrade to Pro for higher caps.