Vercel AI Gateway
Pi's model card for `alibaba/qwen3-max-preview` via Vercel AI Gateway corroborates the first-party data with a 262,144-token context window, 32,768 max output tokens, and a tiered pricing structure of $1.20 per 1M input tokens and $6 per 1M output tokens, plus a $0.24 cache-read rate. The API surface is listed as Anthr Effective compatibility flags include EagerToolInputStreaming, LongCacheRetention, CacheControlOnTools, and Temperature support. Pi's session-cost estimator illustrates a typical agent task at roughly 25 requests with ~5k starting context totaling ~$0.235 effective cost versus ~$0.728 without caching, based on observed