Model details
GPT Chat Latest
GPT Chat Latest is presented in the Microsoft Foundry catalog as a preview chat model designed for advanced, natural, multimodal conversation in enterprise settings. The listing groups it under Azure OpenAI within the Direct from Azure program, where Microsoft curates models for secure deployment, unified billing, and PTU portability across its managed infrastructure. The catalog's positioning emphasizes an interactive, context-aware assistant experience rather than a raw general-purpose model, suggesting a focus on customer support, internal assistants, and similar dialogue-heavy workflows.
The model is described as combining multimodal chat with adaptive reasoning that can switch between concise and deeper step-by-step answers, alongside tuned safety and instruction-following behavior. That combination is aimed at organizations that need a controlled, responsive chat layer for production assistants where reliability, policy alignment, and quick iteration matter more than open-ended generation. Buyers can access it through Azure's standard purchase and deployment flow, with the option to scale on demand or reserve PTUs for predictable performance, which makes it a practical fit for teams already standardized on Microsoft Foundry who want a chat-optimized model with multimodal input and strong instruction adherence.
Quick Info
Powered by- Provider
- Azure Cognitive Services
- Model key
- gpt-chat-latest
- Release date
- May 5, 2026
- Last updated
- May 28, 2026
- Knowledge cutoff
- 2025-12-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $5.00
- Output token cost
- $30.00
Limits
- Input tokens
- 111,616 tokens
- Output tokens
- 16,384 tokens
- Context window
- 128,000 tokens