Kimi K2.5
Kimi K2.5 is Moonshot AI's successor to the K2 family: multimodal inputs, upgraded frontend coding, and a context window of 262.1K tokens, available through AI Gateway via Moonshot AI, Novita AI, Bedrock.
- Input and output price
- Prices from: Input $0.60, Output $3, Per 1M tokens
- 24h uptime
- Loading AI Gateway uptime
import { streamText } from 'ai'
const result = streamText({ model: 'moonshotai/kimi-k2.5', prompt: 'Why is the sky blue?'})Copy link to headingProviders
Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.
| Provider |
|---|
Copy link to headingPlayground
Try out Kimi K2.5 by Moonshot AI. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.
Kimi K2.5
Copy link to headingUptime24 hours
Direct request success rate on AI Gateway and per-provider. Visit the docs for more info.
Copy link to headingThroughput24 hours
P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the docs for more info.
Copy link to headingLatency24 hours
P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the docs for more info.
Copy link to headingAbout Kimi K2.5
Kimi K2.5, released on January 26, 2026, is the generation after the K2 line. Moonshot AI describes K2.5 across agent tasks, coding, visual understanding, and general intelligence benchmarks in its release materials. K2.5 extends both text-based and visual tasks.
Frontend code generation is a highlighted change. Moonshot AI documents more capable frontend coding, including interactive UI with dynamic layouts and animations, beyond bare syntax-level output.
Access K2.5 through AI Gateway by setting the model string to moonshotai/kimi-k2.5. No extra provider accounts are required for gateway-managed access. AI Gateway's observability layer tracks token usage and costs across requests, which helps when usage patterns vary.
Kimi K2.5 is available through AI Gateway at $0.6 per million input tokens and $3 per million output tokens.
Copy link to headingWhat To Consider When Choosing a Provider
- Configuration: Evaluate Kimi K2.5 against your specific use case. The expanded capabilities may not justify the cost relative to K2 variants for every workload.
- Zero Data Retention: Zero Data Retention is available for this model. It is offered on a per-provider and model basis. See the documentation for details.
- Authentication: AI Gateway authenticates requests using an API key or OIDC token. You do not need to manage provider credentials directly.
Copy link to headingWhen to Use Kimi K2.5
Best for
- Interactive frontend code: Generating UI with dynamic layouts, animations, and interactive components
- Multi-capability pipelines: Workloads spanning agent tasks, coding, visual understanding, and general intelligence in one pipeline
- Multimodal or frontend gaps: Kimi-family projects where earlier K2 variants lack the visual or frontend scope you need
- General-purpose assistants: Teams building AI assistants that must handle diverse task types from a single model
Consider alternatives when
- Extended reasoning traces: Kimi K2 Thinking is built for explicit chain-of-thought output
- Sufficient K2 checkpoint: The September 2025 K2 checkpoint covers your workload and K2.5's added scope isn't needed
- Speed-first workloads: Kimi K2 Turbo or K2 Thinking Turbo are better fits when you don't need K2.5's broader capabilities
- Cost-sensitive deployments: K2 variants may meet your quality bar at lower cost per token
Copy link to headingConclusion
Kimi K2.5 adds multimodal inputs and frontend coding emphasis to the Kimi line on AI Gateway, alongside agent, coding, and vision workloads. As of January 26, 2026, it's the K2 successor listed for those combined use cases on AI Gateway.
Your use is subject to Moonshot AI's Terms & Privacy Policies.