SmartBrain API

Enterprise AI as Simple as Utilities

More Links
Model HubConsoleAPI KeysUsage QueryDocs
SmartBrain API. All rights reserved|Privacy Policy|Terms
SmartBrain API
SmartBrain API
  • Model Hub
  • Official Docs
  • Playground
千

qwen/qwen3.8-flash

Online Chat
阿里巴巴

Publish time

2026/8/27

Model Series

千问

Input type

Output type

Input Price

¥1 / 1M tokens

Output Price

¥3 / 1M tokens

Cache Write Price

¥1.25 / 1M tokens

Cache Read Price

¥0.1 / 1M tokens

Context Window

1,000,000

Max Output Length

128,000

Qwen3.8-Flash 是千问最新推出的多模态大模型,兼具强大的理解与生成能力和出色的响应速度。模型原生支持百万级上下文窗口,能够一次性处理超长文档、代码仓库和复杂对话。在编程辅助、智能体协作、图文理解等场景中表现尤为出色——无论是自动修复代码、操作桌面应用,还是分析图表与长视频,都能给出准确、高质量的结果。同时兼容 OpenAI 与 Anthropic 主流接口协议,可无缝接入 Claude Code、Codex 等开发者工具,轻松构建高并发应用与智能工作流。凭借优异的性能与极具竞争力的推理成本,Qwen3.8-Flash 是开发者和企业在 AI 应用中兼顾效果与效率的理想选择。

Providers for qwen/qwen3.8-flash

Zhinao API routes requests to the best-fit provider and automatically fails over to the one with highest availability.

Sort by
通
通义千问
国内

TTFT

1.60s

Throughput

51.07tps

Uptime

100.00%

Provider Model

alibaba/qwen/qwen3.8-flash

Supported Parameters

temperaturetop_ptop_k

Recent Uptime

10月11日 11 PM100.00%
10月9日 7 PM: 100.00%10月9日 7 PM: 100.00%10月9日 8 PM: 100.00%10月9日 8 PM: 100.00%10月9日 9 PM: 100.00%10月9日 9 PM: 100.00%10月9日 10 PM: 99.95%10月9日 10 PM: 100.00%10月9日 11 PM: 100.00%10月9日 11 PM: 99.96%10月10日 8 AM: 100.00%10月10日 9 AM: 100.00%10月10日 9 AM: 100.00%10月10日 10 AM: 100.00%10月10日 10 AM: 100.00%10月10日 11 AM: 100.00%10月10日 11 AM: 100.00%10月10日 12 PM: 99.98%10月10日 12 PM: 100.00%10月10日 1 PM: 100.00%10月10日 1 PM: 99.96%10月10日 2 PM: 100.00%10月10日 2 PM: 100.00%10月10日 3 PM: 100.00%10月10日 3 PM: 100.00%10月10日 4 PM: 100.00%10月10日 4 PM: 100.00%10月10日 5 PM: 100.00%10月10日 5 PM: 100.00%10月10日 6 PM: 100.00%10月10日 6 PM: 99.99%10月10日 7 PM: 100.00%10月10日 7 PM: 100.00%10月10日 8 PM: 100.00%10月10日 8 PM: 99.96%10月10日 9 PM: 100.00%10月10日 9 PM: 100.00%10月10日 10 PM: 100.00%10月10日 10 PM: 100.00%10月10日 11 PM: 100.00%10月10日 11 PM: 100.00%10月11日 8 AM: 100.00%10月11日 9 AM: 100.00%10月11日 9 AM: 100.00%10月11日 10 AM: 100.00%10月11日 10 AM: 100.00%10月11日 11 AM: 99.96%10月11日 11 AM: 100.00%10月11日 12 PM: 100.00%10月11日 12 PM: 100.00%10月11日 1 PM: 100.00%10月11日 1 PM: 100.00%10月11日 2 PM: 100.00%10月11日 2 PM: 100.00%10月11日 3 PM: 100.00%10月11日 3 PM: 100.00%10月11日 4 PM: 100.00%10月11日 4 PM: 100.00%10月11日 5 PM: 99.95%10月11日 5 PM: 100.00%10月11日 6 PM: 100.00%10月11日 6 PM: 100.00%10月11日 7 PM: 100.00%10月11日 7 PM: 100.00%10月11日 8 PM: 100.00%10月11日 8 PM: 100.00%10月11日 9 PM: 99.96%10月11日 9 PM: 100.00%10月11日 10 PM: 100.00%10月11日 10 PM: 100.00%10月11日 11 PM: 100.00%10月11日 11 PM: 100.00%

Reasoning

Toggleable

Supported Response Formats

OpenAI Chat CompletionsAnthropic MessagesOpenAI Responses

Total Context

1,000,000

Max Output

128,000

Input Price

¥1 / 1M tokens

Output Price

¥3 / 1M tokens

Cache Write

¥1.25 / 1M tokens

Cache Read

¥0.1 / 1M tokens

Performance for qwen/qwen3.8-flash

Compare different providers across Zhinao API

Throughput

61.41 tok/s

TTFT

1.70 s

Uptime for qwen/qwen3.8-flash

Uptime for qwen/qwen3.8-flash across all providers

Sample code and API for qwen/qwen3.8-flash

Get API Key

Zhinao API normalizes requests and responses across providers for you

View full docs
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.360.cn/v1",
  apiKey: process.env.ZHINAO_API_KEY,
});

const response = await client.chat.completions.create({
  model: "qwen/qwen3.8-flash",
  messages: [
    { role: "user", content: "Hello, how are you?" }
  ],
  temperature: 0.7,
  max_tokens: 1000,
});

console.log(response.choices[0].message.content);