Qwen

Qwen3.8 Flash

An efficient multimodal workhorse in the Qwen family, with a native million-token context window and strong support for coding assistance, agentic workflows, and visual understanding such as charts. It fits long documents, codebases, high-concurrency applications, and tool-assisted tasks.

  • Raisonnement
  • Utilisation d’outils
  • Sortie structurée
  • Contexte long

La description n’est pas encore disponible dans cette langue ; elle est affichée en anglais.

USD

Tarifs

default
Tarif standard
Entrée$0.1492/1M tokensPrix de complétion$0.4476/1M tokensLecture cache$0.014905/1M tokensÉcriture cache$0.189096/1M tokens

Prix publics en dollars US. Le coût final peut varier selon le groupe de compte et le volume d’utilisation.

API

Exemples de code

curl --location --request POST 'https://api.tokenhot.cn/v1/chat/completions' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
    "model": "qwen3.8-flash",
    "messages": [
        {
            "role": "user",
            "content": "What model are you"
        }
    ]
}'

Qwen

Modèles associés

QwenQwenQwen3.8 MAX

Qwen's most capable flagship multimodal reasoning model, built for long-horizon coding, professional work, and multi-stage agent tasks. It accepts text, images, and video, offers a 1M-token context window with up to 128K output, and supports function calling, structured outputs, web search, and a code interpreter. It is best used as a high-quality workhorse for complex tasks rather than for the lowest-cost, lowest-latency batch workloads.

$1.775/1M tokens
QwenQwenQwen3.7 Plus

Designed for agentic workloads, coding, and office productivity, with a stronger combined focus on long context, tool calling, and structured output than Qwen3.6. Compared with flagship models such as GPT, Claude, and Gemini, it is a cost-effective alternative for Chinese, coding, and multi-step tool workflows.

$0.295/1M tokens
QwenQwenQwen3.7 MAX

Designed for agentic workloads, coding, and office productivity, with a stronger combined focus on long context, tool calling, and structured output than Qwen3.6. Compared with flagship models such as GPT, Claude, and Gemini, it is a cost-effective alternative for Chinese, coding, and multi-step tool workflows.

$1.7744/1M tokens
QwenQwenHappyhorse 1.0 T2V

Built for text-to-video, with a stronger audio-native workflow than 早期视频模型, generating dialogue, sound effects, and background music in one pass. Compared with Veo, Kling, and Runway, its differentiator is synchronized audiovisual generation and character-consistent workflows for short drama, ads, e-commerce, and brand marketing.

$0.135/s

FAQ

Questions fréquentes

À quels usages Qwen3.8 Flash convient-il ?

Qwen3.8 Flash convient aux types d’entrée, de sortie et de capacités indiqués sur cette page. Testez vos charges réelles avant un usage critique.

Comment Qwen3.8 Flash est-il facturé ?

Les tarifs figurent ci-dessus. Le coût réel dépend du volume, du groupe de compte et du contenu de la requête.

Comment appeler Qwen3.8 Flash ?

Utilisez l’identifiant du modèle ci-dessus avec un protocole API pris en charge. Des exemples sont affichés lorsqu’ils sont disponibles.

Commencer

Créez avec Qwen3.8 Flash

Utiliser dans la console
TokenHot

La passerelle d'intelligence de frontière. Une API. 127 modèles. Latence de 0.2s. Payez uniquement ce que vous utilisez.

Tous les systèmes normaux · 99.997% de disponibilité

Entreprise

© 2026 TokenHot Inc. — Conçu pour les créateurs.