DeepSeek

V4 Flash Vision EXP

An experimental multimodal variant of DeepSeek-V4-Flash for image understanding alongside text. A practical choice for visual Q&A, screenshot and document analysis, while its experimental status makes it better suited to evaluation and flexible workflows than strict production-critical paths.

  • Raisonnement
  • Utilisation d’outils
  • Contexte long

La description n’est pas encore disponible dans cette langue ; elle est affichée en anglais.

Utiliser dans la consoledeepseek-v4-flash-vision-exp

USD

Tarifs

default
Tarif standard
Entrée$0.149/1M tokensPrix de complétion$0.298/1M tokensLecture cache$0.0298/1M tokens

Prix publics en dollars US. Le coût final peut varier selon le groupe de compte et le volume d’utilisation.

API

Exemples de code

curl --location --request POST 'https://api.tokenhot.cn/v1/chat/completions' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
    "model": "deepseek-v4-flash-vision-exp",
    "messages": [
        {
            "role": "user",
            "content": "What model are you"
        }
    ]
}'

DeepSeek

Modèles associés

DeepSeekDeepSeekV4 PRO

Built for reasoning, coding, and agentic tool use, with stronger long-context, thinking-mode, and complex-task execution than DeepSeek V3.2. Compared with Chinese peers such as Qwen, GLM, and Kimi, it is especially competitive for math, programming, and logical reasoning tasks.

$1.78/1M tokens
DeepSeekDeepSeekV4 Flash

Built for reasoning, coding, and agentic tool use, with stronger long-context, thinking-mode, and complex-task execution than DeepSeek V3.2. Compared with Chinese peers such as Qwen, GLM, and Kimi, it is especially competitive for math, programming, and logical reasoning tasks.

$0.18/1M tokens
JinaJinaEmbeddings V4

A multimodal embedding model for RAG and search infrastructure, not a chat model. It supports 32K-token text inputs and maps text, images, and visual documents into a shared vector space for knowledge-base retrieval, cross-modal/document search, code retrieval, and semantic matching; dense embeddings default to 2048 dimensions and can be truncated to reduce storage cost.

$0.0374/1M tokens
QwenQwenQwen3.8 MAX

Qwen's most capable flagship multimodal reasoning model, built for long-horizon coding, professional work, and multi-stage agent tasks. It accepts text, images, and video, offers a 1M-token context window with up to 128K output, and supports function calling, structured outputs, web search, and a code interpreter. It is best used as a high-quality workhorse for complex tasks rather than for the lowest-cost, lowest-latency batch workloads.

$1.775/1M tokens

FAQ

Questions fréquentes

À quels usages V4 Flash Vision EXP convient-il ?

V4 Flash Vision EXP convient aux types d’entrée, de sortie et de capacités indiqués sur cette page. Testez vos charges réelles avant un usage critique.

Comment V4 Flash Vision EXP est-il facturé ?

Les tarifs figurent ci-dessus. Le coût réel dépend du volume, du groupe de compte et du contenu de la requête.

Comment appeler V4 Flash Vision EXP ?

Utilisez l’identifiant du modèle ci-dessus avec un protocole API pris en charge. Des exemples sont affichés lorsqu’ils sont disponibles.

Commencer

Créez avec V4 Flash Vision EXP

Utiliser dans la console
TokenHot

La passerelle d'intelligence de frontière. Une API. 127 modèles. Latence de 0.2s. Payez uniquement ce que vous utilisez.

Tous les systèmes normaux · 99.997% de disponibilité

Entreprise

© 2026 TokenHot Inc. — Conçu pour les créateurs.