qwen3-vl-8b-thinking

qwen3-vl-8b-thinking (custom) — 128K context, Reasoning tier. Prix entrée: $0.5/M · Prix sortie: $5/M · Latence 48ms. Routed via Routara OpenAI-compatible endpoint with multi-region failover and metered billing.

Prix entrée: $0.5/M · Prix sortie: $5/M · TTFT: 48ms

Specifications

  • Éditeur: custom
  • Catégorie: Reasoning
  • Fenêtre de contexte: 128K
  • Prix entrée: $0.5 / 1M
  • Prix sortie: $5 / 1M
  • Latence: 48 ms
  • SLA: S+

Typical use cases

  • Génération image/vidéo
  • Chatbots multilingues
  • RAG et appels d’outils

FAQ

  • Tarification de qwen3-vl-8b-thinking sur Routara ? — Entrée $0.5/M, sortie $5/M — facturation à l’usage.
  • qwen3-vl-8b-thinking est-il compatible OpenAI ? — Oui — base_url : https://api.routara.ai/v1
  • qwen3-vl-8b-thinking supporte le streaming ? — Oui si la route est live (stream: true).

Related models

  • MiniMax-File-Upload (/detail/minimax-file-upload)
  • MiniMax-Voice-Clone (/detail/minimax-voice-clone)
  • suno_uploads (/detail/suno-uploads)

Quick integration

  • curl https://api.routara.ai/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"qwen3-vl-8b-thinking","messages":[{"role":"user","content":"Hello"}],"stream":true}'