qwen3-vl-32b-thinking

qwen3-vl-32b-thinking (qwen3-vl-32b-thinking) — 128K context, Reasoning tier. Prix entrée: $0.3/M · Prix sortie: $3/M · Latence 48ms. Routed via Routara OpenAI-compatible endpoint with multi-region failover and metered billing.

Prix entrée: $0.3/M · Prix sortie: $3/M · TTFT: 48ms

Specifications

  • Éditeur: qwen3-vl-32b-thinking
  • Catégorie: Reasoning
  • Fenêtre de contexte: 128K
  • Prix entrée: $0.3 / 1M
  • Prix sortie: $3 / 1M
  • Latence: 48 ms
  • SLA: S+

Typical use cases

  • Génération image/vidéo
  • Chatbots multilingues
  • RAG et appels d’outils

FAQ

  • Tarification de qwen3-vl-32b-thinking sur Routara ? — Entrée $0.3/M, sortie $3/M — facturation à l’usage.
  • qwen3-vl-32b-thinking est-il compatible OpenAI ? — Oui — base_url : https://api.routara.ai/v1
  • qwen3-vl-32b-thinking supporte le streaming ? — Oui si la route est live (stream: true).

Related models

  • audio1.0 (/detail/audio1-0)
  • sao10k/l3-8b-lunaris (/detail/sao10k-l3-8b-lunaris)
  • Sao10K/L3-8B-Stheno-v3.2 (/detail/sao10k-l3-8b-stheno-v3-2)

Quick integration

  • curl https://api.routara.ai/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"qwen3-vl-32b-thinking","messages":[{"role":"user","content":"Hello"}],"stream":true}'