Skip to main content
Velqa

Models available on velqa.dev

All the models below are reachable through velqa.dev's OpenAI-compatible API (https://api.velqa.dev/v1/chat/completions) and Anthropic-compatible API (https://api.velqa.dev/v1/messages). Checkout is Stripe, in USD, by international card. Moroccan-issued cards are not accepted here: customers in Morocco are served by velqa.ma, where dirham checkout by Moroccan card opens soon.

Model table

Model IDSource providerUse caseType
glm-4.7-flashZhipu AIFast chat, everyday tasksFast/Chat
deepseek-v4-flashDeepSeekCheap high-volume, light code reviewFast/Chat
hy3TencentGeneralist and agenticChat/Agentic
glm-4.7Zhipu AIChat and agentic codingChat/Coding
minimax-m3MiniMaxAgentic reasoningCoding
mimo-v2.5XiaomiText generalist, chat and agenticChat/Agentic
kimi-k2.6Moonshot AIAgentic coding, long sessions, 131k contextCoding
glm-5.2Zhipu AILong autonomous tasks, 128k contextCoding
deepseek-v4-proDeepSeekFlagship heavy reasoning, 160k contextPremium
qwen3.7-maxAlibabaPremium reasoning, 256k contextPremium
qwen3.8-maxAlibabaFlagship coding and professional work, 256k contextPremium
kimi-k3Moonshot AIHeavy agentic reasoning, 1,048,576-token contextPremium
bge-m3BAAIMultilingual embeddingsEmbeddings
qwen3-embedding-8bAlibabaHigh-quality multilingual embeddingsEmbeddings
qwen3-reranker-8bAlibabaMultilingual rerankingRerank
whisper-large-v3OpenAI (via DeepInfra)Multilingual audio transcription, max 10 minAudio (STT)
kokoro-ttshexgrad (via DeepInfra)Text-to-speech, 8 languages, no ArabicAudio (TTS)
qwen3-ttsAlibaba (via DeepInfra)Text-to-speech with Arabic and tone controlAudio (TTS)
flux-1-schnellBlack Forest Labs (via DeepInfra)Text-to-image, draft/volume tierImage
flux-2-klein-4bBlack Forest Labs (via DeepInfra)Text-to-image generationImage
flux-2-proBlack Forest Labs (via DeepInfra)Text-to-image, quality tierImage
fastwan-t2v-1.3bFastVideo (via DeepInfra)Text-to-video, 480p, seconds not minutesVideo
wan2.2-t2v-a14bWan-AI (via DeepInfra)Text-to-video generation, 5 secondsVideo

Availability per plan

ModelStarterDevProRecharge Boost
glm-4.7-flashYesYesYesYes
deepseek-v4-flashYesYesYesYes
hy3YesYesYesYes
glm-4.7YesYesYesYes
minimax-m3NoYesYesYes
mimo-v2.5NoYesYesYes
kimi-k2.6NoYesYesYes
glm-5.2NoYesYesYes
deepseek-v4-proNoNoYesYes
qwen3.7-maxNoNoYesYes
qwen3.8-maxNoNoNoYes
kimi-k3NoNoNoYes
bge-m3YesYesYesYes
qwen3-embedding-8bYesYesYesYes
qwen3-reranker-8bYesYesYesYes
whisper-large-v3NoNoNoYes
kokoro-ttsNoNoNoYes
qwen3-ttsNoNoNoYes
flux-1-schnellNoNoNoYes
flux-2-klein-4bNoNoNoYes
flux-2-proNoNoNoYes
fastwan-t2v-1.3bNoNoNoYes
wan2.2-t2v-a14bNoNoNoYes

Boost-only models

kimi-k3 and qwen3.8-max were added on 2026-08-09 and are the first chat models sold outside every subscription. Their per-token cost is beyond what a fixed-price plan can absorb — kimi-k3 alone costs several times an included model's budget for the same turn — so they are sold per token against a prepaid Boost balance instead of being folded into a tier and paid for by everyone else's quota.

  • kimi-k3: heavy agentic reasoning over a 1,048,576-token window. Generation is slow (single-digit tokens per second), so it suits work where reasoning quality outweighs latency, not interactive chat.
  • qwen3.8-max: generalist flagship aimed at code and professional work, 256k context on the primary route.

Calling either one with a subscription key that has no Boost balance is rejected: they are not in that key's allowed-model list. Per-token prices for both are on the pricing page.

Media models added on 2026-08-09

Four at once, all on DeepInfra like the rest of the media catalogue — no new provider relationship.

  • flux-1-schnell ($0.002/image): 15x cheaper than the default tier. This is the drafting model, for the runs where paying $0.03 an attempt makes no sense.
  • flux-2-pro ($0.032/image): clearly better than flux-2-klein-4b for ~7% more. flux-2-klein-9b costs exactly the same and adds nothing, so it was not added.
  • qwen3-tts ($44/M characters): fills the gap kokoro-tts leaves — kokoro speaks 8 languages, none of them Arabic. Also adds forced language and a natural-language tone directive. At 22x kokoro's price it is for the cases kokoro cannot cover, not the default.
  • fastwan-t2v-1.3b ($0.005/second): 30x cheaper than wan2.2-t2v-a14b and, more to the point, ~100x faster — a measured 2.5 seconds of generation against 4m35s for the same 5-second clip. It renders at 480p with less detail: this is the iteration tier, wan2.2 stays the quality tier.

Notes

  • Model identifiers (model_id) are stable — these are the ones to use in your tool configs.
  • The Starter plan is designed for students and exploration; advanced coding models require the Dev plan at minimum.
  • Recharge Boost mode (pay-as-you-go) gives access to every model from a prepaid credit balance, with no subscription.
  • Current USD pricing and checkout availability are shown in the dashboard; customers in Morocco are served by velqa.ma, where dirham checkout by Moroccan card opens soon.