256K contextTool callingReasoning

Pricing

₹ per 1M tokens · as of 29 August 2026
Input₹189.10
Output₹567.30
Cached input₹18.91

Flat rate — every request bills the same. Full catalog and comparisons on the pricing comparison page.

Where it fits

Not sure? Vidman AI Adaptive routes to the best model per request, including this one.

Call it now — OpenAI-compatible
curl https://api.vidman.ai/v1/chat/completions \
  -H "Authorization: Bearer $VIDMAN_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen3.8-2.4t-a95b",
    "messages": [{ "role": "user", "content": "Hello" }]
  }'

Existing OpenAI SDK code works by changing the base URL to https://api.vidman.ai/v1 and the key.