Model Details
About
Qwen: Qwen3.8 2.4T A95B is a chat model built by Qwen. It accepts up to 1M tokens of context per request. Routed through external models — ~3,150 tokens per message (50% markup over upstream cost).
Use via API
curl https://api.free.ai/v1/chat/ \
-H "Authorization: Bearer YOUR_KEY" \
-d '{"model":"qwen/qwen3.8-2.4t-a95b"}'
Compare
FAQ
Qwen: Qwen3.8 2.4T A95B is a chat model built by Qwen. It accepts up to 1M tokens of context per request. Routed through external models — ~3,150 tokens per message (50% markup over upstream cost).
Qwen: Qwen3.8 2.4T A95B works well for general conversation, writing assistance, brainstorming, code help, and analysis. Try the sample prompts above to see its style.
About 3,150 tokens per average message. Free accounts get a 30,000-token daily pool that covers self-hosted chat; premium models are pay-as-you-go, with token top-ups from $1.
It depends on the task. /chat/compare/ lets you send the same prompt to Qwen: Qwen3.8 2.4T A95B and any other model side-by-side - comparison is the fastest way to decide.
Yes. Outputs are yours - Free.ai does not claim rights to anything you generate.
1,000,000 tokens.
Replies stream token-by-token within ~1 second. Total response time depends on length and model size - small models stream faster, frontier models trade speed for depth.
Yes. Signed-in users see every chat in /account/?tab=history. You can also share a one-link copy of any conversation via the Share button.
Free.ai does not train models on your conversations. Self-hosted models stay on our GPUs. Premium models route to the upstream provider for inference.
Yes. POST to /v1/chat/ with model="qwen/qwen3.8-2.4t-a95b" and a messages array. Streaming SSE is supported. Full reference: /api/.
Qwen: Qwen3.8 2.4T A95B is a premium model served by an external provider, so self-hosting is not available. Free.ai exposes it through token-based pricing.
Free accounts get a 30,000-token daily pool. When that runs out, token top-ups start at $1, pay-as-you-go - no subscription required.
Chat with Qwen: Qwen3.8 2.4T A95B
What is Qwen: Qwen3.8 2.4T A95B?
Qwen: Qwen3.8 2.4T A95B is a chat model built by Qwen. It accepts up to 1M tokens of context per request. Routed through external models — ~3,150 tokens per message (50% markup over upstream cost).
Why use Qwen: Qwen3.8 2.4T A95B for chat?
Streaming responses
Replies stream token-by-token within ~1 second of pressing Send. No idle waiting.
Saved history
Signed-in users see every chat in /account/?tab=history with one-click share links.
Compare side by side
Send the same prompt to up to 4 models at /chat/compare/ and judge the outputs side by side.
Commercial use OK
Outputs are yours. Use them in apps, ads, docs, or anything else without attribution.
Sample prompts
Pricing
Premium chat. Cost is per message - typically ~3150 tokens, shown live before you send. Self-hosted chat runs free on your 30,000-token daily pool; premium models are pay-as-you-go, with top-ups from $1.
Compare to alternatives
See all chat models → · Compare up to 4 chat models side-by-side →