Model Details
About
Nous: Hermes 4 70B is a chat model built by Nousresearch. It accepts up to 131K tokens of context per request. Routed through external models — ~208 tokens per message (50% markup over upstream cost).
Use via API
curl https://api.free.ai/v1/chat/ \
-H "Authorization: Bearer YOUR_KEY" \
-d '{"model":"nousresearch/hermes-4-70b"}'
Compare
FAQ
Nous: Hermes 4 70B is a chat model built by Nousresearch. It accepts up to 131K tokens of context per request. Routed through external models — ~208 tokens per message (50% markup over upstream cost).
Nous: Hermes 4 70B works well for general conversation, writing assistance, brainstorming, code help, and analysis. Try the sample prompts above to see its style.
About 208 tokens per average message. Free accounts get a 30,000-token daily pool that covers self-hosted chat; premium models are pay-as-you-go, with token top-ups from $1.
It depends on the task. /chat/compare/ lets you send the same prompt to Nous: Hermes 4 70B and any other model side-by-side - comparison is the fastest way to decide.
Yes. Outputs are yours - Free.ai does not claim rights to anything you generate.
131,072 tokens.
Replies stream token-by-token within ~1 second. Total response time depends on length and model size - small models stream faster, frontier models trade speed for depth.
Yes. Signed-in users see every chat in /account/?tab=history. You can also share a one-link copy of any conversation via the Share button.
Free.ai does not train models on your conversations. Self-hosted models stay on our GPUs. Premium models route to the upstream provider for inference.
Yes. POST to /v1/chat/ with model="nousresearch/hermes-4-70b" and a messages array. Streaming SSE is supported. Full reference: /api/.
Nous: Hermes 4 70B is a premium model served by an external provider, so self-hosting is not available. Free.ai exposes it through token-based pricing.
Free accounts get a 30,000-token daily pool. When that runs out, token top-ups start at $1, pay-as-you-go - no subscription required.
Chat with Nous: Hermes 4 70B
What is Nous: Hermes 4 70B?
Nous: Hermes 4 70B is a chat model built by Nousresearch. It accepts up to 131K tokens of context per request. Routed through external models — ~208 tokens per message (50% markup over upstream cost).
Why use Nous: Hermes 4 70B for chat?
Streaming responses
Replies stream token-by-token within ~1 second of pressing Send. No idle waiting.
Saved history
Signed-in users see every chat in /account/?tab=history with one-click share links.
Compare side by side
Send the same prompt to up to 4 models at /chat/compare/ and judge the outputs side by side.
Commercial use OK
Outputs are yours. Use them in apps, ads, docs, or anything else without attribution.
Sample prompts
Pricing
Premium chat. Cost is per message - typically ~208 tokens, shown live before you send. Self-hosted chat runs free on your 30,000-token daily pool; premium models are pay-as-you-go, with top-ups from $1.
Compare to alternatives
See all chat models → · Compare up to 4 chat models side-by-side →