Chats

No previous chats

NVIDIA ~140 tokens/msg
NVIDIA: Nemotron 3.5 Lightning

Hi! I'm NVIDIA: Nemotron 3.5 Lightning. Ask me anything.

NVIDIA: Nemotron 3.5 Lightning requires purchased tokens. Get Tokens | Sign Up - 30K/day Free | Use Free Model Instead
All models with one subscription - see plans →
~140 tokens/msg Enter to send
Model Details

Model Details

Hosted on NVIDIA
Category Chat
Context 262144 tokens
Cost ~140 tokens/msg
3.8 from 38 users of this category

About

NVIDIA: Nemotron 3.5 Lightning is a chat model built by NVIDIA. It accepts up to 262K tokens of context per request. Routed through external models — ~140 tokens per message (50% markup over upstream cost).

Use via API

curl https://api.free.ai/v1/chat/ \
  -H "Authorization: Bearer YOUR_KEY" \
  -d '{"model":"nvidia/nemotron-3.5-lightning"}'
API Docs

FAQ

NVIDIA: Nemotron 3.5 Lightning is a chat model built by NVIDIA. It accepts up to 262K tokens of context per request. Routed through external models — ~140 tokens per message (50% markup over upstream cost).

NVIDIA: Nemotron 3.5 Lightning works well for general conversation, writing assistance, brainstorming, code help, and analysis. Try the sample prompts above to see its style.

About 140 tokens per average message. Free accounts get a 30,000-token daily pool that covers self-hosted chat; premium models are pay-as-you-go, with token top-ups from $1.

It depends on the task. /chat/compare/ lets you send the same prompt to NVIDIA: Nemotron 3.5 Lightning and any other model side-by-side - comparison is the fastest way to decide.

Yes. Outputs are yours - Free.ai does not claim rights to anything you generate.

262,144 tokens.

Replies stream token-by-token within ~1 second. Total response time depends on length and model size - small models stream faster, frontier models trade speed for depth.

Yes. Signed-in users see every chat in /account/?tab=history. You can also share a one-link copy of any conversation via the Share button.

Free.ai does not train models on your conversations. Self-hosted models stay on our GPUs. Premium models route to the upstream provider for inference.

Yes. POST to /v1/chat/ with model="nvidia/nemotron-3.5-lightning" and a messages array. Streaming SSE is supported. Full reference: /api/.

NVIDIA: Nemotron 3.5 Lightning is a premium model served by an external provider, so self-hosting is not available. Free.ai exposes it through token-based pricing.

Free accounts get a 30,000-token daily pool. When that runs out, token top-ups start at $1, pay-as-you-go - no subscription required.

Love Free.ai? Tell your friends!

Rate this page