⚛️ Nemotron-3-Nano 30B

NVIDIA · Hybrid Mamba-2 + Attention · ZeroGPU · real Triton kernels ✅

First message loads the model (~60s). ZeroGPU shared GPU — response time varies by hardware allocated.

Thinking mode

On: model reasons before answering (slower)

64 2048
0 1.5
0.1 1
Try these