Llama-3.3-Nemotron-Super-49B-v1.5

NVIDIA's 49B reasoning/chat model, running 4-bit quantized on ZeroGPU. Reasoning ON shows a <think> trace before the answer.

Examples
Reasoning (thinking) mode Max new tokens Temperature
128 4096
0.1 1.2