NVIDIA Nemotron 3.5 Lightning model now available on Amazon SageMaker JumpStart
Amazon SageMaker JumpStart now offers NVIDIA's Nemotron 3.5 Lightning model, accelerating persistent agent workloads and rapid task execution.
NVIDIA's Nemotron 3.5 Lightning model is now available on Amazon SageMaker JumpStart, providing AWS customers with access to the fastest open model in its class for persistent agent workloads and rapid task execution. Engineered for persistent agents and high-throughput enterprise automation across domains including personal assistants, financial document processing, cybersecurity triage, and telecom operations, it features a hybrid Mixture-of-Experts (MoE) architecture with 30B total parameters and only 3B active per forward pass, achieving up to 4x the throughput (~410 tokens/sec) and 30% faster task completion over comparable models. Distilled from Nemotron 3 Ultra, it handles up to 1M tokens of context via DFlash speculative decoding and integrates directly with popular agent harnesses. The model is fully open-trained on open datasets, allowing enterprises to post-train for their own tools, workflows, and policies, and deploy with complete ownership across edge, on-premises, or cloud infrastructure. With SageMaker JumpStart, customers can deploy this model in a few clicks to power their specific AI workloads.