Ecosystem
August 18, 2026

NVIDIA Nemotron 3.5 Lightning arrives on Amazon SageMaker JumpStart

AWS has added NVIDIA Nemotron 3.5 Lightning to SageMaker JumpStart, enabling developers to deploy a fast open model with 1 million token context for high-volume agentic AI workloads.

AWS has made NVIDIA Nemotron 3.5 Lightning available through Amazon SageMaker JumpStart, simplifying deployment for high-volume agentic AI workloads. The open model uses a hybrid mixture-of-experts architecture with 30 billion total parameters and 3 billion active parameters, allowing it to run on a single supported GPU.

It supports context windows up to 1 million tokens and DFlash speculative decoding. NVIDIA reports up to four times higher throughput and 30% faster task completion for specialized agent workloads.

Developers can deploy BF16 and NVFP4 variants through SageMaker JumpStart, Hugging Face, or the SageMaker Python SDK without configuring serving infrastructure manually.

#
AWS

Read Our Content

See All Blogs
LLM Models

LLM testing of Muse Spark 1.3

Sarankumar S

September 11, 2026
Read more
LLM Models

Harness engineering for AI agents: The missing layer for production deployment

Sarankumar S

September 11, 2026
Read more