Amazon SageMaker AI now supports serverless model customization for NVIDIA Nemotron 3.5 Lightning, an open-weight model designed for high-volume agentic workloads. Organizations can adapt the model using supervised fine-tuning for domain-specific accuracy, Direct Preference Optimization for preferred response behavior, and reinforcement fine-tuning for specialized tasks.
Nemotron 3.5 Lightning uses a hybrid Mixture-of-Experts architecture with 30 billion total parameters and 3 billion active parameters. SageMaker manages infrastructure provisioning and training orchestration, removing the need to manage training clusters directly.
The capability is available in Northern Virginia, Oregon, Tokyo, and Ireland, with usage-based pricing for customization workloads.




