Home icon

NVIDIA Nemotron 3.5 Lightning now available in Amazon SageMaker JumpStart

Machine Learning Blog



This article announces the availability of NVIDIA Nemotron 3.5 Lightning in Amazon SageMaker JumpStart, a specialized model optimized for high-volume agentic workloads.

  • Delivers up to 4x higher throughput and 30% faster task completion for agent workflows
  • Uses hybrid Mixture-of-Experts architecture with 30B total parameters but only 3B active per forward pass
  • Supports 1M-token context window for long-running sessions without re-grounding
  • Includes DFlash speculative decoding to reduce per-token latency
  • Deploy directly from SageMaker JumpStart, Hugging Face, or SageMaker Python SDK without manual infrastructure configuration
  • Open model available in BF16 and NVFP4 variants for customization and domain-specific fine-tuning
  • Ideal for personal assistants, financial services, cybersecurity, telecom, and retail use cases

Nemotron 3.5 Lightning enables cost-effective agent execution by handling specialized, high-frequency tasks without requiring frontier-model infrastructure.



Go to article

The AWS News Feed is currently looking for gold sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.

Related articles

Aug 11
2026
NVIDIA Nemotron 3.5 Lightning model is now available on Amazon SageMaker JumpStart
Jun 4
2026
NVIDIA Nemotron 3 Ultra now available on Amazon SageMaker JumpStart
Apr 28
2026
NVIDIA Nemotron 3 Nano Omni model now available on Amazon SageMaker JumpStart
Feb 11
2026
NVIDIA Nemotron 3 Nano 30B MoE model is now available in Amazon SageMaker JumpStart

The AWS News Feed is currently looking for silver sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.