NVIDIA Nemotron 3.5 Lightning Now on Amazon SageMaker JumpStart

Venu Kanamatareddy· August 17, 2026 View original

Key takeaways

  • NVIDIA Nemotron 3.5 Lightning is now available on Amazon SageMaker JumpStart.
  • It is an open, 30B MoE model optimized for high-volume agentic workloads.
  • The model offers up to 4x higher throughput and 30% faster task completion.
  • Its availability simplifies deployment for AWS users building AI agents.

Who benefits

Software DevelopmentCloud ServicesCustomer ServiceRoboticsFinancial Services

Summary

NVIDIA's Nemotron 3.5 Lightning, an open model designed for high-volume agentic workloads, is now accessible via Amazon SageMaker JumpStart. This 30B Mixture-of-Experts model offers significantly higher throughput and faster task completion for always-on AI agents.

NVIDIA has announced the availability of its Nemotron 3.5 Lightning model on Amazon SageMaker JumpStart. This powerful open model is specifically engineered to handle demanding agentic workloads, making it suitable for applications requiring continuous, high-volume AI interactions. The model is a 30-billion parameter Mixture-of-Experts (MoE) architecture, with 3 billion parameters actively engaged during inference. The integration into SageMaker JumpStart simplifies deployment for developers and enterprises leveraging AWS infrastructure. Key performance benefits include up to four times higher throughput and a 30% reduction in task completion time, which are critical for maintaining responsive and efficient AI agents. This release aims to accelerate the development and scaling of sophisticated AI agent applications.

Why it matters

Professionals can now easily access and deploy a high-performance AI model optimized for agentic workloads directly within their AWS environment, potentially boosting efficiency and reducing operational costs for AI-driven services.

How to implement this in your domain

  1. 1Access SageMaker JumpStart to find and select the Nemotron 3.5 Lightning model.
  2. 2Deploy the model within your existing AWS infrastructure for agentic applications.
  3. 3Integrate the model into your AI agent workflows to leverage its high throughput.
  4. 4Monitor performance improvements in task completion times and resource utilization.

Original post by Venu Kanamatareddy

"NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to deploy the 30B Mixture-of-Experts model (3B active), which delivers up to 4x higher throughput and up to 30% faster task co…"

View on X

Originally posted by Venu Kanamatareddy on X · view source

Want to go deeper?

Turn these trends into skills with Learnijoy's hands-on AI & tech courses.

Explore courses