Gemma-4-31B-it-assistant and Gemma-4-31B-IT-NVFP4 models now available on Amazon SageMaker JumpStart

Published
September 14, 2026
https://aws.amazon.com/about-aws/whats-new/2026/01/gemma-4-31b-it-assistant-gemma-4-31b-it-nvfp4-jumpstart/

New Models Available on Amazon SageMaker JumpStart

Google DeepMind's Gemma-4-31B-it-assistant and NVIDIA's Gemma-4-31B-IT-NVFP4 models are now available on Amazon SageMaker JumpStart, expanding the portfolio of foundation models available to AWS customers.

Gemma-4-31B-it-assistant

Built for multimodal reasoning, coding, and agentic workflows, this model handles text and image inputs, generating text output with a 256K-token context window and support for over 140 languages. It features a hybrid attention mechanism and native function calling for building autonomous agents.

Gemma-4-31B-IT-NVFP4

This model delivers the same capabilities at a fraction of the memory footprint, quantized to 4-bit FP4 precision, reducing memory usage to ~18.5 GB and achieving approximately 2.5x faster inference while retaining 97–99% of the original model's quality. Ideal for cost-efficient, high-throughput production deployments.

What to do

  • Navigate to the SageMaker JumpStart model catalog in the SageMaker console.
  • Use the SageMaker Python SDK to deploy the models to your AWS account.

For more information, see the Amazon SageMaker JumpStart documentation.




If you need further guidance on AWS, our experts are available at AWS@westloop.io. You may also reach us by submitting the Contact Us form.

Follow our blog

Get the latest insights and advice on AWS services from our experts.

By clicking Sign Up you're confirming that you agree with our Terms and Conditions.
Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.