This product has charges associated with it for support from the seller. Built on openSUSE Linux, this product provides private AI using the Gemma 3 model with 1 billion parameters. MultiCortex HPC (High-Performance Computing) allows you to boost your AI's response quality. This is a plug-and-play, low-cost product with no token fees.
Gemma 3 is an open language artificial intelligence model developed by Google. Designed to be efficient and versatile, Gemma 3 provides advanced capabilities that support a wide range of artificial intelligence applications.
Main Features of Gemma 3
Multimodality. Gemma 3 can process inputs such as text images and short videos allowing the creation of applications that analyze and interpret various types of media
Multilingual support. With compatibility for more than one hundred forty languages Gemma 3 enables global communication and helps adapt applications for different markets and audiences
Extended context windows. The model supports context windows of up to one hundred twenty eight thousand tokens making it possible to process complex information and manage tasks that require deeper context
Computational efficiency. Gemma 3 is built to run efficiently on devices with limited resources and is considered the most powerful artificial intelligence model capable of running on a single graphics processing unit
Advanced reasoning capabilities. The model provides strong performance in mathematics logical reasoning and conversation including structured outputs and function calling which expands the possibilities for interaction and automation
Highlights
Boost your private AI system with MultiCortex HPC and leave the thousands of AI updates to us
Enjoy full technical compliance with complete data control in the hands of your company
No token fees, enhanced performance, and lower consumption of computing resources
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Try this product free for 31 days according to the free trial terms set by the vendor. Usage-based pricing is in effect for usage beyond the free trial terms. Your free trial gets automatically converted to a paid subscription when the trial ends, but may be canceled any time before that.
Pricing is based on actual usage, with charges varying according to how much you consume. Subscriptions have no end date and may be canceled any time.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
If you are an AWS Free Tier customer with a free plan, you are eligible to subscribe to this offer. You can use free credits to cover the cost of eligible AWS infrastructure. See AWS Free Tier for more details. If you created an AWS account before July 15th, 2025, and qualify for the Legacy AWS Free Tier, Amazon EC2 charges for Micro instances are free for up to 750 hours per month. See Legacy AWS Free Tier for more details.
You pay by the hour for the EC2 instance type you run this software on. There is no upfront commitment; billing is usage-based and scales with the hours you use. Each dimension maps to a specific AWS instance size, so pricing rises with the compute, memory, storage, or accelerator capacity of the instance you pick. Options span general-purpose, compute-optimized, memory-optimized, storage-optimized, and GPU or accelerator instances. Smaller instances suit light workloads, while larger and specialized instances handle heavier processing. You choose the instance that fits your performance needs and pay only for the hours it runs.
Top-of-mind questions for buyers
What am I actually paying for with each hourly instance dimension?
Each dimension bills the software running on a specific EC2 instance type, charged per hour that instance runs. The instance size you pick sets the price, based on its CPU, memory, storage, and any GPU or accelerator. You pay separately for the underlying AWS compute.
Am I charged when my instance is stopped or paused?
Hourly software charges apply only while the instance runs. A stopped or paused instance stops accruing software charges. You may still pay AWS for attached storage while the instance is stopped, but the software meters running hours only.
How does this software use both CPU and GPU instance types in the list?
The software supports heterogeneous computing, processing AI across CPUs, GPUs, and other accelerators together. That is why the dimension list includes general-purpose, compute, memory, storage, and GPU or accelerator instances. You pick the instance whose processors match your workload and pay for its running hours.
www.multicortex.ai
Helpful?
Vendor refund policy
If you need to request a refund for software sold by Amazon Web Services, LLC, please contact AWS Customer Service.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
An AMI is a virtual image that provides the information required to launch an instance. Amazon EC2 (Elastic Compute Cloud) instances are virtual servers on which you can run your applications and workloads, offering varying combinations of CPU, memory, storage, and networking resources. You can launch as many instances from as many different AMIs as you need.
Version release notes
Gemma 1b model and all packages installed automatically.
Additional details
Usage instructions
Additional details Usage instructions
<br><br>
To use the product, follow these steps: 1. Open a web browser and go to the application at the following address: https://<EC2_Instance_Public_DNS>/index.html. 2. Log in using the credentials below: - Username: ec2-user - Password: the instance_id of your instance. After launching an EC2 instance, wait about 15 minutes for the system to complete the automatic configuration and download of the LLM template. After this time, you can access the chat interface by typing the instance's IP address followed by port 7000 into your browser. For example: http://10.21.103:7000.
Support
Vendor support
The support service called MultiCortex, developed based on the openSUSE operating system, aims to provide the necessary resources for the efficient execution of AI models that are always up to date and aligned with the latest technological innovations in the area, offering a robust and secure environment for the development and application of AI-based solutions.
AWS infrastructure support
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
A self-hosted, production-ready Gemma 4 26B model, with 4 billion active parameters deployed into your AWS environment with a single click. Because everything runs entirely within your private cloud, your data stays secure, isolated, and fully under your control. Best of all, unlimited tokens.
This product has charges associated with it for seller support. Run & Manage latest LLMs locally, privately, securely and cost-effectively without any vendor lock-in.
This VM solution comes pre-loaded with LLaMA, Mistral, Gemma, DeepSeek, & Qwen models along with Open-WebUI as an intuitive UI to interact with the LLMs and Ollama to install new models as needed.
This product has charges associated with it for seller support. Run & Manage latest LLMs locally, privately, securely and cost-effectively without any vendor lock-in. This VM solution comes with GPU support , pre-loaded with LLaMA, Mistral, Gemma, DeepSeek, & Qwen models along with Open-WebUI as an intuitive UI to interact with the LLMs and Ollama to install new models as needed.
Initial release of MedGemma 27B Multimodal (google/medgemma-27b-it) on AWS Marketplace. This version delivers Google's most capable open medical AI model, supporting both medical image comprehension and clinical text reasoning including FHIR-based EHR data, across radiology, dermatology, histopathology, and ophthalmology modalities.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.