This product has charges associated with it for support from the seller. Built on openSUSE Linux, this product provides private AI using the Qwen 2.5 model with 0.5 billion parameters. MultiCortex HPC (High-Performance Computing) allows you to boost your AI's response quality. This is a plug-and-play, low-cost product with no token fees.
Qwen is the large language model and large multimodal model series of the Qwen Team, Alibaba Group. Now the large language models have been upgraded to Qwen2.5. Both language models and multimodal models are pretrained on large-scale multilingual and multimodal data and post-trained on quality data for aligning to human preferences. Qwen is capable of natural language understanding, text generation, vision understanding, audio understanding, tool use, role play, playing as AI agent, etc.
Qwen 2.5 has the following features:
Pretrained on our latest large-scale dataset, encompassing up to 18T tokens.
Significant improvements in instruction following, generating long texts (over 8K tokens), understanding structured data (e.g, tables), and generating structured outputs especially JSON.
More resilient to the diversity of system prompts, enhancing role-play implementation and condition-setting for chatbots.
Context length support up to 128K tokens and can generate up to 8K tokens.
Multilingual support for over 29 languages, including Chinese, English, French, Spanish, Portuguese, German, Italian, Russian, Japanese, Korean, Vietnamese, Thai, Arabic, and more.
Highlights
Boost your private AI system with MultiCortex HPC and leave the thousands of AI updates to us
Enjoy full technical compliance with complete data control in the hands of your company
No token fees, enhanced performance, and lower consumption of computing resources
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Try this product free for 31 days according to the free trial terms set by the vendor. Usage-based pricing is in effect for usage beyond the free trial terms. Your free trial gets automatically converted to a paid subscription when the trial ends, but may be canceled any time before that.
Pricing is based on actual usage, with charges varying according to how much you consume. Subscriptions have no end date and may be canceled any time.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
If you are an AWS Free Tier customer with a free plan, you are eligible to subscribe to this offer. You can use free credits to cover the cost of eligible AWS infrastructure. See AWS Free Tier for more details. If you created an AWS account before July 15th, 2025, and qualify for the Legacy AWS Free Tier, Amazon EC2 charges for Micro instances are free for up to 750 hours per month. See Legacy AWS Free Tier for more details.
You pay by the hour for the AWS EC2 instance that runs this software. Each dimension maps to one instance type, so pricing scales with the compute you select. Options range from small burstable instances like t3.nano and t2.micro to memory-heavy, storage-heavy, and GPU-accelerated instances such as p5.48xlarge, and large bare-metal or high-memory u-series machines. Larger instances with more CPU, memory, or accelerators carry higher hourly rates. You choose the instance that fits your workload, and billing follows how many hours you run it. There is no upfront commitment.
Top-of-mind questions for buyers
What exactly does one hourly unit cover for a given instance type?
Each unit is one hour of running one AWS EC2 instance of the type named in that dimension. The software runs on that instance, and you pay the software rate plus the underlying AWS infrastructure cost for every hour it runs.
Am I charged when my instance is stopped or idle?
Software charges meter running hours only. A fully stopped instance stops accruing the hourly software rate. You may still pay AWS for attached storage while the instance is stopped, but the software licence bills active running time.
How do I decide which instance type to pick for this model?
Each dimension maps to one instance type, so you match your workload to CPU, memory, or accelerator needs. GPU-accelerated types like p5.48xlarge or g5.xlarge suit heavier inference, while burstable types like t3.nano suit light testing. The product supports processing across GPU and CPU accelerators together.
www.multicortex.ai
Helpful?
Vendor refund policy
If you need to request a refund for software sold by Amazon Web Services, LLC, please contact AWS Customer Service.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
An AMI is a virtual image that provides the information required to launch an instance. Amazon EC2 (Elastic Compute Cloud) instances are virtual servers on which you can run your applications and workloads, offering varying combinations of CPU, memory, storage, and networking resources. You can launch as many instances from as many different AMIs as you need.
Version release notes
Qwen 2.5 0,5b model and all packages installed automatically.
Additional details
Usage instructions
Additional details Usage instructions
<br><br>
To use the product, follow these steps: 1. Open a web browser and go to the application at the following address: https://<EC2_Instance_Public_DNS>/index.html. 2. Log in using the credentials below: - Username: ec2-user - Password: the instance_id of your instance. After launching an EC2 instance, wait about 15 minutes for the system to complete the automatic configuration and download of the LLM template. After this time, you can access the chat interface by typing the instance's IP address followed by port 7000 into your browser. For example: http://10.21.103:7000.
Support
Vendor support
The support service called MultiCortex, developed based on the openSUSE operating system, aims to provide the necessary resources for the efficient execution of AI models that are always up to date and aligned with the latest technological innovations in the area, offering a robust and secure environment for the development and application of AI-based solutions.
AWS infrastructure support
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
A self-hosted production-ready DeepSeek-R1-Distill-Qwen-1.5B model (https://huggingface.co/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B) running seamlessly in your private AWS cloud! With an easy single-click installation, set up all the essential infrastructure in your own cloud environment hassle-free. Plus, you will have quick access to an API endpoint that is ready for your queries and scales automatically based on your needs. Best of all, with the service operating solely in your cloud, your data remains completely secure and confidential, never leaving your private space. Experience peace of mind and unleash the full potential of DeepSeek R1 models today!
A self-hosted, production-ready Qwen 3.6 35B model, with 3 billion active parameters deployed into your AWS environment with a single click. Because everything runs entirely within your private cloud, your data stays secure, isolated, and fully under your control. Best of all, unlimited tokens.
A self-hosted, production-ready Qwen 2.5 7B model deployed into your AWS environment with a single click. Because everything runs entirely within your private cloud, your data stays secure, isolated, and fully under your control. Best of all, unlimited tokens.
This product has charges associated with it for seller support. Run & Manage latest LLMs locally, privately, securely and cost-effectively without any vendor lock-in.
This VM solution comes pre-loaded with LLaMA, Mistral, Gemma, DeepSeek, & Qwen models along with Open-WebUI as an intuitive UI to interact with the LLMs and Ollama to install new models as needed.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.