Unlike standard Deep Learning AMIs, this single image combines RAG pipelines, LLM fine-tuning, medical imaging, and pre-wired Bedrock and SageMaker integrations - ready in minutes, not hours.
Building a production-ready Generative AI environment from scratch means hours of dependency resolution, version-conflict debugging, and manual AWS service integration. This preconfigured Amazon Linux 2023 AMI eliminates that complexity by delivering 20+ pre-tested AI/ML frameworks, AWS-native connectors, and runnable sample workflows in a single image - ready to use within minutes of launch.
Unlike vanilla Deep Learning AMIs or manual pip-install approaches, this stack uniquely combines RAG-ready vector search (FAISS), medical imaging (MONAI), LLM fine-tuning (LoRA/PEFT), and pre-wired Bedrock and SageMaker integrations that no single AWS-managed AMI provides.
Who Benefits
Researchers save iteration time with JupyterLab, RStudio, and pre-loaded sample notebooks
Enterprise teams reduce onboarding friction with a verified, consistent environment across team members
Developers accelerate prototyping with Docker-ready microservices and pre-configured IDE tooling
Key Capabilities
Generative AI - Ready in Minutes
Hugging Face Transformers, LangChain, and LlamaIndex pre-installed and verified compatible
FAISS configured for vector search and RAG pipeline development
LoRA and PEFT frameworks for efficient LLM fine-tuning on GPU instances
Sample RAG chatbot workflow using Bedrock + OpenSearch included as a runnable notebook
AWS-Native Integrations
Bedrock SDK pre-configured for foundation model access
SageMaker SDK ready for training and deployment workflows
OpenSearch connectors for building scalable retrieval-augmented generation applications
Full Development Stack
Visual Studio Code, PyCharm Community Edition, JupyterLab, and RStudio Desktop and Server
PyTorch, TensorFlow, scikit-learn, PySpark, Dask, and Vowpal Wabbit
Docker and Docker Compose for containerized deployments
Anaconda environment management for reproducible experiments
Google Chrome, Git, and AWS CLI preinstalled for immediate productivity
LibreOffice for document editing and reporting
Sample Workflows Included
RAG chatbot with Bedrock + OpenSearch
Multimodal AI pipelines
MONAI medical imaging use cases
Quick-Start Deployment
Subscribe to the AMI and select your instance (GPU: g5, p4d, p5 for LLM fine-tuning; CPU for lighter tasks)
Configure your VPC and security groups to restrict inbound access to necessary ports
Launch the instance and connect via Amazon NICE DCV remote desktop
Open JupyterLab and run the included RAG chatbot notebook to validate your environment
For Bedrock and SageMaker workflows, ensure your IAM role has the appropriate service permissions attached.
Pricing
Software charges apply per hour on top of standard EC2 infrastructure costs. See the Pricing tab on this page for detailed rate information by instance type. To evaluate the AMI at minimal cost, launch on a smaller CPU instance to explore the environment and sample notebooks before scaling to GPU instances for production workloads.
Security Considerations
Amazon NICE DCV connections use TLS encryption for secure remote access. We recommend launching this AMI within a properly configured VPC with security groups restricting inbound access to necessary ports only. For GPU workloads involving sensitive data, enable EBS encryption on attached volumes.
Technical Specifications
Operating System: Amazon Linux 2023
Languages: Python 3.x, R
Recommended Instances: GPU instances (g5, p4d, p5 families) for LLM fine-tuning, PyTorch, and TensorFlow workloads. CPU instances suitable for data preparation, LangChain development, and lighter ML tasks.
RelevanceLab builds cloud-native solutions that accelerate AI adoption on AWS. The team specializes in preconfigured AI/ML environments designed for rapid deployment on AWS infrastructure. For a guided walkthrough or pilot consultation, contact the team through the support channel listed on this page.
Highlights
Unlike standard AWS Deep Learning AMIs, this single image uniquely combines RAG ready vector search (FAISS), medical imaging (MONAI), LLM finetuning (LoRA/PEFT), and pre wired Bedrock and SageMaker integrations eliminating hours of manual dependency resolution and version conflict debugging with 20 plus pretested, verified compatible AI/ML frameworks on Amazon Linux 2023.
AWS Native Integrations with Bedrock, SageMaker, and OpenSearch enable you to build RAG chatbots, multimodal workflows, and scalable AI applications without manual SDK configuration. Includes a runnable RAG chatbot notebook using Bedrock , OpenSearch so you can validate your environment immediately after launch rather than spending time wiring services together.
Complete Developer Productivity Stack with Amazon NICE DCV encrypted remote desktop, JupyterLab, RStudio, Visual Studio Code, PyCharm, Docker, and Anaconda all pre-configured and verified compatible. Launch on a CPU instance to explore the environment and sample workflows, then scale to GPU instances (g5, p4d, p5) for production LLM fine tuning and training workloads.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
The software itself carries no license fee. You pay only for the single hourly compute option, priced per hour on a g4dn.4xlarge GPU-backed instance. This instance provides the GPU acceleration needed to run the preconfigured large language model stack. Your cost scales with the number of hours the instance runs, so you can start and stop it to match usage. There are no separate tiers or add-on dimensions to choose from. You pay standard AWS infrastructure charges for the instance in addition to this listing.
Top-of-mind questions for buyers
What hardware and software do I get with the g4dn.4xlarge hourly rate?
You get a GPU-backed instance running a preconfigured large language model stack. This includes a private model runtime, a web-based chat interface, API access, and GPU acceleration with the supporting driver stack. The stack also supports fine-tuning workspaces and secure model deployment inside your environment.
Am I charged when the instance is stopped or paused?
The hourly software charge applies only while the instance runs. Stopping the instance stops the software charge. You can start and stop it to match usage. Stopped instances may still incur underlying AWS storage fees for attached volumes, billed separately from this listing.
Does the hourly rate cover the AWS instance cost too?
No. The listing charge covers the software stack on the g4dn.4xlarge instance. You pay standard AWS infrastructure charges for running that GPU instance separately. Both appear on your AWS bill. Your total cost combines the software hourly rate and the underlying compute charge.
www.relevancelab.com
Helpful?
Vendor refund policy
NA
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
An AMI is a virtual image that provides the information required to launch an instance. Amazon EC2 (Elastic Compute Cloud) instances are virtual servers on which you can run your applications and workloads, offering varying combinations of CPU, memory, storage, and networking resources. You can launch as many instances from as many different AMIs as you need.
Version release notes
NA
Additional details
Usage instructions
"Quick Usage Summary
Subscribe to the AWS Marketplace product and launch an instance.
Use at least 200 GB for the EBS Volume and select g4dn.4xlarge for smooth performance.
AI and ML libraries are installed in the conda environment named ""genai"". To use them:
a. Open Anaconda Prompt.
b. Run: conda activate genai
c. Run: python --version
d. Import packages such as torch, transformers, langchain, faiss, scikit-learn, pyspark, dask, vowpalwabbit, monai, peft.
Connect via NICE DCV
Open a browser and navigate to https://<your-public-dns-or-IP>:8443
Log in using your Windows Administrator username and password.
You will gain access to the Windows desktop in your browser.
Note: Ensure that TCP port 8443 is allowed in the EC2 security group and Windows firewall.
Development IDEs
Launch Visual Studio Code, Visual Studio 2022, or PyCharm from the Start Menu.
Create new files or open existing projects.
Suitable for Python, R, .NET, and full-stack development.
Anaconda
Open Anaconda Navigator from the Start Menu to manage environments and packages.
Alternatively, use Anaconda Prompt to run commands such as: conda list
JupyterLab and Python
Access JupyterLab from the desktop or Start Menu.
Run Python or R notebooks and import preinstalled libraries.
Activate the genai environment for AI and ML workflows.
Machine Learning and AI Libraries
The environment includes preinstalled libraries for data science, machine learning, and generative AI. Examples:
Transformers version 4.56.2
LangChain version 0.3.27
LlamaIndex version 0.14.3
FAISS version 1.9.0
PyTorch version 2.5.1 with CUDA 12.1 (GPU enabled)
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Pre-configured data science workspace AMI with 15+ tools including RStudio, JupyterLab, and ML frameworks. Unlike DIY setups, deploy a validated environment in under 5 minutes.
Secure, pre-configured RStudio IDE on EC2 Linux. Researchers launch a production-ready R and Python environment with auto-renewing HTTPS in minutes - no server setup required.
Pre-configured RStudio Server on Rocky Linux 9 with Amazon DCV pixel streaming. Deploy in minutes for secure data analysis where no data leaves the instance.
CIS Level 2 hardened Windows Server 2022 AMI with GPU-ready data science tools. Deploy a secure workspace with JupyterLab, RStudio, and PyTorch via AWS RES.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.