This product has charges associated with it for hardening, security configuration, and support.
LocalAI is an open-source, drop-in OpenAI-compatible API server delivered as a single statically linked Go binary - chat, embeddings, images, and audio served from your own VM with no GPU and no cloud egress. Unlike bare LocalAI AMIs that expose port 8080 unauthenticated to the public internet and ship without any TLS termination, this Lynxroute build is ready out of the box: per-instance LOCALAI_API_KEY at first boot, host Nginx with TLS in front of the API, the LocalAI process bound to loopback, UFW firewall pre-configured, and a CIS Level 1 hardened Ubuntu 24.04 LTS base.
MIT license - fully auditable, no vendor lock-in.
This is a repackaged software product wherein additional charges apply for hardening, security configuration, and support.
WHAT IS LOCALAI
LocalAI is an open-source, OpenAI-compatible inference server delivered as a single statically linked Go binary. It exposes the same REST surface as OpenAI - /v1/chat/completions, /v1/embeddings, /v1/audio/transcriptions, /v1/audio/speech, /v1/images/generations, /v1/models, plus a built-in /models gallery API - so any client written for OpenAI (openai-python, openai-node, LangChain OpenAI provider, LlamaIndex, IDE plugins) talks to it unchanged with one base URL swap. Inference runs entirely on CPU using llama.cpp and Whisper.cpp backends; GGUF model files load from local disk under /var/lib/localai/models and persist across restarts - no external database, no Redis, no cloud egress to OpenAI or Anthropic. Bring your own GGUF model (LLaMA family, Mistral, Qwen, Phi, Gemma) or install one in two clicks from the bundled gallery (mudler/LocalAI gallery index). MIT license, no vendor lock-in.
WHAT THIS AMI ADDS
Security hardening:
Per-instance LOCALAI_API_KEY (24-char random) generated at first boot, enforced by LocalAI for every /v1/, /models, /embeddings, /audio, /images call - never baked into the AMI
The same API key authenticates OpenAI-SDK REST clients (Authorization: Bearer) and the Web UI ("Login with API Token") - one credential, one rotation surface
LocalAI bound to 127.0.0.1:8080 - the API process never listens on a public interface
Host Nginx fronts port 443 with a self-signed certificate; HTTP redirects to HTTPS; security headers (X-Content-Type-Options, X-Frame-Options, Referrer-Policy) applied
UFW firewall pre-configured - only TCP 22, 80, 443 are exposed
fail2ban, AppArmor
CVE scan - every image is scanned for vulnerabilities before release
OS hardening (CIS Level 1):
CIS Ubuntu 24.04 LTS Level 1 benchmark applied via ansible-lockdown
CIS Conformance Report at /etc/lynxroute/cis-report.html
CIS Tailored Profile at /usr/share/doc/lynxroute/CIS_TAILORED_PROFILE.md
Highlights
LocalAI security baked in: per-instance LOCALAI_API_KEY at first boot, LocalAI bound to 127.0.0.1, host Nginx with TLS on :443, all auth enforced by LocalAI for /v1/, /models, /embeddings, /audio, /images - unlike bare LocalAI AMIs that expose port 8080 unauthenticated to the public internet, ship no TLS terminator, and leave the model upload and gallery endpoints anonymous.
CIS Level 1 hardened Ubuntu 24.04 LTS: auditd, fail2ban, AppArmor, SSH key-only, IMDSv2 enforced. CVE-scanned before every release. SBOM (CycloneDX) and CIS Conformance Report included.
Drop-in OpenAI replacement on your own VM: /v1/chat, /v1/embeddings, /v1/images, /v1/audio routed by a single Go binary; openai-python, LangChain, and LlamaIndex work unchanged. CPU-only inference - no GPU, no NVIDIA stack, no per-call billing. MIT license - no vendor lock-in, ever.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Try this product free for 5 days according to the free trial terms set by the vendor. Usage-based pricing is in effect for usage beyond the free trial terms. Your free trial gets automatically converted to a paid subscription when the trial ends, but may be canceled any time before that.
LocalAI - Hardened Self-Hosted OpenAI-Compatible API Server
You pay by the hour based on the EC2 instance size you launch. Five instance options are available. The t3.medium and t3.large run on burstable general-purpose hardware for lighter workloads. The m6i.large, m6i.xlarge, and m6i.2xlarge run on general-purpose hardware, scaling up in CPU and memory as you move to the next size. The hourly rate rises with instance capacity, so you match cost to the compute you need. All options ship as the same hardened, self-hosted API server image. AWS bills you hourly through the Marketplace.
Top-of-mind questions for buyers
Am I charged for the software when an instance is stopped or powered off?
Billing meters running instance-hours. A stopped instance stops accruing the hourly software charge. You may still pay underlying AWS fees, such as storage for the attached volume, while the instance is stopped. Only running hours count toward the software rate.
What do I get with each instance option, and how does capacity differ between them?
Each option maps to one running EC2 instance of the named size. The t3.medium and t3.large use burstable CPUs for lighter loads. The m6i.large, m6i.xlarge, and m6i.2xlarge add CPU and memory at each step. All run the same hardened API server image.
Is this pay-as-you-go, or do I commit to a term upfront?
This is usage-based billing with no upfront commitment. You pay only for the hours each instance runs. Start or stop instances any time, and charges follow actual running time. This suits variable workloads where you want cost tied directly to usage.
lynxroute.com
Helpful?
Vendor refund policy
We do not offer refunds for this product. AWS infrastructure charges (EC2, EBS, data transfer) are billed separately by AWS and are not refundable by us.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
An AMI is a virtual image that provides the information required to launch an instance. Amazon EC2 (Elastic Compute Cloud) instances are virtual servers on which you can run your applications and workloads, offering varying combinations of CPU, memory, storage, and networking resources. You can launch as many instances from as many different AMIs as you need.
Version release notes
LocalAI 4.7.1 on Ubuntu 24.04 LTS
LocalAI 4.7.1 on Ubuntu 24.04 LTS - minor update (from 4.6.2): voice-cloning profile endpoints, auto-resolve context size, GPU device selection, and inbound thinking/reasoning aliases - all additive, existing options unchanged. MIT unchanged
Certbot pre-installed - enable a trusted HTTPS certificate with one command: sudo certbot --nginx -d yourdomain.com
Rebuilt on the latest CIS Level 1 hardened Ubuntu 24.04 LTS base
Additional details
Usage instructions
Launch instance (t3.large recommended; t3.medium minimum for the smallest quantized models)
Open Security Group - allow TCP 443 from your IP only
Open https://<PUBLIC_IP>/ in your browser - accept the self-signed certificate warning
Click "Login with API Token", paste the API key from the credentials file
Web UI: Models tab -> pick a model from the gallery -> Install. Or drop a GGUF file into /var/lib/localai/models and run sudo systemctl restart local-ai
REST API (OpenAI-compatible) with curl:
curl -k https://<PUBLIC_IP>/v1/chat/completions
-H "Authorization: Bearer <api-key>"
-H "Content-Type: application/json"
-d '{"model":"<your-model>","messages":[{"role":"user","content":"hi"}]}'
The same API key authenticates Web UI sessions and OpenAI-SDK REST clients (Authorization: Bearer).
Credentials are saved to /root/localai-credentials.txt at first boot.
Models are NOT pre-loaded - install on demand from the gallery or upload your own GGUF files.
Replace the self-signed TLS certificate with a CA-signed certificate for production use:
sudo certbot --nginx -d YOUR_DOMAIN
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
This product includes additional charges for seller support. Root partition and filesystem automatically expands at boot for volumes larger than 8 GiB, for seamless scalability. Preconfigured with Cloud-init and ENA support for enhanced network performance. LocalAI.io provides a cutting-edge solution for running and deploying AI models directly on your local infrastructure. With support for advanced machine learning frameworks, seamless integrations, and optimized performance, LocalAI.io enables businesses to achieve low-latency, cost-effective AI capabilities while maintaining full data sovereignty and control. Ideal for privacy-sensitive applications and edge deployments, LocalAI.io empowers users to scale AI workflows without relying on external cloud providers.
Preconfigured LocalAI virtual machine for running AI models on your own infrastructure. Unified platform for AI Chat, image and voice generation, LLMs and embeddings with OpenAI API compatibility.
Available in both CPU and GPU configurations.
Deepdub GO is a cutting-edge virtual AI studio designed to streamline the post-production dubbing process. This platform empowers creators to produce high-quality localized content quickly and efficiently by leveraging proprietary emotion-based text-to-speech technologies and professional voice creation.
BLEND, tasq.ai company, is a localization and AI translation platform that enables organizations to manage, automate, and control multilingual content workflows across systems and content types.
The platform combines LLMs, quality estimation, AI-assisted review, and human expertise to support scalable, high-quality translation operations.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.