Umans AI is a managed inference platform serving frontier open-weight AI models, with self-hosted model weights and a Zero Data Retention commitment: customer prompts and responses are never stored.
The Pay-as-you-go offer is the most flexible way to access the platform: no upfront commitment and no minimum spend. You only pay for actual token consumption, with 1 credit = $1 of platform usage, billed centrally through your AWS account.
Cas d'usages
AI Coding Assistants,
AI Agents for continuous complex tasks,
Development & Testing,
RAG Applications,
API-based AI Integration
Highlights
OpenAI compatible API : connect coding agents and tools such as Claude Code, Cursor, Zed and OpenCode in minutes, with no GPU infrastructure to manage and no vendor lock in
Service integrity: the model you select is the model you get, no silent model substitution, published quantization
Zero Data Retention: prompts and completions are not retained
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
This listing uses a single usage-based dimension: the AI Inference Credit. One credit equals one dollar of platform usage, and you pay only for what you consume. There are no tiers or fixed quantities to select. You prepay credits into a wallet, then each API request draws down that balance at published per-token rates. Costs scale with actual usage, so heavier coding workloads consume more credits. Your credit balance acts as your spending limit, giving you direct control over how much you use over time.
Top-of-mind questions for buyers
What does one AI Inference Credit actually pay for?
One credit equals one dollar of platform usage. You spend credits on model requests, billed at published per-token rates for the model you call. Server-side web searches also draw from credits and appear in your usage breakdown next to model requests, showing the search backend used.
Does the model I choose change how fast I consume credits?
Yes. Each model has its own per-token price, so heavier models draw credits faster. The premium repository-scale model costs more per token than the lighter, high-interactivity model. Reasoning tokens count toward output too, so deeper thinking consumes more credits per request.
What happens when my credit balance runs low or I send too many requests at once?
Your prepaid balance is your real spending limit. Wallet keys are never auto-paused. If you exceed your tier's concurrency or request window, you get retryable rate-limit errors that coding agents handle by backing off. Your burst limits grow with lifetime top-ups.
app.umans.ai
Helpful?
Vendor refund policy
Usage charges are non-refundable, except for billing errors or unauthorized charges. Refund requests must be submitted to contact@umans.ai within 30 days.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Q-LOG is an innovative awareness platform developed by Celestya Ltd to elevate employee awareness and behavior in the context of cybersecurity and information security.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.