Receive recurring monthly AI credits with priority capacity and a predictable budget. Access frontier open-weight models, billed through AWS Marketplace.
Designed for organizations requiring predictable monthly AI spending without giving up flexibility. Includes reserved AI credits, centralized AWS Marketplace billing and simplified budget management, while allowing additional AI consumption as your needs grow.
Cas d'usages
Enterprise AI Assistants,
Internal Knowledge Search,
Customer Support Automation,
Intelligent Document Processing,
Production AI Applications
Highlights
Predictable monthly AI budget with reserved AI credits
OpenAI-compatible API for frontier open-weight AI models
Enterprise-ready with AWS Marketplace billing and Zero Data Retention
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on the duration and terms of your contract with the vendor. This entitles you to a specified quantity of use for the contract duration. If you choose not to renew or replace your contract before it ends, access to these entitlements will expire.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
This listing bills through a single dimension: Inference Credit units. You reserve a set number of monthly AI credits, where each credit equals one dollar of platform usage. Your requests draw down these credits as you run coding models, with each call billing based on the tokens it consumes. Because pricing works per credit, your cost scales directly with how much you use. You buy credits in the quantity you need and apply them toward inference across the available models.
Top-of-mind questions for buyers
What does one Inference Credit map to when I run coding models?
One credit equals one dollar of platform usage. Each request bills against your credits based on the tokens it consumes, priced per model. Input tokens, output tokens, and cache reads each carry their own per-million-token rate, so a credit's runtime depends on which model you run and how heavy the request is.
What drives how fast I use up my credits?
Token volume drives cost. Each request bills input, output, and cache-read tokens at that model's published per-million-token rate. Heavier or long-context tasks consume more credits. Lower-priced models stretch credits further than premium ones. Reasoning effort also matters, since thinking tokens count toward the output you are billed for.
What happens if I run out of reserved credits during the month?
Your reserved credits draw down as requests consume tokens. When the balance runs low, you top up to keep running. The service issues plain, retryable rate-limit responses rather than hard failures, and burst limits grow with your lifetime top-ups. Bonus credit counts toward your balance but not toward those tier limits.
app.umans.ai+1
Helpful?
Vendor refund policy
Wallet top-ups are final and non-refundable, except for billing errors, duplicate or unauthorized charges, or top-ups not credited to the wallet. Refund requests must be submitted to contact@umans.ai within 30 days of purchase. EU/EEA consumers may withdraw within 14 days and receive a refund for unused credit, subject to applicable consumer protection laws.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Purchase prepaid inference credits and use them across Umans AI portfolio of frontier open-weight models. Top up your organization wallet when needed, with no recurring commitment.
One annual commitment covering all GitLab capabilities on AWS. Reshape your mix of seats and AI usage month to month with no contract amendments, no re-procurement.
Enable your software teams to orchestrate their AI agents across the entire SDLC with GitLab Credits, the consumption currency for GitLab Duo Agent Platform.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.