AudioShake's best-in-class AI separates music, speech, and other audio into its component parts (or "stems") so that it can speed up workflows and open up new opportunities for audio in dubbing, sync licensing, mixing, sonic analysis, vocal synthesis, VR/AR, gaming, and more.
Winner of Sony's Demixing Challenge, AudioShake's patented AI is used by major labels and publishers, film/TV production studios, fitness, gaming, generative AI, and dubbing companies in order to make audio more accessible, interactive, and editable. AudioShake's best-in-class AI separates music, speech, and other audio into its component parts (or "stems") so that it can speed up workflows and open up new opportunities for audio in dubbing, sync licensing, mixing, sonic analysis, vocal synthesis, VR/AR, gaming, and more.
Highlights
Dubbing, Localization, Transcription, Music Removal, Music Stems
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on the duration and terms of your contract with the vendor, and additional usage. You pay upfront or in installments according to your contract terms with the vendor. This entitles you to a specified quantity of use for the contract duration. Usage-based pricing is in effect for overages or additional usage not covered in the contract. These charges are applied on top of the contract price. If you choose not to renew or replace your contract before the contract end date, access to your entitlements will expire.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
You pay per processed minute of audio, so cost scales with how much content you run through each model. The dimensions are independent options, not tiers—you pick the separation task you need. Speech Cleanup and Alignment share the lowest per-minute rate. Dialogue Isolation and Lyric Transcription sit at a mid rate. Film/TV Music & FX, Music Removal, and All Music Instrument Stems share the highest per-minute rate. A separate API Access dimension covers one API token; it includes no usage or calls, so processing minutes are billed under the model dimensions above.
Top-of-mind questions for buyers
What does one "processed minute" mean for billing?
You are charged per minute of audio or video you run through a model. Cost is based on the length of the file processed, not the number of stems you export. Each model dimension meters the same way—minutes of content sent in for separation, transcription, or alignment.
Does the API Access dimension include any processing minutes?
No. The API Access dimension covers one API token only, with no usage or calls included. You use the token to connect your workflow, then pay separately for each model you run under its per-processed-minute rate. The token itself carries no processing allowance.
If I run one file through several models, how does the bill add up?
Each model bills independently at its own per-minute rate. Running the same file through two models charges you twice—once per model—for the same minutes. The dimensions do not bundle; your invoice sums the minutes processed under each model you select.
www.audioshake.ai+1
Helpful?
Vendor refund policy
none
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
Email support is offered Monday - Friday from 9am - 6pm pacific time support@audioshake.ai
AWS infrastructure support
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Utilizes patented artificial intelligence algorithm that won Sony's Demixing Challenge for audio decomposition
Multi-Format Stem Extraction
Generates isolated stems for dubbing, localization, music removal, music stem extraction, and transcription applications
Audio Processing Capabilities
Enables sonic analysis, vocal synthesis, and interactive audio editing for various audio content types
Workflow Integration
Supports integration with dubbing, sync licensing, mixing, VR/AR, gaming, and generative AI applications
AI-Powered Metadata Tagging
Automated asset tagging through AI-driven workflows that accelerate metadata generation and content organization.
Multi-Model AI Ecosystem Integration
Access to ecosystem of over 300 AI models across various categories for content discovery and analysis automation.
Content Monetization Platform
Built-in ecommerce capabilities enabling creation of branded content marketplaces and paid access models for event-specific or general content.
Rights Management and Access Control
Granular access controls and rights management features for managing content permissions across internal and external stakeholders.
System Integration and Interoperability
Flexible integration capabilities with existing DAM solutions, disparate systems, and custom AI models through open architecture.
Video Generation Capabilities
Supports Text-to-Video, Image-to-Video, Video Omni, Motion Control, Avatar, and Effects with native 4K output and intelligent storyboarding
Image Generation with Ultra-HD Output
High-fidelity image generation with native 2K/4K ultra-HD output, batch generation capabilities, and precise control over composition, lighting, and depth-of-field
Multimodal Video Generation
Native multimodal video generation with native audio integration and advanced consistency control
Motion Capture Technology
High-fidelity motion capture for precise and expressive action generation with advanced motion control
Audio Integration
Native audio generation and integration capabilities within video generation workflows
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.