Cut your AWS Athena and Redshift Spectrum scan bill by re-optimizing your S3 data-lake Parquet to the layout your queries actually read. Billed purely per GB of data optimized, with a safe shadow-then-swap that never changes query results.
S4 Scan learns from your Athena query history which columns are read and which predicates filter your tables, then rewrites the underlying S3 Parquet into the physical layout those queries want: partition and row-group pruning, dictionary encoding, zstd, and small-file compaction. Athena and Redshift Spectrum bill per terabyte scanned from S3, so a smaller, better-pruned layout is a direct, recurring bill reduction.
This listing is pay-per-GB: you are billed only for the volume of data S4 Scan optimizes, with no hourly software fee. Safety is the core of the product. S4 Scan never overwrites your source data in place: it writes optimized data to a shadow location, verifies that every query returns byte-for-byte identical results (values, nulls, decimal scale, timestamp time zone), and only then swaps the AWS Glue table pointer. One-command rollback restores the original layout. There is no lock-in: output is standard, Athena-readable Parquet. A dry-run projects the dollar savings on your real tables before you commit. S4 Scan re-optimizes on an EventBridge schedule with no human in the loop.
Highlights
Pay only for value delivered: billed per GB of data optimized, with no hourly software fee.
Safe by construction: optimized data is written to a shadow location, verified query-result-identical (values, nulls, decimal scale, timestamp tz), then swapped via the Glue catalog, with one-command rollback. Source data is never modified in place.
No lock-in, no babysitting: output is standard Athena-readable Parquet, and the AMI re-optimizes on an EventBridge schedule so savings compound automatically.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Try this product free for 14 days according to the free trial terms set by the vendor. Usage-based pricing is in effect for usage beyond the free trial terms. Your free trial gets automatically converted to a paid subscription when the trial ends, but may be canceled any time before that.
Pricing is based on actual usage, with charges varying according to how much you consume. Subscriptions have no end date and may be canceled any time. Alternatively, you can pay upfront for a contract, which typically covers your anticipated usage for the contract duration. Any usage beyond contract will incur additional usage-based costs.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
This listing bills on a single usage dimension: you pay per GB of data optimized, per hour. There are no tiers or instance-size choices to select — your cost scales directly with how much data the tool processes. As it re-optimizes your data-lake layout to cut the bytes queries scan, your charge tracks the volume it works on. Because cost rises with data processed, this pay-per-GB model fits workloads where scan volume is meaningful. You add your own EC2 compute cost separately.
Top-of-mind questions for buyers
What does one GB of "data optimized" mean for billing?
You pay per GB of data the tool re-optimizes, metered per hour. The product rewrites your S3 Parquet into a layout your queries prune better. Billing tracks the volume of data it processes, not the number of queries you run or the bytes those queries scan.
Which parts of my bill does this reduce, and what stays the same?
It cuts the bytes Athena and Redshift Spectrum scan from S3, since both bill per terabyte read. Your per-GB optimization charge and separate EC2 compute cost still apply. At low scan volumes the running cost can exceed the scan savings, so it fits larger scan workloads.
Does the per-GB charge cover the compute it runs on?
No. The per-GB-optimized fee is separate from the EC2 instance cost. The tool runs as an AMI inside your own AWS account and VPC on a t3 or m5 class instance. You pay that EC2 charge on your regular AWS invoice alongside the software fee.
abyo.net
Helpful?
Vendor refund policy
Email support@abyo.net within 30 days of charge for refund requests; refunds are evaluated case by case.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
An AMI is a virtual image that provides the information required to launch an instance. Amazon EC2 (Elastic Compute Cloud) instances are virtual servers on which you can run your applications and workloads, offering varying combinations of CPU, memory, storage, and networking resources. You can launch as many instances from as many different AMIs as you need.
Version release notes
Adds a CloudFormation Quick Launch delivery option. Software identical to version 1.1.2.
Additional details
Usage instructions
Deploy via deploy/cloudformation/s4scan-quickstart.yaml with EnableMetering=true. Billed per GB of data optimized, reported hourly to the AWS Marketplace Metering Service under the gb_optimized dimension.
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Self-contained EC2 AMI of the S4 transparent S3 compression gateway with CPU codecs (zstd / gzip) preinstalled. Launch on any general-purpose or compute-optimized instance (t3 / m6i / m7i / c6i / c7i), point your S3 clients at it, and cut S3 storage bytes 50-80 percent for compressible data with zero application changes.
Self-contained EC2 AMI of the S4 transparent S3 compression gateway with NVIDIA nvCOMP GPU codecs preinstalled. Launch on a GPU instance (g4dn / g5 / g6), point your S3 clients at it, and cut S3 storage bytes 50-80 percent for compressible data with zero application changes.
Drop-in S3-compatible gateway that transparently compresses every object (CPU zstd or GPU nvCOMP), cutting S3 storage bytes 50-80 percent for compressible data with zero application changes. Includes pre-deployment savings estimation and measured-savings reporting.
Drop-in S3-compatible gateway that transparently compresses every object, cutting S3 storage bytes 50-80 percent for compressible data with zero application changes. This edition bills by measured savings: you pay per GB of backend storage avoided, per hour, at roughly one third of the avoided storage cost.
Cut your CloudWatch custom-metric bill: govern metric cardinality at ingest, then auto-baseline and roll up savings across your whole AWS Organization.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.