The Yellowbrick Data Warehouse is a cloud native, scale-out SQL database designed for the most demanding batch, ad hoc, real-time and mixed workloads. Fully elastic clusters with separate storage and compute run complex queries at any scale with sub-second response times.
Yellowbrick innovates in three key areas:
A radically redesigned database engine, using Direct Data Accelerator technology that results in superior efficiency and significantly higher price performance.
Runs in your VPC, but delivered as a SaaS experience, eliminating security concerns and giving you full control.
Providing deployment flexibility by offering an identical data warehouse for on-premises use cases.
Yellowbrick is designed to scale cost-effectively to meet the growing needs of your business from small deployments of a few 100GB, all the way up to running multi-petabyte enterprise-grade data warehouses. It's a database that you can trust to store and be the system of record for your most valuable enterprise data, or to underpin innovative data applications for thousands of users. Our customers trust Yellowbrick systems to generate business-critical, auditable financial reports that your business depends on. It requires almost no management, tuning, diagnosing or handholding and is familiar to modern developers accustomed to working with PostgreSQL.
Please reach out to aws.marketplace@yellowbrick.com for Purchasing Inquiries and Private Offers. Speaking to a Yellowbrick representative before purchase is required. This will ensure the best experience, including the pricing for Fixed Capacity, EPOD and other options. Pricing provided applies to US-based deployments only, international pricing provided in custom quote. Please email aws.marketplace@yellowbrick.com to initiate a private offer.
Highlights
Hyper-focussed on efficiency, using our patented Direct Data Accelerator technology, means Yellowbrick can guarantee to be 75% cheaper than your current cloud data warehouse solution (see website for terms and conditions).
Deliver complex analytics for thousands of enterprise users with fast query response, using your existing BI and integration tools, or underpin SaaS applications with thousands of users meeting guaranteed performance SLAs.
The only modern, cloud-native data warehouse that runs truly hybrid, in your VPC or data center, with built-in replication for Disaster Recovery or cross-geo data sharing.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on the duration and terms of your contract with the vendor. This entitles you to a specified quantity of use for the contract duration. If you choose not to renew or replace your contract before it ends, access to these entitlements will expire.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
You buy this data warehouse as a contract, priced per vCPU per hour on a 24x7 basis. The three options are sizing tiers that scale with your data volume and compute needs. Entry Level covers about 1TB with 16 vCPUs. Departmental covers about 10TB with 96 vCPUs. Enterprise covers 100TB or more with 384 vCPUs. As you move up the tiers, both the data capacity and the vCPU count increase together. Each price includes estimated AWS infrastructure cost. You pay for the software based on the vCPU compute you run.
Top-of-mind questions for buyers
What counts as a vCPU for billing on these tiers?
A vCPU is a unit of compute your data warehouse runs on. Each tier includes a set count: 16 for Entry Level, 96 for Departmental, and 384 for Enterprise. Your software cost is driven solely by how many vCPUs run. Changing node size can change the vCPU count and your price.
Am I charged for vCPUs I do not use, and can unused capacity roll forward?
You pay a flat rate for use up to your chosen maximum vCPU count for the period. If you do not run the software 24x7, that is fine, but there are no unused vCPU hours to roll forward. You pay for capacity availability, not actual runtime, under this contract.
Does the price include the AWS infrastructure cost, or do I pay AWS separately?
Each tier price includes an estimated AWS infrastructure cost bundled with the software. Yellowbrick's own model normally has you pay cloud providers directly for infrastructure and Yellowbrick for software. On this listing, the two are combined into one per vCPU per hour figure. 24x7 global support is included at no extra charge.
yellowbrick.com
Helpful?
Vendor refund policy
Subject to MSA and contract terms.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
https://support.yellowbrick.io/ Customers can create a service ticket at any time for any issue, incident, or question. Yellowbrick Data's Customer Success team is available 24x7. Customers can open a service ticket via email, phone or secure web portal (support.yellowbrick.io). Customers have full access to the Yellowbrick Customer Support Center, which includes a knowledge base, software downloads, and a comprehensive documentation library. Phone: 877-4YB-DATA (877-492-3282)
AWS infrastructure support
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Cloud-native, scale-out SQL database with Direct Data Accelerator technology for optimized query processing and efficiency
Elastic Cluster Infrastructure
Fully elastic clusters with separate storage and compute architecture enabling complex query execution at any scale with sub-second response times
Deployment Flexibility
Hybrid deployment capability supporting public cloud VPC, private cloud, on-premises, and edge deployments with identical data warehouse functionality across environments
PostgreSQL Compatibility
SQL database compatible with PostgreSQL, enabling developers to use familiar tools and interfaces for application development and integration
Built-in Disaster Recovery
Integrated replication functionality for disaster recovery and cross-geographic data sharing across hybrid deployments
Query Engine Architecture
Vectorized query engine with optimizer designed to handle complex queries, joins, and machine-generated queries efficiently
SQL Compatibility
Postgres-compatible SQL support enabling standard SQL operations across analytical workloads
Data Format Support
Apache Iceberg table format support for efficient data storage and management
Infrastructure Architecture
Decoupled metadata, storage, and compute architecture enabling isolated testing environments and independent scaling
Performance Optimization
Tunable data layout, caching mechanisms, and query plan control for production workload optimization
Open Table Format Support
Supports open table formats including Apache Iceberg and Parquet for consistent analytics and governance across hybrid-cloud environments without requiring ETL processes.
Multi-Engine Query Optimization
Provides multi-engine optimization across Presto SQL and Apache Spark to execute queries across structured, semi-structured, and unstructured data with workload-specific optimization.
Vector Database Integration
Integrates vector search and multi-model capabilities through Cassandra (Astra DB) and Milvus to support advanced retrieval-augmented generation (RAG), similarity search, and real-time operational workloads.
Enterprise Security and Compliance
Offers VPC-based deployments, AWS PrivateLink connectivity, and compliance support for FedRAMP (Medium) and HIPAA for AWS GovCloud environments.
Unified Access Control and Governance
Implements integrated data fabric with governance, lineage, and data quality capabilities, including AWS Lake Formation integration and Common Policy Gateway (CPG) for unified access control with real-time policy synchronization and audit tracking.
Ease of use when converting from Netezza and the performance and uptime of their appliances are incredible
What do you dislike about the product?
Cloud installation is not as straight forward and IP usage is high
What problems is the product solving and how is that benefiting you?
When Netezza went end of life, we needed a migration path and their appliance made that very easy and cost effective!
Consulting
An outstanding enterprise-class data platform
Reviewed on Jun 21, 2024
Review provided by G2
What do you like best about the product?
Yellowbrick is an outstanding high performance data platform that is straightforward to implement and maintain. It has proven itself to be an enterprise-class analytics platform with rich workload management and security features capable of supporting hundreds of users across data science, business intelligence, data engineering, and AI.
Our implementation of Yellowbrick involved a large scale migration of existing data assets and analytics products - an effort that was well supported through methodology backed by automated tooling. Interoperability with common data warehouse technologies further allowed us to leverage existing investments in key components such as data acquisition pipelines and end user analytics while upgrading data warehouse capabilities.
What do you dislike about the product?
Support for geospatial data is not yet available in the version implemented, although we look forward to adopting this in the near future as the roadmap is realised.
What problems is the product solving and how is that benefiting you?
Yellowbrick Data Warehouse is a modern, high-performance solution designed for fast and cost-effective query execution on large datasets. It efficiently handles high query concurrency, making it an ideal central data platform for numerous users with diverse workload requirements.
Oswaldo V.
Excelent option to have a DWH engine with low cost and efficency
Reviewed on Jun 14, 2024
Review provided by G2
What do you like best about the product?
Yellow Brick Appliance, is a very efficient database that can help companies to store a big amount of that and retrieve in fastest ways. Different use cases can be implemented.
What do you dislike about the product?
In some cases the compatibility with ETL tools are not ready, always you can use YB tools, however is not something easy according the data strategy.
What problems is the product solving and how is that benefiting you?
Offer data to business users in a quick and easy way
Financial Services
Simple, Speed and Powerful processing
Reviewed on May 30, 2024
Review provided by G2
What do you like best about the product?
The cluster implementation process is simple and quicker, per Postgresql engine, SQL is easy to use while supports much complex queries and quick to integrate with variour client softwares. The speed of processing, export & importing data is phenominal.
What do you dislike about the product?
The product do not support user defined functions.
What problems is the product solving and how is that benefiting you?
Enterprise Data Warehouse, Data analytics on large data-sets, faster unloads and ETL processings.
Victor R.
Simplicity, Speed and Processing Power
Reviewed on Feb 17, 2022
Review provided by G2
What do you like best about the product?
everything is documented, and the way of working is very simple since its base is Postgrestq there is a lot of documentation
What do you dislike about the product?
the updates so frequently, it seems as if they find bugs very soon that arise
What problems is the product solving and how is that benefiting you?
Speed in processing large data (groupings, sums, etc), Incredibly fast uploads as well as downloads
Recommendations to others considering the product:
It is an important evolution of traditional DWH, easy to use and simple to implement