Overview
This dataset contains 8,507,968 full-text U.S. court opinion records organized into 38 decade-based CSV files, from the 1650s to 26Q1. Each file contains records whose filing date falls within that decade, making it easy to work with a specific era, sample across time periods, or process the corpus incrementally — load the 1990s for a focused study, or stream all 38 files for the complete 86-billion-character collection. Opinion texts average roughly 10,000 characters (about 1,600 words) per record.
Each record includes:
case_name — the name of the case
cleaned_text — full narrative opinion text, cleaned and whitespace-normalized
citations — the case's citations as reported
date_filed — filing date
citations_norm — normalized citation string for matching and deduplication
Files are UTF-8, RFC-4180 CSVs with header rows (e.g. cases_1990s_no_empty_cases.csv), exported directly from PostgreSQL for reliable import into any database or data pipeline. Every record contains non-empty opinion text — empty and whitespace-only entries are excluded. Text fields preserve original paragraph structure, and the identical schema across all 38 files means one import routine handles the entire collection.
Highlights
- 8.5 million full-text U.S. court opinion records split into 38 decade files spanning the 1650s to 26Q1 — study a single era or process the full 86-billion-character corpus file by file.
- Identical schema in every file: case name, filing date, citations, normalized citation string, and complete cleaned opinion text. Empty and whitespace-only records excluded.
- UTF-8, RFC-4180 CSVs with header rows, exported directly from PostgreSQL — decade-sized files are far easier to load, sample, and parallelize than a single 81 GB download.
Details
Introducing multi-product solutions
You can now purchase comprehensive solutions tailored to use cases and industries.
Features and programs
Financing for AWS Marketplace purchases
Pricing
Dimension | Description | Cost/month |
|---|---|---|
Product Access | Dimension that grants access to the product for subscribers. | $250.00 |
Vendor refund policy
Because this product is downloadable data that cannot be returned once accessed, approved refunds are limited to 50% of the purchase price. Requests must be sent to support@maconapps.com within 14 days of purchase and include your AWS account ID and reason; refunds are not available after 14 days. If the dataset is materially defective (corrupted, inaccessible, or materially different from the listing), we will remedy it or issue a full refund.
How can we make this page better?
Legal
Vendor terms and conditions
Content disclaimer
Delivery details
AWS Data Exchange (ADX)
AWS Data Exchange is a service that helps AWS easily share and manage data entitlements from other organizations at scale.
Additional details
You will receive access to the following data sets.
Data set name | Type | Historical revisions | Future revisions | Sensitive information | Data dictionaries | Data samples |
|---|---|---|---|---|---|---|
U.S. Court Opinions by Decade — 8.5M Full-Text Records in 38 Files (1658–Present) | All historical revisions | All future revisions | Not included | Not included |
Similar products

![1950 Census Population Schedules, Enumeration District Maps, and En[...]](https://d1ewbp317vsrbd.cloudfront.net/6ae40ca4-eac2-41b1-a4df-7e6bb2cbf1a3.png)
