The Brain Language Metrics on Earnings Calls Transcripts (BLMECT) dataset has the objective of monitoring several language metrics for the quarterly earnings call transcripts of 4500+ US stocks.
With this dataset we aim at providing additional building blocks to asset managers to build investment strategies based on alternative data.
Brain Language Metrics on Earnings Calls Transcripts
Overview
The exploitation of textual unstructured content (news, company filings, earnings calls etc) in financial analysis is quickly expanding across both quantitative and discretionary strategies as reflected in the growing number of academic papers and products in this domain.
The Brain Language Metrics on Earnings Calls Transcripts (BLMECT) dataset has the objective of monitoring several language metrics for the quarterly earnings call transcripts of 4500+ US stocks.
The dataset is composed of two parts. Part one includes several language metrics for the most recent earnings call transcript for each stock, namely:
Financial sentiment
Percentage of words belonging to financial domain classified by language types: constraining, litigious, uncertainty and interesting language. Readability score
Lexical metrics such as lexical density and richness of text
Text statistics such as the transcript length
Part two includes the differences between the most recent earnings call transcript and the previous one:
Difference of the various language metrics (e.g. delta sentiment, delta readability score, delta percentage of a specific language type etc.)
Similarity metrics between documents, also with respect to a specific language type (for example similarity with respect to “litigious” language or “uncertainty ” language)
The metrics calculation is reported separately for the following sections of the transcript:
Management Discussion
Analysts’ Questions
Management Answers to Analysts’ Questions
Historical Trial
The dataset contains historical data from January 2012 to July 2021 that can be freely accessed for 2 months. For a live feed please contact us at support@braincompany.co and we will make accessible a customized version of the product on AWS Data Exchange according to Client requirements.
Feed Details
The dataset is updated with a daily frequency since new earnings calls transcripts are published every day for some of the universe stocks. Clearly the data for each stock will change on a quarterly basis when new earnings calls are published. The historical dataset is available from year 2012.
The content of this dataset is not to be intended as investment advice. The material is provided for informational purposes only and does not constitute an offer to sell, a solicitation to buy, or a recommendation or endorsement for any security or strategy, nor does it constitute an offer to provvaluee investment advisory or other services by Brain. Brain makes no guarantees regarding the accuracy and completeness of the information expressed in the dataset.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
This is a free listing with a single pricing dimension: Product Access (Units). It grants subscribers access to the dataset at no charge. There are no tiers, usage add-ons, or size-based options to compare. You subscribe once to unlock access. As a History Trial, it lets you evaluate the earnings call language metrics dataset, which is delivered daily as CSV files in an S3 bucket. Because pricing is free with one access dimension, there is no cost that scales with usage, volume, or term.
Top-of-mind questions for buyers
What does the Product Access unit grant me for this History Trial?
One Product Access unit unlocks the earnings call language metrics dataset for you. The data covers quarterly earnings call transcripts for roughly 7000 global stocks. Historical data reaches back to 2012. You receive it as daily CSV files published in an S3 bucket.
Does my cost change as more transcripts or data are added?
No. This is a free listing with one flat access dimension. Cost does not scale with the number of transcripts, stocks, or data volume. The dataset updates daily as new earnings calls publish, and each stock refreshes quarterly, at no added charge to you.
What language metrics does the trial dataset actually deliver?
You get financial sentiment scores, readability scores, and lexical metrics like density and richness. It also includes text statistics, document similarity metrics, and differences between documents. Metrics report separately for management discussion, analyst questions, and management answers.
braincompany.co
Helpful?
Vendor refund policy
No refunds are offered for this product, for more information please contact us at support@braincompany.co
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
We are going to use an extra brain to help combat clinical error! MedicineOne has created Clinical Brain, a clinical intelligence platform to support medical decisions and help reduce clinical error.
BonData is the Organizational Context Intelligence Layer for the modern enterprise. In 24 hours, with zero human in the loop, our signature engine bonds every system, database, and document into one living Context Layer that humans and agents operate from on day one.