This data set contains recordings of voice commands without a wake word in Italian (it_IT) of 65 participants of age 6-14 (e.g., "Ehi Alexa, raccontami una barzelletta.").
This data set contains recordings of voice commands without a wake word in Italian (it_IT) of 65 participants of age 6-14 (e.g., "Ehi Alexa, raccontami una barzelletta.").
The participants recorded their voices remotely using our in-house platform solution. Each participant has recorded on average 66 utterances (minimum 50, maximum 74).
This data set contains the voice command only. The data set of wake words is available in a separate data package. You can access a sample of the wake word and voice command data set here. Recordings are grouped by participant IDs and can be mapped to the phrase text. The data set comes with a complete phrase list and speaker metadata (see below).
Please contact us in case you have any questions about the data set.
Our Data - Ready for People
Each individual data set has undergone a thorough QA process. Prior to uploading the data sets, data integrity and completeness has been verified. It is our goal to provide high-quality, ready-to-use products to our clients.
Audio metadata:
Audio format: Wave
Encoding: pcm_s16le
Sampling rate: 44.1 kHz
Bit depth: 16 bits
Bit rate: 706 kb/s (constant)
Channels: 1
For each recording, the following speaker metadata is provided:
Childhood Location
Participant Gender
Record Location
Participant Age
Dialect
Nationality
Key Demographic Information
Age: 6-14 years, on average 10 years old, at the median 10 years old
Globalme is a language technologies company with headquarters in Vancouver, BC Canada. Our mission is simple: we ensure your products and services are Ready for People by helping train and test the latest AI technologies.
If you want to learn more about Globalme and our services, please visit our homepage.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on the duration and terms of your contract with the vendor. This entitles you to a specified quantity of use for the contract duration. If you choose not to renew or replace your contract before it ends, access to these entitlements will expire.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
This listing uses a contract pricing model with one dimension: Product Access, billed in units. You buy units to grant access to the product, which provides Italian-language voice command data for youth speakers. There are no tiers or instance sizes to choose from. Pricing scales with the number of units you commit to under the contract term.
Top-of-mind questions for buyers
What counts as one unit under the Product Access dimension?
One unit represents a measure of access to the dataset that this listing grants. The listing provides Italian-language voice command data from youth speakers. You buy units to unlock that access. The marketplace table does not further break down what one unit maps to, so confirm the exact per-unit contents with the vendor.
How does my cost change if I need more access to the dataset?
Cost scales with the number of units you commit to under the contract. Adding units raises the total; the per-unit basis stays the same. There are no tiers or instance sizes that change the rate. To adjust unit quantity mid-term, contact the vendor.
What kind of data does this product provide?
You get human-sourced voice command data recorded in Italian by youth speakers. This type of speech dataset is used to train and improve voice and speech AI systems. The seller focuses on ethically sourced, multilingual training data across speech, text, image, and video.
globalme.net
Helpful?
Vendor refund policy
We do not offer any refund at this time. If you have questions regarding your subscription or the product, please contact us.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
This data set contains recordings of the wake word "Alexa" in Italian (it_IT) of 65 participants of age 6-14 used in voice commands (e.g., "Ehi Alexa, raccontami una barzelletta.").
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the nova-3 model which can each transcribe a set of languages. See version details for more information.
You will be billed $0.0077/min as described by https://deepgram.com/pricing. Private pricing available upon request.
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the aura-2 model which can each speak a set of languages and voices. See version details for more information. Deepgram charges are billed per request as described by https://deepgram.com/pricing
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the nova-3 model which can each transcribe a set of languages. See version details for more information.
You will be billed $0.0092/min as described by https://deepgram.com/pricing. Private pricing available upon request.
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the flux model which can each transcribe a set of languages. See version details for more information.
You will be billed $0.0078/min as described by https://deepgram.com/pricing.
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the flux model which can each transcribe a set of languages. See version details for more information.
You will be billed $0.0077/min as described by https://deepgram.com/pricing. Private pricing available upon request.
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.