This data set contains recordings of voice commands without a wake word in Italian (it_IT) of 135 participants of age 15-65 (e.g., "Ehi Alexa, raccontami una barzelletta.").
This data set contains recordings of voice commands without a wake word in Italian (it_IT) of 135 participants of age 15-65 (e.g., "Ehi Alexa, raccontami una barzelletta.").
The participants recorded their voices remotely using our in-house platform solution. Each participant has recorded on average 72 utterances (minimum 51, maximum 77).
This data set contains the voice command only. The data set of wake words is available in a separate data package. You can access a sample of the wake word and voice command data set here. Recordings are grouped by participant IDs and can be mapped to the phrase text. The data set comes with a complete phrase list and speaker metadata (see below).
Please contact us in case you have any questions about the data set.
Our Data - Ready for People
Each individual data set has undergone a thorough QA process. Prior to uploading the data sets, data integrity and completeness has been verified. It is our goal to provide high-quality, ready-to-use products to our clients.
Audio metadata:
Audio format: Wave
Encoding: pcm_s16le
Sampling rate: 44.1 kHz
Bit depth: 16 bits
Bit rate: 706 kb/s (constant)
Channels: 1
For each recording, the following speaker metadata is provided:
Childhood Location
Participant Gender
Record Location
Participant Age
Dialect
Nationality
Key Demographic Information
Age: 15-65 years, on average 34 years old, at the median 31 years old
Globalme is a language technologies company with headquarters in Vancouver, BC Canada. Our mission is simple: we ensure your products and services are Ready for People by helping train and test the latest AI technologies.
If you want to learn more about Globalme and our services, please visit our homepage.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on the duration and terms of your contract with the vendor. This entitles you to a specified quantity of use for the contract duration. If you choose not to renew or replace your contract before it ends, access to these entitlements will expire.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
This listing uses a contract pricing model with a single dimension: Product Access (Units). You buy units that grant access to the product. Pricing scales with the number of units you purchase, rather than through separate tiers or plan levels. There is one billing structure here, so you select the unit quantity that matches your needs. The product provides Italian adult voice command data to support AI training and speech applications.
Top-of-mind questions for buyers
What does one unit of Product Access represent for this dataset?
A unit is the billing measure that grants you access to the Italian adult voice command dataset. You buy the number of units that matches your project needs. The listing does not define units by recordings or speakers, so contact the vendor for the exact per-unit scope.
How does my cost change if I need more access to the dataset?
Cost scales with the number of units you purchase. There are no separate tiers or plan levels here. To increase access, you add more units. Each added unit adds to your total under the same single pricing structure.
What kind of data does this product provide?
This product provides Italian adult voice command data. It supports AI training and speech applications, such as voice recognition and voice-driven interfaces. The vendor sources multilingual speech data for AI systems.
globalme.net
Helpful?
Vendor refund policy
We do not offer any refund at this time. If you have questions regarding your subscription or the product, please contact us.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
This data set contains recordings of the wake word "Alexa" in Italian (it_IT) of 135 participants of age 15-65 used in voice commands (e.g., "Ehi Alexa, raccontami una barzelletta.").
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the nova-3 model which can each transcribe a set of languages. See version details for more information.
You will be billed $0.0077/min as described by https://deepgram.com/pricing. Private pricing available upon request.
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the aura-2 model which can each speak a set of languages and voices. See version details for more information. Deepgram charges are billed per request as described by https://deepgram.com/pricing
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the nova-3 model which can each transcribe a set of languages. See version details for more information.
You will be billed $0.0092/min as described by https://deepgram.com/pricing. Private pricing available upon request.
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the flux model which can each transcribe a set of languages. See version details for more information.
You will be billed $0.0078/min as described by https://deepgram.com/pricing.
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.
Deepgram is the enterprise Voice AI platform for building and scaling real time voice applications on AWS. This product listing contains multiple versions of the flux model which can each transcribe a set of languages. See version details for more information.
You will be billed $0.0077/min as described by https://deepgram.com/pricing. Private pricing available upon request.
Our APIs for Nova Speech to Text (STT) are natively available in the new SageMaker Bi-Directional Streaming API. Additional native touchpoints with Amazon Bedrock, Lex, and Amazon Connect make it simple to compose full voice experiences with the cloud services your teams already trust.