OLAC Record: Mixer 3 Speech

OLAC Record
oai:www.ldc.upenn.edu:LDC2023S02

Metadata

Title: Mixer 3 Speech

Access Rights: Licensing Instructions for Subscription & Standard Members, and Non-Members: http://www.ldc.upenn.edu/language-resources/data/obtaining

Bibliographic Citation: Huang, Shudong, Kevin Walker, and David Graff. Mixer 3 Speech LDC2023S02. Web Download. Philadelphia: Linguistic Data Consortium, 2023

Contributor: Huang, Shudong

Walker, Kevin

Graff, David

Date (W3CDTF): 2023

Date Issued (W3CDTF): 2023-03-15

Description: *Introduction* Mixer 3 Speech was developed by the Linguistic Data Consortium (LDC) and comprises 3,200 hours of audio recordings of conversational telephone speech involving 3,875 speakers and 26 distinct languages. This material was collected by LDC from 2005-2007 as part of the Mixer project, and recordings in this corpus were used in NIST Speaker Recognition Evaluation (SRE) and NIST Language Recognition Evaluation (LRE) corpora, including 2006 SRE and 2007 LRE. Researchers interested in applying those benchmark test sets should consult the respective NIST Evaluation Plans for guidelines on allowable training data for those tests. Data from 2006 SRE and 2007 LRE are available in the LDC Catalog: 2006 NIST Speaker Recognition Evaluation Training Set (LDC2011S09), 2006 NIST Speaker Recognition Evaluation Test Set Part 1 (LDC2011S10), 2006 NIST Speaker Recognition Evaluation Test Set Part 2 (LDC2012S01), 2007 NIST Language Recognition Evaluation Supplemental Training Set (LDC2009S05) and 2007 NIST Language Recognition Evaluation Test Set (LDC2009S04). *Data* The audio recordings were generated using LDC's computer telephony system capable of collecting speech from the telephone network. Recruited speakers were connected through a robot operator to carry on casual conversations lasting up to 10 minutes. Subjects fluent in languages other than English were asked to complete at least one non-English call. The documentation for this release contains information about the number of calls per subject and the number of calls per language. It also includes certain speaker demographic information, such as date of birth, level of education, native language, other language capability, place of birth, place of residence and occupation. The Mixer 3 collection contains 19,595 telephone recordings. The raw digital audio content for each call side was captured as a separate channel, then merged to be presented as a 2-channel file; the files are formatted as 8kHz, 8000 samples/second, u-law encoded NIST SPHERE files. *Samples* Please listen to this audio sample. *Updates* None at this time. *Additional Licensing Instructions* This members-only corpus is available to current members. Contact ldc@ldc.upenn.edu for information about becoming a member.

Extent: Corpus size: 172618870 KB

Format: Sampling Rate: 8000

Sampling Format: mulaw

Identifier: LDC2023S02

https://catalog.ldc.upenn.edu/LDC2023S02

ISLRN: 823-474-406-019-9

DOI: 10.35111/s9jz-3210

Language: English

Mandarin Chinese

Min Nan Chinese

Amharic

Bengali

Persian

Hindi

Italian

Japanese

Georgian

Central Khmer; Khmer

Korean

Lao

Panjabi

Western Panjabi

Russian

Spanish

Tamil

Tagalog

Thai

Tigrinya

Urdu

Uzbek

Vietnamese

Wu Chinese

Language (ISO639): eng

cmn

nan

amh

ben

fas

hin

ita

jpn

kat

khm

kor

lao

pan

pnb

rus

spa

tam

tgl

tha

tir

urd

uzb

vie

wuu

Medium: Distribution: Web Download

Publisher: Linguistic Data Consortium

Publisher (URI): https://www.ldc.upenn.edu

Relation (URI): https://catalog.ldc.upenn.edu/docs/LDC2023S02

Rights Holder: Portions © 2005-2007, 2011, 2012, 2023 Trustees of the University of Pennsylvania

Type (DCMI): Sound

Type (OLAC): primary_text

OLAC Info

Archive: The LDC Corpus Catalog

Description: http://www.language-archives.org/archive/www.ldc.upenn.edu

GetRecord: OAI-PMH request for OLAC format

GetRecord: Pre-generated XML file

OAI Info

OaiIdentifier: oai:www.ldc.upenn.edu:LDC2023S02

DateStamp: 2024-02-01

GetRecord: OAI-PMH request for simple DC format

Search Info
Citation: Huang, Shudong; Walker, Kevin; Graff, David. 2023. Linguistic Data Consortium.
Terms: area_Africa area_Asia area_Europe country_BD country_CN country_ES country_ET country_GB country_GE country_IN country_IT country_JP country_KH country_KR country_LA country_PH country_PK country_RU country_TH country_VN dcmi_Sound iso639_amh iso639_ben iso639_cmn iso639_eng iso639_fas iso639_hin iso639_ita iso639_jpn iso639_kat iso639_khm iso639_kor iso639_lao iso639_nan iso639_pan iso639_pnb iso639_rus iso639_spa iso639_tam iso639_tgl iso639_tha iso639_tir iso639_urd iso639_uzb iso639_vie iso639_wuu olac_primary_text

http://www.language-archives.org/item.php/oai:www.ldc.upenn.edu:LDC2023S02
Up-to-date as of: Wed Oct 29 7:02:10 EDT 2025

Metadata
Title:		Mixer 3 Speech
Access Rights:		Licensing Instructions for Subscription & Standard Members, and Non-Members: http://www.ldc.upenn.edu/language-resources/data/obtaining
Bibliographic Citation:		Huang, Shudong, Kevin Walker, and David Graff. Mixer 3 Speech LDC2023S02. Web Download. Philadelphia: Linguistic Data Consortium, 2023
Contributor:		Huang, Shudong
		Walker, Kevin
		Graff, David
Date (W3CDTF):		2023
Date Issued (W3CDTF):		2023-03-15
Description:		Introduction Mixer 3 Speech was developed by the Linguistic Data Consortium (LDC) and comprises 3,200 hours of audio recordings of conversational telephone speech involving 3,875 speakers and 26 distinct languages. This material was collected by LDC from 2005-2007 as part of the Mixer project, and recordings in this corpus were used in NIST Speaker Recognition Evaluation (SRE) and NIST Language Recognition Evaluation (LRE) corpora, including 2006 SRE and 2007 LRE. Researchers interested in applying those benchmark test sets should consult the respective NIST Evaluation Plans for guidelines on allowable training data for those tests. Data from 2006 SRE and 2007 LRE are available in the LDC Catalog: 2006 NIST Speaker Recognition Evaluation Training Set (LDC2011S09), 2006 NIST Speaker Recognition Evaluation Test Set Part 1 (LDC2011S10), 2006 NIST Speaker Recognition Evaluation Test Set Part 2 (LDC2012S01), 2007 NIST Language Recognition Evaluation Supplemental Training Set (LDC2009S05) and 2007 NIST Language Recognition Evaluation Test Set (LDC2009S04). Data The audio recordings were generated using LDC's computer telephony system capable of collecting speech from the telephone network. Recruited speakers were connected through a robot operator to carry on casual conversations lasting up to 10 minutes. Subjects fluent in languages other than English were asked to complete at least one non-English call. The documentation for this release contains information about the number of calls per subject and the number of calls per language. It also includes certain speaker demographic information, such as date of birth, level of education, native language, other language capability, place of birth, place of residence and occupation. The Mixer 3 collection contains 19,595 telephone recordings. The raw digital audio content for each call side was captured as a separate channel, then merged to be presented as a 2-channel file; the files are formatted as 8kHz, 8000 samples/second, u-law encoded NIST SPHERE files. Samples Please listen to this audio sample. Updates None at this time. Additional Licensing Instructions This members-only corpus is available to current members. Contact ldc@ldc.upenn.edu for information about becoming a member.
Extent:		Corpus size: 172618870 KB
Format:		Sampling Rate: 8000
Format:		Sampling Format: mulaw
Identifier:		LDC2023S02
		https://catalog.ldc.upenn.edu/LDC2023S02
		ISLRN: 823-474-406-019-9
		DOI: 10.35111/s9jz-3210
Language:		English
		Mandarin Chinese
		Min Nan Chinese
		Amharic
		Bengali
		Persian
		Hindi
		Italian
		Japanese
		Georgian
		Central Khmer; Khmer
		Korean
		Lao
		Panjabi
		Western Panjabi
		Russian
		Spanish
		Tamil
		Tagalog
		Thai
		Tigrinya
		Urdu
		Uzbek
		Vietnamese
		Wu Chinese
Language (ISO639):		eng
		cmn
		nan
		amh
		ben
		fas
		hin
		ita
		jpn
		kat
		khm
		kor
		lao
		pan
		pnb
		rus
		spa
		tam
		tgl
		tha
		tir
		urd
		uzb
		vie
		wuu
Medium:		Distribution: Web Download
Publisher:		Linguistic Data Consortium
Publisher (URI):		https://www.ldc.upenn.edu
Relation (URI):		https://catalog.ldc.upenn.edu/docs/LDC2023S02
Rights Holder:		Portions © 2005-2007, 2011, 2012, 2023 Trustees of the University of Pennsylvania
Type (DCMI):		Sound
Type (OLAC):		primary_text
OLAC Info
Archive:		The LDC Corpus Catalog
Description:		http://www.language-archives.org/archive/www.ldc.upenn.edu
GetRecord:		OAI-PMH request for OLAC format
GetRecord:		Pre-generated XML file
OAI Info
OaiIdentifier:		oai:www.ldc.upenn.edu:LDC2023S02
DateStamp:		2024-02-01
GetRecord:		OAI-PMH request for simple DC format
Search Info
Citation:		Huang, Shudong; Walker, Kevin; Graff, David. 2023. Linguistic Data Consortium.
Terms:		area_Africa area_Asia area_Europe country_BD country_CN country_ES country_ET country_GB country_GE country_IN country_IT country_JP country_KH country_KR country_LA country_PH country_PK country_RU country_TH country_VN dcmi_Sound iso639_amh iso639_ben iso639_cmn iso639_eng iso639_fas iso639_hin iso639_ita iso639_jpn iso639_kat iso639_khm iso639_kor iso639_lao iso639_nan iso639_pan iso639_pnb iso639_rus iso639_spa iso639_tam iso639_tgl iso639_tha iso639_tir iso639_urd iso639_uzb iso639_vie iso639_wuu olac_primary_text