Matches in ScholarlyData for { <https://w3id.org/scholarlydata/inproceedings/lrec2008/papers/57> ?p ?o. }
Showing items 1 to 13 of
13
with 100 items per page.
- 57 creator krzysztof-marasek.
- 57 creator ryszard-gubrynowicz.
- 57 type InProceedings.
- 57 label "Design and Data Collection for Spoken Polish Dialogs Database".
- 57 sameAs 57.
- 57 abstract "Spoken corpora provide a critical resource for research, development and evaluation of spoken dialog systems. This paper describes the telephone spoken dialog corpus for Polish created by Polish-Japanese Institute of Information Technology team within the LUNA project (IST 033549). The main goal of this project is to create a robust natural spoken language understanding (SLU) toolkit, which can be used to improve the speech-enabled telecom services in multilingual context (Italian, French and Polish). The corpus has been collected at the call center of Warsaw Transport Authority, manually transcribed and richly annotated on acoustic, syntactic and semantic levels. The most frequent users requests concern city traffic information (public transportation stops, routes, schedules, trip planning etc.). The collected database consists of two parts: 500 human-human dialogs of approx. 670 minutes long with a vocabulary of ca. 8,000 words and 500 human-machine dialogs recorded via the use of Wizard-of-Oz paradigm. The syntactic and semantic annotation is carried out by another team (Mykowiecka et al., 2007). This database is the first one collected for spontaneous Polish speech recorded through telecommunication lines and will be used for development and evaluation of automatic speech recognition (ASR) and robust natural spoken language understanding (SLU) components.".
- 57 hasAuthorList authorList.
- 57 hasTopic Linguistics.
- 57 isPartOf proceedings.
- 57 keyword "Corpus (creation, annotation, etc.)".
- 57 keyword "Dialogue & Natural Interactivity".
- 57 keyword "Speech resource/database".
- 57 title "Design and Data Collection for Spoken Polish Dialogs Database".