SPID Corpus

The SPID Corpus


	Contents:
	Overview
	Map Task
	Signal Examples
	State of the Corpus
	Download

The "SPID" (SPontaneous In-car Dialogues) corpus allows investigating - for the first time ever - the communication between speakers inside a driving car and the Lombard effect that emerges at different driving speeds. Currently, speech recordings are made in order to analyze the effects of an in-car-communication (ICC) system on speech production at different driving noise levels. ICC systems are meant to improve the communication of passengers inside a car and thus help increase driving safety.

At the heart of the SPID corpus is an acoustic ambiance simulation. That is, the speakers sit inside a stationary car and hear realistic driving noises of exactly this car. The acoustic simulation is further complemented by a visual (screen-based) projection of real driving situations that match with the driving noises. On this basis, the Lombard effect can be investigated under highly sophisticated and at the same time highly controlled laboratory conditions; and in addition, the noise can be entirely removed again from the signals after the recordings by means of adaptive cancellation and suppression approaches. Thereby, Lombardaffected speech features like intonation, stress, and formants can be analyzed in full detail, without the corresponding measurements being distorted by background noise.

The recordings themselves were based on the map-task paradigm, which elicits spontaneous speech with a number of selected target words included (names of streets, places, persons etc.). One speaker sat in the front passenger seat and the other one behind him/her on the backseat. The two speakers had their own microphone. Recordings were made in driving simulations at 50 km/h (city) and 130 km/h (highway), as well as in a silent reference condition.

Key Features

The corpus consists of approximately 25 hours of spontaneous dialogues whose individual durations vary depending on how long it took the dialogue partners to solve their Map Task. The speaker sample included 8 male and 8 female native speakers of Standard German. They lived for a long time in Northern Germany and were between 22 and 31 years old (average age: 26.5 years) at the time of the recording. Pairs of speakers were always of the same gender.

State of the Corpus

You can find the current state of the corpus here.

Download

The corpus can be used for non-commercial research purposes. Details can be found here.

Examples


Without an ICC system at 130 km/h.		With an ICC system at 130 km/h.

Creators of the Data Base

The data base was created as a joint work between Kiel University (CAU) and the University of Southern Denmark (SDU, Mads Clausen Institute). Involved researchers are:

Tina John (CAU)
Rabea Landgraf (CAU)
Christian Lüke (CAU)
Oliver Niebuhr (SDU)
Gerhard Schmidt (CAU)
Anne Theiss (CAU)

Corresponding Publications

{loadpapers:authors=31:categories=any:years=any:months=any:lab=any:style=div}

Visit of the Hans Böckler Foundation

The Hans Böckler Foundation offers students not only financial support, but also a wide range of seminars. We had the pleasure of taking part in an exciting seminar on the topic of ‘Data channels in the seabed - insights into the underwater infrastructure of the future’. We spent one day at the Faculty of Engineering at Kiel University. We were given presentations on various topics. Such as the geology of the seabed and the submarine cable incident between Finland and Estonia.

We were particularly impressed by the opportunity to take a look behind the scenes and experience the work of the students and researchers up close. We were allowed to visit the clean room and the special ‘paddling pool’ where experiments are tested directly in water. We were deeply impressed by the university and its diverse research opportunities. Thank you very much for your hospitality and the exciting insights!

Text and photo by Luise Artmann