The SPID Corpus

The "SPID" (SPontaneous In-car Dialogues) corpus allows investigating - for the first time ever - the communication between speakers inside a driving car and the Lombard effect that emerges at different driving speeds. Currently, speech recordings are made in order to analyze the effects of an in-car-communication (ICC) system on speech production at different driving noise levels. ICC systems are meant to improve the communication of passengers inside a car and thus help increase driving safety.

At the heart of the SPID corpus is an acoustic ambiance simulation. That is, the speakers sit inside a stationary car and hear realistic driving noises of exactly this car. The acoustic simulation is further complemented by a visual (screen-based) projection of real driving situations that match with the driving noises. On this basis, the Lombard effect can be investigated under highly sophisticated and at the same time highly controlled laboratory conditions; and in addition, the noise can be entirely removed again from the signals after the recordings by means of adaptive cancellation and suppression approaches. Thereby, Lombardaffected speech features like intonation, stress, and formants can be analyzed in full detail, without the corresponding measurements being distorted by background noise.

The recordings themselves were based on the Map-Task paradigm, which elicits spontaneous speech with a number of selected target words included (names of streets, places, persons etc.). One speaker sat in the front passenger seat and the other one behind him/her on the backseat. The two speakers had their own microphone. Recordings were made in driving simulations at 50 km/h (city) and 130 km/h (highway), as well as in a silent reference condition.

 

Key Features

The corpus consists of approximately 25 hours of spontaneous dialogues whose individual durations vary depending on how long it took the dialogue partners to solve their Map Task. The speaker sample included 8 male and 8 female native speakers of Standard German. They lived for a long time in Northern Germany and were between 22 and 31 years old (average age: 26.5 years) at the time of the recording. Pairs of speakers were always of the same gender.

 

State of the Corpus

You can find the current state of the corpus here.

 

Download

The corpus can be used for non-commercial research purposes. Details can be found here.

 

Examples

 
Without an ICC system at 130 km/h.   With an ICC system at 130 km/h.

 

Creators of the Data Base

The data base was created as a joint work between Kiel University (CAU) and the University of Southern Denmark (SDU, Mads Clausen Institute). Involved researchers are:

  • Tina John (CAU)
  • Rabea Landgraf (CAU)
  • Christian Lüke (CAU)
  • Oliver Niebuhr (SDU)
  • Gerhard Schmidt (CAU)
  • Anne Theiss (CAU)

 

Corresponding Publications

   

R. Landgraf, T. John, G. Schmidt, M. Gimm, O. Niebuhr: Man höre und staune – Perzeptive Evidenz für die Funktionalität eines ICC-Systems anhand von Dialogdaten aus einer multimodalen Fahrsimulation, DAGA 2018

   

R. Landgraf, G. Schmidt, J. Köhler-Kaeß, O. Niebuhr, and T. John: More Noise, Less Talk - The Impact of Driving Noise and ICC Systems on Acoustic-prosodic Parameters in Dialogue, Proc. DAGA, Kiel, Germany, open access, 2017

   

R. Landgraf, J. Köhler-Kaeß, C. Lüke, O. Niebuhr, G. Schmidt: Can You Hear Me Now? Reducing the Lombard Effect in a Driving Car Using an In-Car Communication System, Proc. Speech Prosody, pp. 479 - 483, 2016

   

A. Warhadpande, C. Lüke, A. Theiß, G. Schmidt: Improvement by Adding Video Feature in an Acoustic Ambiance Simulation for Automobiles, 7th Biennial Workshop on DSP for In-Vehicle Systems and Safety 2015, Berkeley, CA, USA

   

O. Niebuhr, B. Peters, R. Landgraf, G. Schmidt: The Kiel Corpora of "Speech & Emotion" - A Summary, Proc. DAGA 2015, March 16-19, 2015, Nürnberg, Germany

   

R. Landgraf, O. Niebuhr, G. Schmidt, T. John, C. Lüke, A. Theiß: Von der Straße ins Labor: Die Modifikation der Sprachproduktion bei lauten Fahrgeräuschen, DAGA 2015, Nürnberg, Germany

   

C. Lüke, A. Theiß, G. Schmidt, O. Niebuhr, T. John: Creation of a Lombard Speech Database using an Acoustic Ambiance Simulation with Loudspeakers, 6th Biennial DSP Workshop for In-Vehicle Systems 2013, Seoul, Korea

Website News

07.08.2019: Talk from Juan Rafael Orozco-Arroyave added.

11.07.2019: First free KiRAT version released - a game for Parkinson patients

25.06.2019: About 30 pupils from the Isarnwohld-Schule in Gettorf visited us.

02.05.2019: Christin Baasch finished sucessfully her defense on the evaluation of Parkinson speech.

30.11.2018: New student project on driver distraction added.

Recent Publications

   

J. Reermann, E. Elzenheimer and G. Schmidt: Real-time Biomagnetic Signal Processing for Uncooled Magnetometers in Cardiology, IEEE Sensors Journal, Volume 15, Number 10, Pages 4237-4249, June 2019, doi: 10.1109/JSEN.2019.2893236

Contact

Prof. Dr.-Ing. Gerhard Schmidt

E-Mail: gus@tf.uni-kiel.de

Christian-Albrechts-Universität zu Kiel
Faculty of Engineering
Institute for Electrical Engineering and Information Engineering
Digital Signal Processing and System Theory

Kaiserstr. 2
24143 Kiel, Germany

How to find us

Recent News

First Water Test with CASSy

On Thursday, August 29th, we tested CASSy the first time in real water. With the help of our collegues from the Bundeswehr (special thanks to Dr. Jan Abshagen and Dr. Arne Stoltenberg) we were able to put CASSy into the Baltic sea. After fixing some engine problems, we had your first "cruise" - a short trip from Kiel to Mönkeberg and back. Everything worked fine. Now, we feel confident to go ...


Read more ...