---
_id: '9456'
abstract:
- lang: eng
  text: In this paper we present our experimental results about classifying audio
    data into broad acoustic categories. The reverberated sound samples from indoor
    recordings are grouped into four classes, namely speech, music, acoustic events
    and noise. We investigated a total of 188 acoustic features and achieved for the
    best configuration a classification accuracy better than 98\%. This was achieved
    by a 42-dimensional feature vector consisting of Mel-Frequency Cepstral Coefficients,
    an autocorrelation feature and so-called track features that measure the length
    of ''traces'' of high energy in the spectrogram. We also found a 4-feature configuration
    with a classification rate of about 90\% allowing for broad acoustic category
    classification with low computational effort.
author:
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Markus
  full_name: Bartek, Markus
  last_name: Bartek
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: 'Schmalenstroeer J, Bartek M, Haeb-Umbach R. Investigations into Features for
    Robust Classification into Broad Acoustic Categories. In: <i>37. Deutsche Jahrestagung
    Fuer Akustik (DAGA 2011)</i>. ; 2011.'
  apa: Schmalenstroeer, J., Bartek, M., &#38; Haeb-Umbach, R. (2011). Investigations
    into Features for Robust Classification into Broad Acoustic Categories. <i>37.
    Deutsche Jahrestagung Fuer Akustik (DAGA 2011)</i>.
  bibtex: '@inproceedings{Schmalenstroeer_Bartek_Haeb-Umbach_2011, title={Investigations
    into Features for Robust Classification into Broad Acoustic Categories}, booktitle={37.
    Deutsche Jahrestagung fuer Akustik (DAGA 2011)}, author={Schmalenstroeer, Joerg
    and Bartek, Markus and Haeb-Umbach, Reinhold}, year={2011} }'
  chicago: Schmalenstroeer, Joerg, Markus Bartek, and Reinhold Haeb-Umbach. “Investigations
    into Features for Robust Classification into Broad Acoustic Categories.” In <i>37.
    Deutsche Jahrestagung Fuer Akustik (DAGA 2011)</i>, 2011.
  ieee: J. Schmalenstroeer, M. Bartek, and R. Haeb-Umbach, “Investigations into Features
    for Robust Classification into Broad Acoustic Categories,” 2011.
  mla: Schmalenstroeer, Joerg, et al. “Investigations into Features for Robust Classification
    into Broad Acoustic Categories.” <i>37. Deutsche Jahrestagung Fuer Akustik (DAGA
    2011)</i>, 2011.
  short: 'J. Schmalenstroeer, M. Bartek, R. Haeb-Umbach, in: 37. Deutsche Jahrestagung
    Fuer Akustik (DAGA 2011), 2011.'
date_created: 2019-04-25T14:36:02Z
date_updated: 2023-10-26T08:15:44Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2011/ScBaHa11-1.pdf
oa: '1'
publication: 37. Deutsche Jahrestagung fuer Akustik (DAGA 2011)
quality_controlled: '1'
status: public
title: Investigations into Features for Robust Classification into Broad Acoustic
  Categories
type: conference
user_id: '460'
year: '2011'
...
---
_id: '11892'
abstract:
- lang: eng
  text: For an environment to be perceived as being smart, contextual information
    has to be gathered to adapt the system's behavior and its interface towards the
    user. Being a rich source of context information speech can be acquired unobtrusively
    by microphone arrays and then processed to extract information about the user
    and his environment. In this paper, a system for joint temporal segmentation,
    speaker localization, and identification is presented, which is supported by face
    identification from video data obtained from a steerable camera. Special attention
    is paid to latency aspects and online processing capabilities, as they are important
    for the application under investigation, namely ambient communication. It describes
    the vision of terminal-less, session-less and multi-modal telecommunication with
    remote partners, where the user can move freely within his home while the communication
    follows him. The speaker diarization serves as a context source, which has been
    integrated in a service-oriented middleware architecture and provided to the application
    to select the most appropriate I/O device and to steer the camera towards the
    speaker during ambient communication.
author:
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: Schmalenstroeer J, Haeb-Umbach R. Online Diarization of Streaming Audio-Visual
    Data for Smart Environments. <i>IEEE Journal of Selected Topics in Signal Processing</i>.
    2010;4(5):845-856. doi:<a href="https://doi.org/10.1109/JSTSP.2010.2050519">10.1109/JSTSP.2010.2050519</a>
  apa: Schmalenstroeer, J., &#38; Haeb-Umbach, R. (2010). Online Diarization of Streaming
    Audio-Visual Data for Smart Environments. <i>IEEE Journal of Selected Topics in
    Signal Processing</i>, <i>4</i>(5), 845–856. <a href="https://doi.org/10.1109/JSTSP.2010.2050519">https://doi.org/10.1109/JSTSP.2010.2050519</a>
  bibtex: '@article{Schmalenstroeer_Haeb-Umbach_2010, title={Online Diarization of
    Streaming Audio-Visual Data for Smart Environments}, volume={4}, DOI={<a href="https://doi.org/10.1109/JSTSP.2010.2050519">10.1109/JSTSP.2010.2050519</a>},
    number={5}, journal={IEEE Journal of Selected Topics in Signal Processing}, author={Schmalenstroeer,
    Joerg and Haeb-Umbach, Reinhold}, year={2010}, pages={845–856} }'
  chicago: 'Schmalenstroeer, Joerg, and Reinhold Haeb-Umbach. “Online Diarization
    of Streaming Audio-Visual Data for Smart Environments.” <i>IEEE Journal of Selected
    Topics in Signal Processing</i> 4, no. 5 (2010): 845–56. <a href="https://doi.org/10.1109/JSTSP.2010.2050519">https://doi.org/10.1109/JSTSP.2010.2050519</a>.'
  ieee: 'J. Schmalenstroeer and R. Haeb-Umbach, “Online Diarization of Streaming Audio-Visual
    Data for Smart Environments,” <i>IEEE Journal of Selected Topics in Signal Processing</i>,
    vol. 4, no. 5, pp. 845–856, 2010, doi: <a href="https://doi.org/10.1109/JSTSP.2010.2050519">10.1109/JSTSP.2010.2050519</a>.'
  mla: Schmalenstroeer, Joerg, and Reinhold Haeb-Umbach. “Online Diarization of Streaming
    Audio-Visual Data for Smart Environments.” <i>IEEE Journal of Selected Topics
    in Signal Processing</i>, vol. 4, no. 5, 2010, pp. 845–56, doi:<a href="https://doi.org/10.1109/JSTSP.2010.2050519">10.1109/JSTSP.2010.2050519</a>.
  short: J. Schmalenstroeer, R. Haeb-Umbach, IEEE Journal of Selected Topics in Signal
    Processing 4 (2010) 845–856.
date_created: 2019-07-12T05:30:16Z
date_updated: 2023-10-26T08:10:18Z
department:
- _id: '54'
doi: 10.1109/JSTSP.2010.2050519
intvolume: '         4'
issue: '5'
keyword:
- audio streaming
- audio visual data streaming
- context information speech
- face identification
- face recognition
- image segmentation
- middleware
- multimodal telecommunication
- online diarization
- service oriented middleware architecture
- sessionless telecommunication
- software architecture
- speaker identification
- speaker localization
- speaker recognition
- steerable camera
- telecommunication computing
- temporal segmentation
- terminal-less telecommunication
- video streaming
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2010/ScHa10.pdf
oa: '1'
page: 845-856
publication: IEEE Journal of Selected Topics in Signal Processing
quality_controlled: '1'
status: public
title: Online Diarization of Streaming Audio-Visual Data for Smart Environments
type: journal_article
user_id: '460'
volume: 4
year: '2010'
...
---
_id: '11900'
author:
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Volker
  full_name: Leutnant, Volker
  last_name: Leutnant
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: 'Schmalenstroeer J, Leutnant V, Haeb-Umbach R. Audio-Visual Data Processing
    for Ambient Communication. In: <i>1st International Workshop on Distributed Computing
    in Ambient Environments within 32nd Annual Conference on Artificial Intelligence</i>.
    ; 2009.'
  apa: Schmalenstroeer, J., Leutnant, V., &#38; Haeb-Umbach, R. (2009). Audio-Visual
    Data Processing for Ambient Communication. <i>1st International Workshop on Distributed
    Computing in Ambient Environments within 32nd Annual Conference on Artificial
    Intelligence</i>.
  bibtex: '@inproceedings{Schmalenstroeer_Leutnant_Haeb-Umbach_2009, title={Audio-Visual
    Data Processing for Ambient Communication}, booktitle={1st International Workshop
    on Distributed Computing in Ambient Environments within 32nd Annual Conference
    on Artificial Intelligence}, author={Schmalenstroeer, Joerg and Leutnant, Volker
    and Haeb-Umbach, Reinhold}, year={2009} }'
  chicago: Schmalenstroeer, Joerg, Volker Leutnant, and Reinhold Haeb-Umbach. “Audio-Visual
    Data Processing for Ambient Communication.” In <i>1st International Workshop on
    Distributed Computing in Ambient Environments within 32nd Annual Conference on
    Artificial Intelligence</i>, 2009.
  ieee: J. Schmalenstroeer, V. Leutnant, and R. Haeb-Umbach, “Audio-Visual Data Processing
    for Ambient Communication,” 2009.
  mla: Schmalenstroeer, Joerg, et al. “Audio-Visual Data Processing for Ambient Communication.”
    <i>1st International Workshop on Distributed Computing in Ambient Environments
    within 32nd Annual Conference on Artificial Intelligence</i>, 2009.
  short: 'J. Schmalenstroeer, V. Leutnant, R. Haeb-Umbach, in: 1st International Workshop
    on Distributed Computing in Ambient Environments within 32nd Annual Conference
    on Artificial Intelligence, 2009.'
date_created: 2019-07-12T05:30:25Z
date_updated: 2023-11-15T15:03:08Z
ddc:
- '004'
department:
- _id: '54'
file:
- access_level: open_access
  content_type: application/pdf
  creator: schmalen
  date_created: 2023-11-15T15:02:34Z
  date_updated: 2023-11-15T15:02:34Z
  file_id: '48934'
  file_name: SchLeuHae09.pdf
  file_size: 98062
  relation: main_file
file_date_updated: 2023-11-15T15:02:34Z
has_accepted_license: '1'
language:
- iso: eng
oa: '1'
publication: 1st International Workshop on Distributed Computing in Ambient Environments
  within 32nd Annual Conference on Artificial Intelligence
quality_controlled: '1'
status: public
title: Audio-Visual Data Processing for Ambient Communication
type: conference
user_id: '460'
year: '2009'
...
---
_id: '11806'
abstract:
- lang: eng
  text: Microphone arrays represent the basis for many challenging acoustic sensing
    tasks. The accuracy of techniques like beamforming directly depends on a precise
    knowledge of the relative positions of the sensors used. Unfortunately, for certain
    use cases manually measuring the geometry of an array is not feasible due to practical
    constraints. In this paper we present an approach to unsupervised shape calibration
    of microphone array networks. We developed a hierarchical procedure that first
    performs local shape calibration based on coherence analysis and then employs
    SRP-PHAT in a network calibration method. Practical experiments demonstrate the
    effectiveness of our approach especially for highly reverberant acoustic environments.
author:
- first_name: Marius
  full_name: Hennecke, Marius
  last_name: Hennecke
- first_name: Thomas
  full_name: Ploetz, Thomas
  last_name: Ploetz
- first_name: Gernot A.
  full_name: Fink, Gernot A.
  last_name: Fink
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: 'Hennecke M, Ploetz T, Fink GA, Schmalenstroeer J, Haeb-Umbach R. A hierarchical
    approach to unsupervised shape calibration of microphone array networks. In: <i>IEEE/SP
    15th Workshop on Statistical Signal Processing (SSP 2009)</i>. ; 2009:257-260.
    doi:<a href="https://doi.org/10.1109/SSP.2009.5278589">10.1109/SSP.2009.5278589</a>'
  apa: Hennecke, M., Ploetz, T., Fink, G. A., Schmalenstroeer, J., &#38; Haeb-Umbach,
    R. (2009). A hierarchical approach to unsupervised shape calibration of microphone
    array networks. <i>IEEE/SP 15th Workshop on Statistical Signal Processing (SSP
    2009)</i>, 257–260. <a href="https://doi.org/10.1109/SSP.2009.5278589">https://doi.org/10.1109/SSP.2009.5278589</a>
  bibtex: '@inproceedings{Hennecke_Ploetz_Fink_Schmalenstroeer_Haeb-Umbach_2009, title={A
    hierarchical approach to unsupervised shape calibration of microphone array networks},
    DOI={<a href="https://doi.org/10.1109/SSP.2009.5278589">10.1109/SSP.2009.5278589</a>},
    booktitle={IEEE/SP 15th Workshop on Statistical Signal Processing (SSP 2009)},
    author={Hennecke, Marius and Ploetz, Thomas and Fink, Gernot A. and Schmalenstroeer,
    Joerg and Haeb-Umbach, Reinhold}, year={2009}, pages={257–260} }'
  chicago: Hennecke, Marius, Thomas Ploetz, Gernot A. Fink, Joerg Schmalenstroeer,
    and Reinhold Haeb-Umbach. “A Hierarchical Approach to Unsupervised Shape Calibration
    of Microphone Array Networks.” In <i>IEEE/SP 15th Workshop on Statistical Signal
    Processing (SSP 2009)</i>, 257–60, 2009. <a href="https://doi.org/10.1109/SSP.2009.5278589">https://doi.org/10.1109/SSP.2009.5278589</a>.
  ieee: 'M. Hennecke, T. Ploetz, G. A. Fink, J. Schmalenstroeer, and R. Haeb-Umbach,
    “A hierarchical approach to unsupervised shape calibration of microphone array
    networks,” in <i>IEEE/SP 15th Workshop on Statistical Signal Processing (SSP 2009)</i>,
    2009, pp. 257–260, doi: <a href="https://doi.org/10.1109/SSP.2009.5278589">10.1109/SSP.2009.5278589</a>.'
  mla: Hennecke, Marius, et al. “A Hierarchical Approach to Unsupervised Shape Calibration
    of Microphone Array Networks.” <i>IEEE/SP 15th Workshop on Statistical Signal
    Processing (SSP 2009)</i>, 2009, pp. 257–60, doi:<a href="https://doi.org/10.1109/SSP.2009.5278589">10.1109/SSP.2009.5278589</a>.
  short: 'M. Hennecke, T. Ploetz, G.A. Fink, J. Schmalenstroeer, R. Haeb-Umbach, in:
    IEEE/SP 15th Workshop on Statistical Signal Processing (SSP 2009), 2009, pp. 257–260.'
date_created: 2019-07-12T05:28:37Z
date_updated: 2023-10-26T08:09:22Z
department:
- _id: '54'
doi: 10.1109/SSP.2009.5278589
keyword:
- acoustic sensing tasks
- array geometry
- calibration
- coherence analysis
- hierarchical procedure
- local shape calibration
- microphone array networks
- microphone arrays
- network calibration method
- sensor arrays
- SRP-PHAT
- unsupervised shape calibration
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2009/HePlFiScHa09.pdf
oa: '1'
page: 257-260
publication: IEEE/SP 15th Workshop on Statistical Signal Processing (SSP 2009)
quality_controlled: '1'
status: public
title: A hierarchical approach to unsupervised shape calibration of microphone array
  networks
type: conference
user_id: '460'
year: '2009'
...
---
_id: '11899'
abstract:
- lang: eng
  text: In this paper we present a system for identifying and localizingspeakers using
    distant microphone arrays and a steerablepan-tilt-zoom camera. Audio and video
    streams are processedin real-time to obtain the diarization information {grqq}who
    speakswhen and where'' with low latency to be used in advanced videoconferencing
    systems or user-adaptive interfaces. A key featureof the proposed system is to
    first glean information about thespeaker{\rq}s location and identity from the
    audio and visual datastreams separately and then to fuse these data in a probabilisticframework
    employing the Viterbi algorithm. Here, visual evidenceof a person is utilized
    through a priori state probabilities,while location and speaker change information
    are employedvia time-variant transition probablities. Experiments show thatvideo
    information yields a substantial improvement comparedto pure audio-based diarization.
author:
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Martin
  full_name: Kelling, Martin
  last_name: Kelling
- first_name: Volker
  full_name: Leutnant, Volker
  last_name: Leutnant
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: 'Schmalenstroeer J, Kelling M, Leutnant V, Haeb-Umbach R. Fusing Audio and
    Video Information for Online Speaker Diarization. In: <i>Interspeech 2009</i>.
    ; 2009.'
  apa: Schmalenstroeer, J., Kelling, M., Leutnant, V., &#38; Haeb-Umbach, R. (2009).
    Fusing Audio and Video Information for Online Speaker Diarization. <i>Interspeech
    2009</i>.
  bibtex: '@inproceedings{Schmalenstroeer_Kelling_Leutnant_Haeb-Umbach_2009, title={Fusing
    Audio and Video Information for Online Speaker Diarization}, booktitle={Interspeech
    2009}, author={Schmalenstroeer, Joerg and Kelling, Martin and Leutnant, Volker
    and Haeb-Umbach, Reinhold}, year={2009} }'
  chicago: Schmalenstroeer, Joerg, Martin Kelling, Volker Leutnant, and Reinhold Haeb-Umbach.
    “Fusing Audio and Video Information for Online Speaker Diarization.” In <i>Interspeech
    2009</i>, 2009.
  ieee: J. Schmalenstroeer, M. Kelling, V. Leutnant, and R. Haeb-Umbach, “Fusing Audio
    and Video Information for Online Speaker Diarization,” 2009.
  mla: Schmalenstroeer, Joerg, et al. “Fusing Audio and Video Information for Online
    Speaker Diarization.” <i>Interspeech 2009</i>, 2009.
  short: 'J. Schmalenstroeer, M. Kelling, V. Leutnant, R. Haeb-Umbach, in: Interspeech
    2009, 2009.'
date_created: 2019-07-12T05:30:24Z
date_updated: 2023-10-26T08:10:10Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2009/ScKeLeHa09.pdf
oa: '1'
publication: Interspeech 2009
quality_controlled: '1'
status: public
title: Fusing Audio and Video Information for Online Speaker Diarization
type: conference
user_id: '460'
year: '2009'
...
---
_id: '11893'
author:
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: 'Schmalenstroeer J, Haeb-Umbach R. Joint Speaker Segmentation, Localization
    and Identification for Streaming Audio. In: <i>Interspeech 2007</i>. ; 2007.'
  apa: Schmalenstroeer, J., &#38; Haeb-Umbach, R. (2007). Joint Speaker Segmentation,
    Localization and Identification for Streaming Audio. <i>Interspeech 2007</i>.
  bibtex: '@inproceedings{Schmalenstroeer_Haeb-Umbach_2007, title={Joint Speaker Segmentation,
    Localization and Identification for Streaming Audio}, booktitle={Interspeech 2007},
    author={Schmalenstroeer, Joerg and Haeb-Umbach, Reinhold}, year={2007} }'
  chicago: Schmalenstroeer, Joerg, and Reinhold Haeb-Umbach. “Joint Speaker Segmentation,
    Localization and Identification for Streaming Audio.” In <i>Interspeech 2007</i>,
    2007.
  ieee: J. Schmalenstroeer and R. Haeb-Umbach, “Joint Speaker Segmentation, Localization
    and Identification for Streaming Audio,” 2007.
  mla: Schmalenstroeer, Joerg, and Reinhold Haeb-Umbach. “Joint Speaker Segmentation,
    Localization and Identification for Streaming Audio.” <i>Interspeech 2007</i>,
    2007.
  short: 'J. Schmalenstroeer, R. Haeb-Umbach, in: Interspeech 2007, 2007.'
date_created: 2019-07-12T05:30:17Z
date_updated: 2023-10-26T08:10:02Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2007/ScHa07.pdf
oa: '1'
publication: Interspeech 2007
quality_controlled: '1'
status: public
title: Joint Speaker Segmentation, Localization and Identification for Streaming Audio
type: conference
user_id: '460'
year: '2007'
...
---
_id: '11901'
author:
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Volker
  full_name: Leutnant, Volker
  last_name: Leutnant
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: 'Schmalenstroeer J, Leutnant V, Haeb-Umbach R. Amigo Context Management Service
    with Applications in Ambient Communication Scenarios. In: <i>AMI-07 - European
    Conference on Ambient Intelligence</i>. ; 2007.'
  apa: Schmalenstroeer, J., Leutnant, V., &#38; Haeb-Umbach, R. (2007). Amigo Context
    Management Service with Applications in Ambient Communication Scenarios. <i>AMI-07
    - European Conference on Ambient Intelligence</i>.
  bibtex: '@inproceedings{Schmalenstroeer_Leutnant_Haeb-Umbach_2007, title={Amigo
    Context Management Service with Applications in Ambient Communication Scenarios},
    booktitle={AMI-07 - European Conference on Ambient Intelligence}, author={Schmalenstroeer,
    Joerg and Leutnant, Volker and Haeb-Umbach, Reinhold}, year={2007} }'
  chicago: Schmalenstroeer, Joerg, Volker Leutnant, and Reinhold Haeb-Umbach. “Amigo
    Context Management Service with Applications in Ambient Communication Scenarios.”
    In <i>AMI-07 - European Conference on Ambient Intelligence</i>, 2007.
  ieee: J. Schmalenstroeer, V. Leutnant, and R. Haeb-Umbach, “Amigo Context Management
    Service with Applications in Ambient Communication Scenarios,” 2007.
  mla: Schmalenstroeer, Joerg, et al. “Amigo Context Management Service with Applications
    in Ambient Communication Scenarios.” <i>AMI-07 - European Conference on Ambient
    Intelligence</i>, 2007.
  short: 'J. Schmalenstroeer, V. Leutnant, R. Haeb-Umbach, in: AMI-07 - European Conference
    on Ambient Intelligence, 2007.'
date_created: 2019-07-12T05:30:27Z
date_updated: 2023-10-26T08:13:09Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2007/ScLeHa07.pdf
oa: '1'
publication: AMI-07 - European Conference on Ambient Intelligence
quality_controlled: '1'
status: public
title: Amigo Context Management Service with Applications in Ambient Communication
  Scenarios
type: conference
user_id: '460'
year: '2007'
...
---
_id: '11933'
author:
- first_name: Ernst
  full_name: Warsitz, Ernst
  last_name: Warsitz
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
citation:
  ama: 'Warsitz E, Haeb-Umbach R, Schmalenstroeer J. Zweistufige Sprache/Pause-Detektion
    in stark gestoerter Umgebung. In: <i>33. Deutsche Jahrestagung Fuer Akustik (DAGA
    2007)</i>. ; 2007.'
  apa: Warsitz, E., Haeb-Umbach, R., &#38; Schmalenstroeer, J. (2007). Zweistufige
    Sprache/Pause-Detektion in stark gestoerter Umgebung. <i>33. Deutsche Jahrestagung
    Fuer Akustik (DAGA 2007)</i>.
  bibtex: '@inproceedings{Warsitz_Haeb-Umbach_Schmalenstroeer_2007, title={Zweistufige
    Sprache/Pause-Detektion in stark gestoerter Umgebung}, booktitle={33. Deutsche
    Jahrestagung fuer Akustik (DAGA 2007)}, author={Warsitz, Ernst and Haeb-Umbach,
    Reinhold and Schmalenstroeer, Joerg}, year={2007} }'
  chicago: Warsitz, Ernst, Reinhold Haeb-Umbach, and Joerg Schmalenstroeer. “Zweistufige
    Sprache/Pause-Detektion in Stark Gestoerter Umgebung.” In <i>33. Deutsche Jahrestagung
    Fuer Akustik (DAGA 2007)</i>, 2007.
  ieee: E. Warsitz, R. Haeb-Umbach, and J. Schmalenstroeer, “Zweistufige Sprache/Pause-Detektion
    in stark gestoerter Umgebung,” 2007.
  mla: Warsitz, Ernst, et al. “Zweistufige Sprache/Pause-Detektion in Stark Gestoerter
    Umgebung.” <i>33. Deutsche Jahrestagung Fuer Akustik (DAGA 2007)</i>, 2007.
  short: 'E. Warsitz, R. Haeb-Umbach, J. Schmalenstroeer, in: 33. Deutsche Jahrestagung
    Fuer Akustik (DAGA 2007), 2007.'
date_created: 2019-07-12T05:31:04Z
date_updated: 2023-10-26T08:12:49Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2007/WaHaSc07.pdf
oa: '1'
publication: 33. Deutsche Jahrestagung fuer Akustik (DAGA 2007)
quality_controlled: '1'
status: public
title: Zweistufige Sprache/Pause-Detektion in stark gestoerter Umgebung
type: conference
user_id: '460'
year: '2007'
...
---
_id: '11902'
author:
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Ernst
  full_name: Warsitz, Ernst
  last_name: Warsitz
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: 'Schmalenstroeer J, Warsitz E, Haeb-Umbach R. Projekt Amigo - Sprachsignalverarbeitung
    im vernetzten Haus. In: <i>33. Deutsche Jahrestagung Fuer Akustik (DAGA 2007)</i>.
    ; 2007.'
  apa: Schmalenstroeer, J., Warsitz, E., &#38; Haeb-Umbach, R. (2007). Projekt Amigo
    - Sprachsignalverarbeitung im vernetzten Haus. <i>33. Deutsche Jahrestagung Fuer
    Akustik (DAGA 2007)</i>.
  bibtex: '@inproceedings{Schmalenstroeer_Warsitz_Haeb-Umbach_2007, title={Projekt
    Amigo - Sprachsignalverarbeitung im vernetzten Haus}, booktitle={33. Deutsche
    Jahrestagung fuer Akustik (DAGA 2007)}, author={Schmalenstroeer, Joerg and Warsitz,
    Ernst and Haeb-Umbach, Reinhold}, year={2007} }'
  chicago: Schmalenstroeer, Joerg, Ernst Warsitz, and Reinhold Haeb-Umbach. “Projekt
    Amigo - Sprachsignalverarbeitung Im Vernetzten Haus.” In <i>33. Deutsche Jahrestagung
    Fuer Akustik (DAGA 2007)</i>, 2007.
  ieee: J. Schmalenstroeer, E. Warsitz, and R. Haeb-Umbach, “Projekt Amigo - Sprachsignalverarbeitung
    im vernetzten Haus,” 2007.
  mla: Schmalenstroeer, Joerg, et al. “Projekt Amigo - Sprachsignalverarbeitung Im
    Vernetzten Haus.” <i>33. Deutsche Jahrestagung Fuer Akustik (DAGA 2007)</i>, 2007.
  short: 'J. Schmalenstroeer, E. Warsitz, R. Haeb-Umbach, in: 33. Deutsche Jahrestagung
    Fuer Akustik (DAGA 2007), 2007.'
date_created: 2019-07-12T05:30:28Z
date_updated: 2023-10-26T08:13:01Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2007/ScWaHa07.pdf
oa: '1'
publication: 33. Deutsche Jahrestagung fuer Akustik (DAGA 2007)
quality_controlled: '1'
status: public
title: Projekt Amigo - Sprachsignalverarbeitung im vernetzten Haus
type: conference
user_id: '460'
year: '2007'
...
---
_id: '11894'
abstract:
- lang: eng
  text: In this paper we consider the problem of detecting speaker changes in audio
    signals recorded by distant microphones. It is shown that the possibility to exploit
    the spatial separation of speakers more than makes up the degradation in detection
    accuracy due to the increased source-to-sensor distance compared to close-talking
    microphones. Speaker direction information is derived from the filter coefficients
    of an adaptive Filter-and-Sum Beamformer and is combined with BIC analysis. The
    experimental results reveal significant improvements compared to BIC-only change
    detection, be it with the distant or close-talking microphone.
author:
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
citation:
  ama: 'Schmalenstroeer J, Haeb-Umbach R. Online Speaker Change Detection by Combining
    BIC with Microphone Array Beamforming. In: <i>Interspeech 2006</i>. ; 2006.'
  apa: Schmalenstroeer, J., &#38; Haeb-Umbach, R. (2006). Online Speaker Change Detection
    by Combining BIC with Microphone Array Beamforming. <i>Interspeech 2006</i>.
  bibtex: '@inproceedings{Schmalenstroeer_Haeb-Umbach_2006, title={Online Speaker
    Change Detection by Combining BIC with Microphone Array Beamforming}, booktitle={Interspeech
    2006}, author={Schmalenstroeer, Joerg and Haeb-Umbach, Reinhold}, year={2006}
    }'
  chicago: Schmalenstroeer, Joerg, and Reinhold Haeb-Umbach. “Online Speaker Change
    Detection by Combining BIC with Microphone Array Beamforming.” In <i>Interspeech
    2006</i>, 2006.
  ieee: J. Schmalenstroeer and R. Haeb-Umbach, “Online Speaker Change Detection by
    Combining BIC with Microphone Array Beamforming,” 2006.
  mla: Schmalenstroeer, Joerg, and Reinhold Haeb-Umbach. “Online Speaker Change Detection
    by Combining BIC with Microphone Array Beamforming.” <i>Interspeech 2006</i>,
    2006.
  short: 'J. Schmalenstroeer, R. Haeb-Umbach, in: Interspeech 2006, 2006.'
date_created: 2019-07-12T05:30:19Z
date_updated: 2023-10-26T08:09:54Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2006/ScHa06.pdf
oa: '1'
publication: Interspeech 2006
quality_controlled: '1'
status: public
title: Online Speaker Change Detection by Combining BIC with Microphone Array Beamforming
type: conference
user_id: '460'
year: '2006'
...
---
_id: '11802'
author:
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
citation:
  ama: 'Haeb-Umbach R, Schmalenstroeer J. Speech Processing in the Networked Home
    Environment - A View on the Amigo Project. In: <i>Interspeech 2005</i>. ; 2005.'
  apa: Haeb-Umbach, R., &#38; Schmalenstroeer, J. (2005). Speech Processing in the
    Networked Home Environment - A View on the Amigo Project. <i>Interspeech 2005</i>.
  bibtex: '@inproceedings{Haeb-Umbach_Schmalenstroeer_2005, place={Lisboa}, title={Speech
    Processing in the Networked Home Environment - A View on the Amigo Project}, booktitle={Interspeech
    2005}, author={Haeb-Umbach, Reinhold and Schmalenstroeer, Joerg}, year={2005}
    }'
  chicago: Haeb-Umbach, Reinhold, and Joerg Schmalenstroeer. “Speech Processing in
    the Networked Home Environment - A View on the Amigo Project.” In <i>Interspeech
    2005</i>. Lisboa, 2005.
  ieee: R. Haeb-Umbach and J. Schmalenstroeer, “Speech Processing in the Networked
    Home Environment - A View on the Amigo Project,” 2005.
  mla: Haeb-Umbach, Reinhold, and Joerg Schmalenstroeer. “Speech Processing in the
    Networked Home Environment - A View on the Amigo Project.” <i>Interspeech 2005</i>,
    2005.
  short: 'R. Haeb-Umbach, J. Schmalenstroeer, in: Interspeech 2005, Lisboa, 2005.'
date_created: 2019-07-12T05:28:32Z
date_updated: 2023-10-26T08:09:00Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2005/HaSc05-2.pdf
oa: '1'
place: Lisboa
publication: Interspeech 2005
quality_controlled: '1'
status: public
title: Speech Processing in the Networked Home Environment - A View on the Amigo Project
type: conference
user_id: '460'
year: '2005'
...
---
_id: '11801'
author:
- first_name: Reinhold
  full_name: Haeb-Umbach, Reinhold
  id: '242'
  last_name: Haeb-Umbach
- first_name: Joerg
  full_name: Schmalenstroeer, Joerg
  id: '460'
  last_name: Schmalenstroeer
citation:
  ama: 'Haeb-Umbach R, Schmalenstroeer J. A Comparison of Particle Filtering Variants
    for Speech Feature Enhancement. In: <i>Interspeech 2005</i>. ; 2005.'
  apa: Haeb-Umbach, R., &#38; Schmalenstroeer, J. (2005). A Comparison of Particle
    Filtering Variants for Speech Feature Enhancement. <i>Interspeech 2005</i>.
  bibtex: '@inproceedings{Haeb-Umbach_Schmalenstroeer_2005, title={A Comparison of
    Particle Filtering Variants for Speech Feature Enhancement}, booktitle={Interspeech
    2005}, author={Haeb-Umbach, Reinhold and Schmalenstroeer, Joerg}, year={2005}
    }'
  chicago: Haeb-Umbach, Reinhold, and Joerg Schmalenstroeer. “A Comparison of Particle
    Filtering Variants for Speech Feature Enhancement.” In <i>Interspeech 2005</i>,
    2005.
  ieee: R. Haeb-Umbach and J. Schmalenstroeer, “A Comparison of Particle Filtering
    Variants for Speech Feature Enhancement,” 2005.
  mla: Haeb-Umbach, Reinhold, and Joerg Schmalenstroeer. “A Comparison of Particle
    Filtering Variants for Speech Feature Enhancement.” <i>Interspeech 2005</i>, 2005.
  short: 'R. Haeb-Umbach, J. Schmalenstroeer, in: Interspeech 2005, 2005.'
date_created: 2019-07-12T05:28:31Z
date_updated: 2023-10-26T08:09:13Z
department:
- _id: '54'
language:
- iso: eng
main_file_link:
- open_access: '1'
  url: https://groups.uni-paderborn.de/nt/pubs/2005/HaSc05-1.pdf
oa: '1'
publication: Interspeech 2005
quality_controlled: '1'
status: public
title: A Comparison of Particle Filtering Variants for Speech Feature Enhancement
type: conference
user_id: '460'
year: '2005'
...
