Dr. Malte Belz erforscht die Phonetik und Variation des Deutschen in spontaner Rede, insbesondere wie Menschen tatsächlich sprechen — nicht wie es die Grammatik vorschreibt. Sein aktueller Fokus liegt auf der Messung und Annotation von Sprechflüssigkeit (etwa Füllpartikeln wie „äh" und „ähm"), der Artikulation von Konsonanten und Vokalen sowie der Frage, wie situative Faktoren wie Gesprächspartner oder Kommunikationskanal die Aussprache beeinflussen. Für Unternehmen und öffentliche Institutionen ist diese Arbeit relevant für die Entwicklung von Spracherkennungssystemen, Sprachsynthese und Dialogsystemen, die natürlichere Interaktion ermöglichen sollen — sowie für die Verbesserung von Sprachunterricht und Aussprachetraining. Seine Methoden kombinieren Korpusanalyse mit instrumenteller Phonetik (elektromagnetische Artikulographie) und standardisierter Annotation, um objektive Messverfahren für bislang schwer fassbare Phänomene der gesprochenen Sprache zu entwickeln.
🔒 Das System hat 379 mögliche Industrie-Partner gefunden — Firmen, Scores und Begründungen sind nur für eingeloggte Nutzer:innen sichtbar. Anmelden
Dr. Malte Belz
HU-FIS-Profil ↗Förderer: DFG Sonderforschungsbereich Zeitraum: 01/2024 - 12/2027 Projektleitung: Prof. Dr. Anke Lüdeling, Dr. Malte Belz, Prof. Dr. Christine Mooshammer
Language Resources and Evaluation · DOI
This paper introduces a multi-layer corpus architecture with multiple tokenizations using the open source historical, diachronic corpus of German called Register in Diachronic German Science. The corpus contains herbal texts printed between the fifteenth and nineteenth centuries and is concerned with the development of a German scientific register, independent of Latin. We will discuss difficulties of transcribing, normalizing and annotating historical texts and will thereby argue for the advantages of multiple layers and multiple tokenizations. A virtually infinite number of annotations can be added to the corpus, without the need for deciding between or discarding interpretations. Thus, this flexible architecture enables multiple normalizations and types of annotation and is open to a wide range of research questions in the humanities. We provide case studies concerning the exploitation of our different normalizations as well as structural, register-specific and linguistic annotations. The corpus architecture allows for its reuse as a resource for corpus-based research approaches.
oder extralinguistische Gerusche wie Lachen, Husten, Ruspern, oder Pfeifen. Abbildung 1.1 stellt die in dieser Arbeit entwickelte und verwendete Formenhierarchie dar.
International Journal of Learner Corpus Research · DOI
Abstract In this article, we explore the disfluencies of advanced learners and native speakers of German in spontaneous speech. We focus on the frequency, form, and place of silent and filled pauses as well as self-repairs. Frequency significantly differs for silent pauses only. As to form, the distribution for both filled pauses and repair types significantly differs between the groups, while the proportion of within-repair hesitations (‘interregna’) is similar. For the neighbouring tokens of filled pauses, learners adhere to the pattern of their native language English, which is significantly different from the pattern we find for native German. Our results indicate that for some aspects of disfluencies, it seems that learners can adapt to a native-like pattern, while others are imported from the L1. Still others are significantly different from both the target and the native pattern. We present different possible explanations for all these cases.