Nöth / Horák / Sojka | Text, Speech, and Dialogue | Buch | 978-3-032-37248-2 | www.sack.de

Buch, Englisch, Format (B × H): 155 mm x 235 mm

Reihe: Lecture Notes in Artificial Intelligence

Nöth / Horák / Sojka

Text, Speech, and Dialogue

29th International Conference, TSD 2026, Brno, Czech Republic, September 1–4, 2026, Proceedings
Erscheinungsjahr 2026
ISBN: 978-3-032-37248-2
Verlag: Springer

29th International Conference, TSD 2026, Brno, Czech Republic, September 1–4, 2026, Proceedings

Buch, Englisch, Format (B × H): 155 mm x 235 mm

Reihe: Lecture Notes in Artificial Intelligence

ISBN: 978-3-032-37248-2
Verlag: Springer


    This conference volume constitutes the proceedings of the 29th International Conference on Text, Speech, and Dialogue, TSD 2026, held in Brno, Czech Republic, during September 1–4, 2026.

    The 54 full papers included in this volume were carefully reviewed and selected from 111 submissions. The papers focus on Large Language Models, Speech Recognition, Speech Disorder Analysis, Retrieval Augmented Generation, or Response Quality Assessment.

Nöth / Horák / Sojka Text, Speech, and Dialogue jetzt bestellen!

Zielgruppe


Research

Weitere Infos & Material


.- Text.

.- Using Explainable AI to Identify Spanish SDOH Keywords.
.- SYN_r25: a Corpus Linked to a Database of Multi-Word Expressions.
.- Is this Funny to You?: A Reader-Aware Humour Classification Dataset.
.- Context as a Key: Quantifying Verbatim Data Leakage Across Model Scale, Alignment, and Reasoning Architectures.
.- Large Language Models for Sparse Entity Alignment over Knowledge Graphs.
.- Evaluation of Transformer Language Models for Hate Speech Detection in Croatian Online Text.
.- Large Language Models for Norwegian Bokmål-Nynorsk Translation: Scaling Laws and the Limits of Back-Translation.
.- LLMs as Linguistic Experts: The Case of Morphological Segmentation.
.- “Chi nas dal soch el sent de legn” - Auditing Text Corpora for Lombard.
.- Benchmarking Pragmatic Reasoning of Large Language Models in the Czech Language.
.- UzbekSpell: An Annotated Benchmark Corpus for Error Detection and Correction in Uzbek.
.- Size Matters: Foundation Model for Czech HTML Documents.
.- Online Punctuation and Capitalization Restoration Using Seq2Seq Approach.
.- Location-Aware Language Models via Secondary Embeddings.
.- Gender Bias in Greek Pronoun Resolution: Evaluating Multilingual and Language-Specific LLMs.
.- Data Synthesis and Parameter-Efficient Fine-Tuning for Low-Resource NMT: A Case Study on Q’eqchi’ Mayan.
.- Building Data Infrastructure and LLM Benchmarks for the Masurian Dialect.
.- Large-Scale Computational Morphology and Lexical Resources for Latvian and Latgalian.
.- Evaluating the Effect of Fine-Tuning on Usability Annotation of English-Slovak Translation.
.- Pseudonymisation for Morphologically Rich Languages.

.- Speech.

.- SYNODA: Synthetic Data-Oriented Domain Adaptation for Automatic Speech Recognition.
.- OLaPh: Optimal Language Phonemizer.
.- Syllable Stress Detection to Evaluate Pathological Speech and L2 Language Level Pronunciation.
.- Evaluating the Sufficiency of Verbal Picture Description Task for Cognitive Impairment Detection.
.- On the Synthesis of Dysarthric Speech: Evaluating Zero-Shot and Fine-Tuning Approaches.
.- Optimizing Streaming Zipformer for Czech Humanoid Robotics: Balancing Accuracy, Latency, and Ego-Noise.
.- On the Influence of VITS Initialisation on Speech Quality.
.- On the Robustness to Recording Condition of Speaker and Accent Embeddings for a Voice Cloning System.
.- Leveraging Zero-Shot TTS for Data Augmentation: A Comparison of VITS and StyleTTS2 in Low-Resource and Low-Quality Conditions.
.- Cross-Attention Fusion of Acoustic and Linguistic Features for Parkinson’s Disease Detection.
.- SofiaFala ECOA: A Framework for Multimodal Corpora Construction for Speech Disorder Analysis and Assistive AI Systems.
.- Acoustic Volume Mismatch in Distant Streaming ASR: Case Study and Mitigation via Wide-Range Volume Perturbation.
.- Cross-lingual Text-to-speech Translation for Regional, Low-resource Languages in Vietnam.
.- Contextual Biasing for Streaming Automatic Speech Recognition via Phoneme-Fused Representation and Detect-Then-Inject Decoding.
.- LibriConvo: Simulating Conversations from Read Literature for ASR and Diarization.
.- Early Alzheimer’s Detection Using Siamese Conformer Networks on Slovak Confrontation Naming Tasks.
.- Tremor Vision: Cross-Domain Multiclass Plosive Burst Detection from Audio-Derived Spectrogram Images for Parkinsonian DDK Speech Analysis.
.- On the Explainability of Speech Impairment Regarding Dysarthria Progression in Parkinson’s Disease.
.- Benchmarking Pretrained Speech Models for Arabic Dialect Identification with Cross-Corpus Evaluation.
.- Cross-Database Performance of Speech-Based Graph Neural Networks for Parkinson’s Disease Detection.

.- Dialogue.

.- Communicative Function Classification using Voice Activity Projection.
.- GRASS - HRI: A Corpus of Spontaneous Austrian German Dialogues with the Social Robot Furhat.
.- Utterance-Level Spoken Embeddings for Dialogue State Tracking.
.- Turn-by-Turn Acoustic-Prosodic Alignment and its Relationship to Lexical and Semantic Similarity in Casual Conversations.
.- Multimodal Speech Recognition in High-Noise Factory Floors for Human Robot Collaboration.
.- Analysis and Implications of Alignment Between LLM-as-a-Judge and Human Evaluators in a Generative AI Chatbot: A Real-World Study.
.- Modelling Thematic Structure in Clinical Dialogues for Interpretable Depression Detection.
.- Your Retriever Already Knows: Predicting RAG Retrieval Sufficiency from Score Distributions.
.- Reading between the Lines: Leveraging Large Language Models for Global Dementia and Depression Assessment from Clinical Interviews.
.- An Effective Structured Generation Approach to Sentence Stress Detection with Speech Language Models.
.- From Paper to Speech: ASR-Driven Digitalization of the PICNIR Cognitive Test.
.- Blackwell: An Explainable Anamnesis-based Multi-Agent Framework for Remote Clinical Diagnosis.
.- Comparing Acoustic and Linguistic Features for Automated Deception Detection.
.- OnkoRAG: Modelling Cross-Lingual German-Arabic Medical Question Answering.



Ihre Fragen, Wünsche oder Anmerkungen
Vorname*
Nachname*
Ihre E-Mail-Adresse*
Kundennr.
Ihre Nachricht*
Lediglich mit * gekennzeichnete Felder sind Pflichtfelder.
Wenn Sie die im Kontaktformular eingegebenen Daten durch Klick auf den nachfolgenden Button übersenden, erklären Sie sich damit einverstanden, dass wir Ihr Angaben für die Beantwortung Ihrer Anfrage verwenden. Selbstverständlich werden Ihre Daten vertraulich behandelt und nicht an Dritte weitergegeben. Sie können der Verwendung Ihrer Daten jederzeit widersprechen. Das Datenhandling bei Sack Fachmedien erklären wir Ihnen in unserer Datenschutzerklärung.