Hoe mensen praten en elkaar begrijpen heeft mij altijd gefascineerd. De grote vraag voor mij is hoe we technologie kunnen ontwikkelen die gesproken taal begrijpt: hoe maken we automatische spraakherkenning intelligenter? Naast wat er gezegd wordt, zit er ook heel veel informatie in hoe iets gezegd wordt: aspecten van fysieke, emotionele, en mentale gesteldheid klinken door in de stem, bewust en onbewust. Mijn interesse gaat vooral uit naar het automatisch interpreteren van juist die impliciete informatie met als doel om bijvoorbeeld conversational agents (zoals Siri) gepaster te laten reageren op kinderen of ouderen, of om bijvoorbeeld apps te ontwikkelen die op afstand ondersteuning bieden aan mensen met depressie.

Na mijn studie Taalwetenschap (richting Taal en Spraaktechnologie) in Utrecht, ben ik bij TNO beland en heb ik daar automatische emotieherkenning in spraak onderzocht. Vervolgens ben ik naar Universiteit Twente, Human Media Interaction gegaan waar ik nu nog steeds werk aan de automatische analyse van nonverbale aspecten in spraak communicatie (o.a. lachen, backchanneling) in mens-mens, en mens-machine interactie. Naast het doen van onderzoek geef ik ook onderwijs over speech processing, affective computing, en interaction technology.

Expertises

  • Computer Science

    • Robotics
    • Robot
    • Annotation
    • Speech Recognition
    • Speech Emotion Recognition
    • Convolutional Network
    • Open Source
    • Temporal Feature

Organisaties

Mijn onderzoek richt zich vooral op het automatisch analyseren en interpreteren van nonverbale aspecten in spraak communicatie die iets zeggen over hoe het gesprek gaat, en wat iemands fysieke, socio-emotionele, en mentale gesteldheid is. Mijn doel is automatisch spraakherkenning intelligenter te maken. Ik heb o.a. gewerkt aan automatische detectie van lachen, automatische emotie herkenning in spraak, en het automatisch genereren van backchannels voor artificiele agents. Op dit moment begeleid ik een aantal PhD studenten die onderzoek doen naar multimodale emotie expressie bij ouderen, en responsible design voor kind-robot interactie. Ook begeleid ik master studenten in hun onderzoek naar technologie ten behoeve van kwetsbare mensen (bijvoorbeeld mensen met dementie, mensen met meervoudige beperkingen), en mens-robot interactie.

Je kunt meer lezen over mijn onderzoek hier https://www.utwente.nl/en/research/researchers/featured-scientists/truong/index/ en op mijn persoonlijke website http://khiettruong.space/

Publicaties

2025

The role of voice and appearance in gender perception of speaking robots (2025)In 34th IEEE International Conference on Robot and Human Interactive Communication, RO-MAN 2025 (pp. 178-183) (IEEE International Workshop on Robot and Human Communication (ROMAN); Vol. 2025). IEEE. van Veen, S., Willemse, C., Garcia Goo, H. & Truong, K. P.https://doi.org/10.1109/RO-MAN63969.2025.11217687Do I Sound as Capable as I Look?: Impact of Robot Communication Style and Appearance on User Perception (2025)In Proceedings of the 25th ACM International Conference on Intelligent Virtual Agents (pp. 1-4). Article 44. ACM Press. Kavuza, N. F., Garcia Goo, H. & Truong, K. P.https://doi.org/10.1145/3717511.3749305Remembering past emotions: How emotion expressions are linked to memory reappraisal (2025)PLoS ONE, 20(9). Article e0332575. Nazareth, D. S., Truong, K. P., Heylen, D., Kok, P. & Westerhof, G. J.https://doi.org/10.1371/journal.pone.0332575Enhancing Transcripts of Open-Source Automatic Speech Recognition Models Through Fine-Tuning with Laughter and Speech-Laugh (2025)In Interspeech 2025 (pp. 4513-4517). Ho, P. H., Bălan, D. A., Heylen, D. K. J. & Truong, K. P.https://doi.org/10.21437/Interspeech.2025-2193Capturing the complexity of laughter: Acquisition, annotation and analysis of laughter data in social signal processing (2025)[Thesis › PhD Thesis - Research UT, graduation UT]. University of Twente. Jansen, M.-P.https://doi.org/10.3990/1.9789036566940Age Against the Machine: How Age Relates to Listeners' Ability to Recognize Emotions in Robots' Semantic-Free Utterances (2025)IEEE transactions on affective computing (E-pub ahead of print/First online). Goo, H. G., Ermers, L., Janse, E., Kolkmeier, J., Schadenberg, B., Evers, V. & Truong, K. P.https://doi.org/10.1109/TAFFC.2025.3568595Being Sorry is the Hardest Thing: How Robots can Apologize and Learn from Mistakes to Restore People’s Trust (2025)In CHI EA 2025 - Extended Abstracts of the 2025 CHI Conference on Human Factors in Computing Systems. Article 94. Association for Computing Machinery (ACM). Goo, H. G., Schadenberg, B. R., Kolkmeier, J., Truong, K. P. & Evers, V.https://doi.org/10.1145/3706599.3719739Ask and You Shall Find: How Suggestions by a Conversational Robot Assist Children with Information Search (2025)In Social Robotics - 16th International Conference, ICSR + AI 2024, Proceedings (pp. 417-430) (Lecture Notes in Computer Science; Vol. 15563 LNAI). Springer. Beelen, T., Ordelman, R., Truong, K. P., Evers, V. & Huibers, T.https://doi.org/10.1007/978-981-96-3525-2_35Benchmarking State-of-the-Art Automatic Speech Recognition systems for Dutch (2025)[Contribution to conference › Poster] 3rd Dutch Speech Tech Day 2025. Bălan, D. A., Truong, K. P. & Ordelman, R. J. F.

Onderzoeksprofielen

Verbonden aan opleidingen

Vakken collegejaar 2026/2027

Vakken in het huidig collegejaar worden toegevoegd op het moment dat zij definitief zijn in het Osiris systeem. Daarom kan het zijn dat de lijst nog niet compleet is voor het gehele collegejaar.

Vakken collegejaar 2025/2026

Vakken collegejaar 2024/2025

Lopende projecten

Advancing technology for multimodal analysis of emotion expression in dementia

Multimodal analysis of emotional expression in spoken memories of older adults, lifestory books, reminiscence therapy

Children and AI: talking trust and responsible spoken search

CHATTERS

Responsible design in child-robot-media interaction, spoken interaction between child and conversational agent

4TU Humans & Technology: Smart Social Systems and Spaces for Living Well

Social signal processing and affective computing in speech

Voltooide projecten

EU-FP7 SQUIRREL (Clearing Clutter Bit by Bit)

Robot that helps children tidying up, social signal processing in child-robot interaction

COMMIT P3 SENSEI

Exercise intensity detection through voice, running app

EU-FP7 SSPNet (Social Signal Processing Network)

Automatic analysis of laughter, backchannel generation, speech synchrony

QR codeScan de QR-code of
Download vCard