MultiCom logo

The Potential Field Multimodal Communication (MultiCom) is a research focus at Goethe University within the university’s Research Profile and the profile area Universality and Diversity.


Description

Communication is inherently multimodal. In spoken languages, visual and other modalities complement speech through manual and facial gestures. Sign languages—languages with a fully grammaticalized communication system in the visual domain—can also use multiple channels of expression simultaneously, including manual signs, facial expressions, body posture, and head movements. Written language often integrates visual elements such as emojis or images, and animal communication also exhibits multimodality, for example when vocal and manual cues are combined.

Within theoretical linguistics, multimodality is a relatively recent field of research that seeks to integrate such phenomena into formal linguistic theory. A central goal is to capture the formal components of visual communication by identifying meaning-bearing elements within the continuous signals of body, head, arm, and hand movements. A key challenge is the treatment of iconic and depictive elements—vocal or visual expressions whose form reflects aspects of their meaning, such as speech-accompanying gestures—which have long been difficult to integrate into existing formal models.

The Potential Field Multimodal Communication aims to strengthen multimodality research at Goethe University and within the Rhine–Main Universities (RMU) by creating spaces for structured exchange across disciplinary boundaries. Through its network of associated researchers, its Short-Term Collaboration Program, and the MultiCom Research Directory, it brings together researchers from different perspectives, supports new collaborations, and increases the visibility of existing research activities and resources.

The Potential Field is deliberately open and exploratory in scope. It seeks to foster dialogue across fields such as linguistics, literary studies, theatre and performance studies, anthropology, film and media studies, cultural studies, history, philosophy, psychology and neurocognition, musicology, computer science, graphic illustration, and related areas.


MultiCom Research Directory

The MultiCom Research Directory provides a growing overview of multimodality-related research at Goethe University and across the Rhine–Main Universities. It brings together research programs, projects, corpora, datasets, and sustained resources, while linking to pages maintained by the respective researchers and project teams.

Explore the research directory →


People

Speakers

Cornelia Ebert
 
Cornelia Ebert
MultiCom interests:
Gestures, co-speech demonstrations. 
 
Frank Kügler
 
MultiCom interests:
Prosody, co-speech gesture, gesture-speech integration, (multimodal) prominence, multimodal encoding of information structure.
 

Associated Researchers

Markus Bader
 
MultiCom interests:
Interplay of linguistic and visual features when describing visual stimuli (pictures, videos).
Kathryn Barnes
 
MultiCom interests:
Depiction in visual and spoken modality, multimodal interaction.
 
MultiCom interests:
Co-speech gestures, gestures in Austronesian languages, stresslessness and gestures.
 
MultiCom interests:
Brain mechanisms underlying multimodal language processing.
Alina Gregori
 
MultiCom interests:
Co-speech gestures, gesture–prosody link, multimodal prominence, temporal alignment, iconicity.
Stefan Hinterwimmer
 
MultiCom interests:
Interaction of speech and gesture, interaction of text and pictorial content, Internet memes.
Andrey Logutov
 
MultiCom interests:
Multimodality in media, multimodal imagination, popular song.
Nikolaus Müller-Schöll
 
MultiCom interests:
Multimodal theatre, gesture, juxtaposition of the elements in script-based theatre.
Simone Pfeifer
 
MultiCom interests:
Multimodal curating, multimodality as process and public engagement, (p)reenactments, memes and appropriation.
Boris Podoroga
 
MultiCom interests:
Visual and audiovisual communication, embodied and machinic perception, cinema and technical media, drone vision, interaction between moving images and bodily affect.
 
MultiCom interests:
Interaction of text and images, multimodal language processing, visual salience and language production.
Anna Pressler
 
MultiCom interests:
Prosody, multimodal interaction.
Manfred Sailer
 
MultiCom interests:
Idioms based on movement or body parts, non-linguistic perlocutionary effects, placeholder expressions (“thingamajig”, “uhm”, …).
 
MultiCom interests:
Multimodal / metamodal representations of print, print-speech interactions and recalibration effects.
 
MultiCom interests:
Expressivity in gesture, gesture semantics, multimodal response strategies.
 

Short-Term Collaboration Program

The Short-Term Collaboration Program is at the heart of the Potential Field. It brings together different scholarly perspectives and creates space for new conversations on multimodal communication. Many of the collaborations grew out of encounters at the MultiCom kick-off workshop, where researchers from a wide range of fields and research traditions entered into dialogue.

We are delighted to support eleven such collaborations. Together, they involve perspectives from linguistics, psychology and cognitive neuroscience, theatre and performance studies, social and cultural anthropology, design anthropology, media and sound studies, philosophy of technology, photography, and visual culture.

The program has fostered both new exchanges across disciplinary boundaries and collaborations between researchers working in different areas of the same discipline. Neither kind of exchange can be taken for granted: bringing distinct theoretical, empirical, artistic, and methodological approaches into conversation can open up new ways of understanding multimodal communication.

Funded Short-Term Collaborations

+ The interplay of linguistic prominence and visual salience in sports reporting
Markus Bader and Yvonne Portele

Live sports reporting provides a unique opportunity to study the joint effect of linguistic features and visual information on the choice of word order and referential expressions. Multimodality differs for speaker and hearer between radio and TV reports because the speaker always relies on visual input whereas only TV reports provide the hearer with access to the visual input. In preparation of a larger research project, we plan a pilot experiment requiring participants to describe picture stories showing sport events to determine how linguistic prominence and visual salience interact during spontaneous language production.

Keywords: Linguistic prominence, visual salience, language production, sports reporting, TV and radio.

+ Limits of linguistic iconicity
Kathryn Barnes and Frank Kügler

This project investigates the limits of linguistic iconicity. Segmental and prosodic properties of ideophones are systematically manipulated, and participants are asked to assess whether the resulting forms allow for an iconic or non-iconic interpretation. The aim is to determine at what point a modification causes the iconic meaning to be lost.

Keywords: Iconicity, ideophones, perception.

+ Prosody–gesture coordination in a language without word-level prominence
Christoph Bracks, Alina Gregori, and Frank Kügler

The central question is how gestures in a language—in this case, Totoli—align with prosodic landmarks when there are no word- or sentence-level prosodic prominence markers such as lexical stress, lexical tone, or pitch accents. Gesture–prosody coordination has primarily been studied in languages with phrasal prominence, where tones and accents have been established as anchor points for gestures. In Totoli, we aim to investigate how gesture–prosody coordination interacts with pragmatic prominence, specifically focus and the givenness of discourse referents.

Keywords: Prosody–gesture coordination, prominence, prosodic and gestural landmarks.

+ Prominence in the production and perception of multimodal artistic performances
Cornelia Ebert, Frank Kügler, Nikolaus Müller-Schöll, and Anna Pressler

This short-term collaboration focuses on transferring a research approach from linguistics to the study of the production and perception of theatre and performance art. It starts from the question of whether the partly conscious and partly unconscious direction of visual and auditory attention in the performing arts can be described and analyzed more precisely using approaches from linguistic research on prominence. To this end, we plan to apply a range of concepts and analytical tools from linguistic prominence research to this new domain.

Keywords: Prominence, multimodal artistic performances.

+ When Gestures Stop Making Sense: Adapting the Semantic Satiation Paradigm to Gesture Processing
Cornelia Ebert, Andrey Logutov, Markus Steinbach, and Patrick Trettenbrein

Semantic satiation refers to the observation that repeated exposure to a meaningful item can temporarily reduce access to its meaning. While this effect has been studied mainly with spoken and written words, the proposed project will, for the first time, adapt the semantic satiation paradigm to gesture comprehension by repeatedly presenting gesture videos and measuring subsequent semantic processing in a classical behavioural paradigm such as, for example a semantic categorization task. By comparing different gesture types, such as emblematic and iconic gestures, the study will provide a first test of whether meaning “fatigue” extends to visual-manual communication and to which kinds of visual expressions. Additionally, our collaboration, bringing together perspectives from linguistics, psychology, as well as from literary and media studies, provides an interdisciplinary resource for future research on repetition and semantic processing.

Keywords: Semantic satiation, gesture studies, language processing, multimodality, psycholinguistics.

+ Remembering what is backgrounded: Extending sentence recall to multimodal communication
Cornelia Ebert, Yvonne Portele, and Sebastian Walter

The proposed project extends the sentence recall paradigm to multimodal communication in order to investigate how gestural information is retained and reproduced. Two pilot studies will evaluate the feasibility of the method and test whether recall behavior is sensitive to theoretically motivated differences between gesture types and gesture positions. The project lays the groundwork for a novel production-based approach to studying the semantics and pragmatics of gesture.

Keywords: Multimodal language production, gesture semantics, psycholinguistics, methodological advancement.

+ Semantic Satiation in Word and Gesture Processing: A Pilot Study
Christian Fiebach, Jack E. Taylor, and Cornelia Ebert

“Semantic satiation” refers to the empirical observation that a large number of repetitions of a word leads to the loss of its meaning in the speaker, reader, or listener. While this effect has been studied with visual and auditory word presentation, it has so far not been examined in a multimodal communication setting. Here, we propose to use EEG for a more direct test of semantic satiation than in previous work, and apply for funding to support a pilot EEG study with visual word presentation to test this novel approach. This work will provide a basis for studying semantic satiation in written, auditory (spoken words), visual (gestures and signs), and multimodal (verbal plus gestural) communication settings.

Keywords: Semantic satiation, word processing, repetition, EEG, machine learning regression, encoding model.

+ Imaginary Multimodal Machines: A Participatory Museum
Andrey Logutov, Boris Podoroga, and Tomáš Dvořák

Imaginary Multimodal Machines: A Participatory Museum will develop a beta online catalogue of existing fictional, speculative, metaphorical, and unrealized devices for producing, organizing, and communicating knowledge. Drawing on literature, cinema, philosophy, scientific speculation, media theory, and technological history, the catalogue will examine how imagined machines combine text, images, numbers, diagrams, sound, gesture, and spatial organization. Invited members of the MultiCom network will be able to propose additional entries through a moderated, token-protected submission form.

Keywords: Multimodal communication, imaginary media, epistemic machines, media archaeology, knowledge infrastructures.

+ Workshop on the Multimodal Appreciation Toolkit
Simone Pfeifer, Andrew Gilbert, and Carla J. Maier

The workshop brings together researchers from MultiCom and authors of the Multimodal Appreciation Toolkit developed at the HU Berlin. The workshop will create a space for valuing and assessing multimodal works within interdisciplinary settings and will lay the groundwork for future collaboration within the network.

Keywords: Evaluation, multimodal appreciation, methodological challenges, immersion.

+ Text–image interactions in visual narratives
Yvonne Portele and Sebastian Walter

The proposed project investigates how visual and textual information interact in the comprehension of visual narratives, employing an interdisciplinary approach that merges formal-semantic analysis with psycholinguistic experimentation. To explore the roles each modality plays and how their combination might facilitate (or hinder) language comprehension, we will conduct a pilot study comparing comprehension across unimodal and multimodal conditions. Results will lay the foundation for a planned eye-tracking study investigating multimodal language processing and comprehension in visual narratives.

Keywords: Visual narratives, multimodal meaning construction, semantics, psycholinguistics.

+ Recalibration Effects in Thai: Using Unique Features of Thai to Understand the Locus of Auditory-Visual Interactions in Word Perception
Jack E. Taylor and Christoph Bracks

Research on European languages using Latin script has demonstrated recalibration effects, in which decoding of linguistic information (e.g., identifying a presented pseudoword as either /aba/ or /ada/) is biased by information from another sensory modality (e.g., vision of written text). Two unusual features of Thai make it an insightful language for studying recalibration phenomena from a new perspective: (a) Thai phonology exhibits a three-way distinction between phonemic categories on a voice-onset-time continuum, where previous research has relied on languages with only a two-way distinction, and (b) this three-way distinction on a phonetic continuum is reflected by corresponding similarity on a graphetic (visual character similarity) continuum. We propose a study adapting the standard recalibration study design to study recalibration effects in Thai, providing a foundation for a new perspective from which to study multimodal interactions in visual and auditory word recognition.

Keywords: Thai, perception, recalibration, phonology, orthography.