Modeling human social vision with cinematic stimuli: An integrative approach

dc.contributor.authorSantavirta, Severi
dc.contributor.authorParanko, Birgitta
dc.contributor.authorSeppälä, Kerttu
dc.contributor.authorHyönä, Jukka
dc.contributor.authorNummenmaa, Lauri
dc.contributor.organizationfi=psykiatria|en=Psychiatry|
dc.contributor.organizationfi=tyks, vsshp|en=tyks, varha|
dc.contributor.organizationfi=PET-keskus|en=Turku PET Centre|
dc.contributor.organizationfi=psykologia|en=Psychology|
dc.contributor.organization-code1.2.246.10.2458963.20.14646305228
dc.contributor.organization-code1.2.246.10.2458963.20.16217176722
dc.contributor.organization-code1.2.246.10.2458963.20.15586825505
dc.converis.publication-id526838500
dc.converis.urlhttps://research.utu.fi/converis/portal/Publication/526838500
dc.date.accessioned2026-07-31T20:11:45Z
dc.description.abstract<p>Sociability is central for humans. Visual information ranging from low-level physical features (e.g., luminance) to mid-level semantic information (e.g., face recognition) and high-level social inference (e.g., emotional valence of social interactions) is constantly sampled for navigating the social world. In this study, we utilized large-scale eye tracking during natural vision for mapping how different levels of visual information guide the perception of socially relevant features (social vision) simultaneously. In three experiments, participants (<em>N</em> = 166) watched full-length films and short movie clips with varying social content (total duration: 193 minutes) during eye tracking. To model the association between perceptual features and spatiotemporal gaze parameters (gaze position, gaze synchronization, pupil size and blinking), we extracted 39 stimulus features from the movies, including low-level audiovisual features (e.g., luminance, motion), presence and location of mid-level semantic categories (e.g., faces, objects), and high-level social information (e.g., body movements, pleasantness). Integrative analysis techniques with cross-validation were developed to simultaneously associate the perceptual features with the gaze behavior. Pupil size was modulated by luminance, scene cuts, and emotional arousal while gaze position was most accurately predicted by a combination of the presence of human faces, local motion, and entropy. Faces and eyes were prioritized over other semantic categories, and blinking rate decreased during periods of attentional engagement. Altogether, the results show that human social vision is primarily guided by low-level physical features and mid-level semantic categories, while high-level social features such as emotional arousal primarily modulate pupillary responses.<br></p>
dc.identifier.jour-issn1534-7362
dc.identifier.urihttps://www.utupub.fi/handle/11111/62820
dc.identifier.urlhttps://doi.org/10.1167/jov.26.5.9
dc.identifier.urnURN:NBN:fi-fe20260728112972
dc.language.isoen
dc.okm.affiliatedauthorSantavirta, Severi
dc.okm.affiliatedauthorParanko, Birgitta
dc.okm.affiliatedauthorSeppälä, Kerttu
dc.okm.affiliatedauthorHyönä, Jukka
dc.okm.affiliatedauthorNummenmaa, Lauri
dc.okm.affiliatedauthorDataimport, tyks, vsshp
dc.okm.discipline3126 Surgery, anesthesiology, intensive care, radiologyen_GB
dc.okm.discipline3126 Kirurgia, anestesiologia, tehohoito, radiologiafi_FI
dc.okm.internationalcopublicationnot an international co-publication
dc.okm.internationalityInternational publication
dc.okm.typeA1 ScientificArticle
dc.publisherAssociation for Research in Vision and Ophthalmology (ARVO)
dc.publisher.countryUnited Statesen_GB
dc.publisher.countryYhdysvallat (USA)fi_FI
dc.publisher.country-codeUS
dc.relation.doi10.1167/jov.26.5.9
dc.relation.ispartofjournalJournal of Vision
dc.relation.issue5
dc.relation.volume26
dc.titleModeling human social vision with cinematic stimuli: An integrative approach
dc.year.issued2026

Tiedostot

Näytetään 1 - 1 / 1
Ladataan...
Name:
i1534-7362-26-5-9_1779877921.22324.pdf
Size:
14.07 MB
Format:
Adobe Portable Document Format