Academic Journal

Effects of a Safety User Interface Bundle on Verification Intentions in Generative AI Chat Use Among Older Chinese Adults: Randomized Vignette Survey.

Λεπτομέρειες βιβλιογραφικής εγγραφής
Τίτλος: Effects of a Safety User Interface Bundle on Verification Intentions in Generative AI Chat Use Among Older Chinese Adults: Randomized Vignette Survey.
Συγγραφείς: Yu J; Faculty of Science, University of Auckland, Auckland, New Zealand., Chen J; Medical Services Management Department, Peking University People's Hospital (PKUPH), Beijing, China., Ren A; Healthcare & Education Research Center, Chengdu Gongyun Education & Management Research Institute, Chengdu, China., Duan H; School of Public Administration and Policy, Renmin University of China, Beijing, China., Meng H; Department of Economics and Management, Sichuan University of Architectural Technology, Chengdu, China., Gao Z; Department of Human Resource Management, Beijing Geriatric Hospital, 118 Wenquan Road, Haidian District, Beijing, 100095, China, 86 15201400966.
Πηγή: Journal of medical Internet research [J Med Internet Res] 2026 Aug 14; Vol. 28, pp. e94140. Date of Electronic Publication: 2026 Aug 14.
Τύπος έκδοσης: Journal Article; Randomized Controlled Trial
Γλώσσα: English
Στοιχεία περιοδικού: Publisher: JMIR Publications Country of Publication: Canada NLM ID: 100959882 Publication Model: Electronic Cited Medium: Internet ISSN: 1438-8871 (Electronic) Linking ISSN: 14388871 NLM ISO Abbreviation: J Med Internet Res Subsets: MEDLINE
Imprint Name(s): Publication: <2011- > : Toronto : JMIR Publications
Original Publication: [Pittsburgh, PA? : s.n., 1999-
Ιατρικοί όροι (MeSH): Generative Artificial Intelligence* , Intention* , Safety* , User-Computer Interface*, Aged ; Female ; Humans ; Male ; Middle Aged ; China ; Cross-Sectional Studies ; Surveys and Questionnaires ; Trust ; East Asian People
Περίληψη: Background: Generative AI chat systems are increasingly used for everyday information seeking, but plausible errors and omissions can mislead users when outputs are accepted without scrutiny. Interface-level safety cues may help users calibrate trust and engage in verification; yet, evidence in older Chinese adults remains limited.
Objective: This study aimed to test whether adding a safety user interface (UI) bundle to a generative AI chat interface increases verification intention among older Chinese adults and to examine selected secondary outcomes, including reliance intention, trust calibration, perceived trustworthiness, comprehension, usability/readability, cognitive load, and a behavioral proxy of verification.
Methods: We conducted a cross-sectional survey with an embedded randomized UI vignette experiment between May 22, 2025, and September 3, 2025. Chinese adults aged ≥60 years were recruited through community sites, outpatient clinic waiting areas, and WeChat (Tencent Holdings Ltd) groups, and randomized 1:1 to view screenshots of a baseline chat UI or a safety UI bundle containing generic source-label cues, and an uncertainty and verification nudge. Each participant completed 2 scenarios (service/travel decision and general well-being related to sleep/fatigue), followed by measures of verification intention (primary), reliance intention, trust calibration index, comprehension (0-8), perceived trustworthiness, usability/readability, cognitive load (0-10), manipulation checks, and a behavioral proxy (expanding optional "source information"). Analyses used intention-to-treat regression models with covariate adjustment.
Results: Of 214 consenting respondents who started the survey, 200 were included in the analysis (100 per arm). The safety UI bundle increased verification intention (mean 4.72, SD 0.63 vs 4.41, SD 0.59 on a 7-point scale; adjusted β=0.293, 95% CI 0.128-0.457; P<.001). Reliance intention did not increase (mean 4.97, SD 0.54 vs 5.03, SD 0.58; adjusted β=-0.105, 95% CI -0.239 to 0.029; P=.13). Trust calibration improved (trust calibration index: mean -0.29, SD 1.43 vs 0.29, SD 1.43; adjusted β=-0.567, 95% CI -1.005 to -0.129; P=.01). Expansion of optional source information was numerically higher, although the adjusted CI included the null (42% vs 27%; adjusted odds ratio [OR]=1.76, 95% CI 0.95-3.27; P=.07). Comprehension remained high and similar across arms (mean 6.33, SD 1.14 vs 6.32, SD 1.08; adjusted β=-0.132, 95% CI -0.428 to 0.163; P=.38). Perceived trustworthiness was modestly lower in the Safety UI arm (mean 5.20, SD 0.61 vs 5.39, SD 0.66; adjusted β=-0.199, 95% CI -0.382 to -0.016; P=.03). Usability/readability was unchanged, and cognitive load did not increase. Manipulation checks indicated higher cue recognition in the Safety UI arm.
Conclusions: In a randomized static-vignette survey of older Chinese adults, a brief safety UI bundle was associated with higher verification intention and a trust calibration index consistent with lower overreliance risk, without detectable reductions in comprehension or usability/readability. Because the intervention was tested as a bundle using screenshots and generic source labels, findings should be interpreted as evidence for a practical interface-level strategy rather than proof that any single cue caused the observed effects.
(© Jun'an Yu, Jun Chen, Anjie Ren, Hui Duan, Hua Meng, Zhuo Gao. Originally published in the Journal of Medical Internet Research (https://www.jmir.org).)
References: Patterns (N Y). 2022 Feb 24;3(4):100455. (PMID: 35465233)
Front Public Health. 2025 Sep 04;13:1637270. (PMID: 40977766)
JMIR Aging. 2023 May 16;6:e44564. (PMID: 37191976)
BMC Psychol. 2024 May 8;12(1):255. (PMID: 38720382)
Front Public Health. 2024 Nov 19;12:1435329. (PMID: 39628811)
Front Psychol. 2024 Apr 17;15:1382693. (PMID: 38694439)
Hum Factors. 2025 Oct;67(10):1062-1083. (PMID: 40104968)
J Med Internet Res. 2023 Jun 14;25:e47184. (PMID: 37314848)
PeerJ Comput Sci. 2025 Mar 27;11:e2773. (PMID: 40567759)
J Am Med Inform Assoc. 2024 Nov 1;31(11):2730-2739. (PMID: 39325508)
Hum Factors. 2024 Jan;66(1):126-144. (PMID: 35344676)
World Wide Web. 2021;24(5):1857-1884. (PMID: 34366701)
Contributed Indexing: Keywords: China; generative AI; human-computer interaction; large language models; older adults; randomized experiment; safety cues; trust calibration; user interface; verification; vignette survey
Entry Date(s): Date Created: 20260814 Date Completed: 20260814 Latest Revision: 20260817
Update Code: 20260817
PubMed Central ID: PMC13475784
DOI: 10.2196/94140
PMID: 42600074
Βάση Δεδομένων: MEDLINE
Περιγραφή
ISSN:1438-8871
DOI:10.2196/94140