• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

相似文献

1
Describe Now: User-Driven Audio Description for Blind and Low Vision Individuals.即时描述:面向盲人和低视力者的用户驱动音频描述
DIS (Des Interact Syst Conf). 2025 Jul;2025:458-474. doi: 10.1145/3715336.3735685. Epub 2025 Jul 4.
2
Prescription of Controlled Substances: Benefits and Risks管制药品的处方:益处与风险
3
VIIDA and InViDe: computational approaches for generating and evaluating inclusive image paragraphs for the visually impaired.VIIDA和InViDe:为视障人士生成和评估包容性图像段落的计算方法。
Disabil Rehabil Assist Technol. 2025 Jul;20(5):1470-1495. doi: 10.1080/17483107.2024.2437567. Epub 2024 Dec 11.
4
VideoA11y: Method and Dataset for Accessible Video Description.VideoA11y:无障碍视频描述的方法与数据集。
Proc SIGCHI Conf Hum Factor Comput Syst. 2025 Apr-May;2025. doi: 10.1145/3706598.3714096. Epub 2025 Apr 25.
5
A Comprehensive and Modality Diverse Cervical Spine and Back Musculoskeletal Physical Exam Curriculum for Medical Students.面向医学生的全面且多模态的颈椎和背部肌肉骨骼物理检查课程
J Educ Teach Emerg Med. 2025 Jul 31;10(3):SG1-SG8. doi: 10.21980/J8RQ0N. eCollection 2025 Jul.
6
How lived experiences of illness trajectories, burdens of treatment, and social inequalities shape service user and caregiver participation in health and social care: a theory-informed qualitative evidence synthesis.疾病轨迹的生活经历、治疗负担和社会不平等如何影响服务使用者和照顾者参与健康和社会护理:一项基于理论的定性证据综合分析
Health Soc Care Deliv Res. 2025 Jun;13(24):1-120. doi: 10.3310/HGTQ8159.
7
Stigma Management Strategies of Autistic Social Media Users.自闭症社交媒体用户的污名管理策略
Autism Adulthood. 2025 May 28;7(3):273-282. doi: 10.1089/aut.2023.0095. eCollection 2025 Jun.
8
Exploring the use of smartphone applications during navigation-based tasks for individuals who are blind or who have low vision: future directions and priorities.探索为盲人或视力低下者在基于导航的任务中使用智能手机应用程序:未来方向与优先事项。
Disabil Rehabil Assist Technol. 2025 Aug 25:1-29. doi: 10.1080/17483107.2025.2544942.
9
Development of a personalized conversational health agent to enhance physical activity for blind and low-vision individuals.开发个性化对话式健康代理以增强盲人和低视力个体的身体活动。
Mhealth. 2025 Jul 10;11:29. doi: 10.21037/mhealth-24-60. eCollection 2025.
10
Aspects of Genetic Diversity, Host Specificity and Public Health Significance of Single-Celled Intestinal Parasites Commonly Observed in Humans and Mostly Referred to as 'Non-Pathogenic'.人类常见且大多被称为“非致病性”的单细胞肠道寄生虫的遗传多样性、宿主特异性及公共卫生意义
APMIS. 2025 Sep;133(9):e70036. doi: 10.1111/apm.70036.

本文引用的文献

1
VideoA11y: Method and Dataset for Accessible Video Description.VideoA11y:无障碍视频描述的方法与数据集。
Proc SIGCHI Conf Hum Factor Comput Syst. 2025 Apr-May;2025. doi: 10.1145/3706598.3714096. Epub 2025 Apr 25.
2
Supporting Novices Author Audio Descriptions via Automatic Feedback.通过自动反馈支持新手作者音频描述
Proc SIGCHI Conf Hum Factor Comput Syst. 2023;2023:1-18. doi: 10.1145/3544548.3581023.
3
You Described, We Archived: A Rich Audio Description Dataset.你描述,我们存档:一个丰富的音频描述数据集。
J Technol Pers Disabil. 2023 May;11:192-208. Epub 2024 Jan 19.

即时描述:面向盲人和低视力者的用户驱动音频描述

Describe Now: User-Driven Audio Description for Blind and Low Vision Individuals.

作者信息

Cheema Maryam, Seifi Hasti, Fazli Pooyan

机构信息

Arizona State University Tempe, Arizona, USA.

出版信息

DIS (Des Interact Syst Conf). 2025 Jul;2025:458-474. doi: 10.1145/3715336.3735685. Epub 2025 Jul 4.

DOI:10.1145/3715336.3735685
PMID:40918301
原文链接:https://pmc.ncbi.nlm.nih.gov/articles/PMC12413206/
Abstract

Audio descriptions (AD) make videos accessible for blind and low vision (BLV) users by describing visual elements that cannot be understood from the main audio track. AD created by professionals or novice describers is time-consuming and offers little customization or control to BLV viewers on description length and content and when they receive it. To address this gap, we explore user-driven AI-generated descriptions, enabling BLV viewers to control both the timing and level of detail of the descriptions they receive. In a study, 20 BLV participants activated audio descriptions for seven different video genres with two levels of detail: concise and detailed. Our findings reveal differences in the preferred frequency and level of detail of ADs for different videos, participants' sense of control with this style of AD delivery, and its limitations. We discuss the implications of these findings for the development of future AD tools for BLV users.

摘要

音频描述(AD)通过描述主音频轨道中无法理解的视觉元素,使盲人和低视力(BLV)用户能够访问视频。由专业人员或新手描述者创建的音频描述耗时且在描述长度、内容以及BLV观众接收描述的时间方面几乎没有提供定制或控制权。为了弥补这一差距,我们探索了用户驱动的人工智能生成描述,使BLV观众能够控制他们收到的描述的时间和细节程度。在一项研究中,20名BLV参与者为七种不同视频类型激活了音频描述,有两种细节级别:简洁和详细。我们的研究结果揭示了不同视频的音频描述在首选频率和细节级别、参与者对这种音频描述传递方式的控制感以及其局限性方面的差异。我们讨论了这些研究结果对未来为BLV用户开发音频描述工具的意义。