Aster describes what a lecture video shows but never says, so blind and low-vision students can follow it
-
Updated
Aug 14, 2026 - JavaScript
Aster describes what a lecture video shows but never says, so blind and low-vision students can follow it
Detecção de cenas para que deficientes visuais consigam saber tudo que tem ao seu redor e suas respectivas posições no ambiente.
A system for generating music playlists based on the results of audio content analysis. The MusAV dataset is used as a music audio collection, music descriptors are extracted using Essentia, and a simple user interface is used to generate playlists based on these descriptors.
Omni Describer — AI-powered audio description for video, images and PDFs. Accessible descriptions for blind and visually impaired users.
Script and manifest files to create an HLS program containing the Elephants Dream video with captions, subtitles, and audio description
"DANTE-AD: Dual-Vision Attention Network for Long-Term Audio Description" CVPR Workshop AI4CC 2025
Sync fan-made audio description tracks with your video files. Accessibility-first, one command.
One video, two audio tracks, two output devices - so people who need different languages, or audio description, can watch the same screen together.
This Python script replaces the audio in a video file (MP4) with a provided audio file (MP3 or WAV).
Create audio descriptions for videos using ai
Human-in-the-loop AI audio-description authoring for short-form video — model-agnostic (OpenAI/Claude/Gemini/local), React + Flask.
Датасет включает более 200 произведений искусства с аннотациями и профессионально подготовленными тифлокомментариями на русском языке.
Automatyczne tworzenie audiodeskrypcji do filmów z użyciem Google Gemini 2.5 i Google AI TTS
DescribeAT Admin Portal - content management for audio descriptions (React/Vite). Public open-source release.
Generate audio descriptions for videos using AI to make screen content accessible to blind and low-vision viewers.
Demonstration WebVTT tracks that are copyright © to third-party publishers.
Local, explainable film accessibility conformance pre-check engine. IBM AI Builders Challenge July 2026.
Echo Labs is a San Francisco-based company building AI-powered media accessibility for higher education. Its platform audits, captions, and audio describes entire institutional video libraries within 24 hours, producing ADA / Title II and WCAG 2.1 AA compliant closed captions and audio descriptions at claimed accuracy above 99 percent and roughly…
Some scripts used to convert AD scripts to many formats
Add a description, image, and links to the audio-description topic page so that developers can more easily learn about it.
To associate your repository with the audio-description topic, visit your repo's landing page and select "manage topics."