Aster describes what a lecture video shows but never says, so blind and low-vision students can follow it
-
Updated
Aug 14, 2026 - JavaScript
Aster describes what a lecture video shows but never says, so blind and low-vision students can follow it
Detecção de cenas para que deficientes visuais consigam saber tudo que tem ao seu redor e suas respectivas posições no ambiente.
A system for generating music playlists based on the results of audio content analysis. The MusAV dataset is used as a music audio collection, music descriptors are extracted using Essentia, and a simple user interface is used to generate playlists based on these descriptors.
Omni Describer — AI-powered audio description for video, images and PDFs. Accessible descriptions for blind and visually impaired users.
Script and manifest files to create an HLS program containing the Elephants Dream video with captions, subtitles, and audio description
"DANTE-AD: Dual-Vision Attention Network for Long-Term Audio Description" CVPR Workshop AI4CC 2025
Sync fan-made audio description tracks with your video files. Accessibility-first, one command.
Automated Audio Description (AD) sync toolkit with acoustic cross-correlation, DTW alignment, ITU-R BS.775 2.0 downmixing, and commercial excision.
One video, two audio tracks, two output devices - so people who need different languages, or audio description, can watch the same screen together.
Create audio descriptions for videos using ai
Датасет включает более 200 произведений искусства с аннотациями и профессионально подготовленными тифлокомментариями на русском языке.
Automatyczne tworzenie audiodeskrypcji do filmów z użyciem Google Gemini 2.5 i Google AI TTS
This Python script replaces the audio in a video file (MP4) with a provided audio file (MP3 or WAV).
API-first, multi-tenant workflow for human-reviewed video audio description — FastAPI, Next.js, Python workers, TypeScript SDK/CLI, PostgreSQL, S3 and SQS.
Demonstration WebVTT tracks that are copyright © to third-party publishers.
Local, explainable film accessibility conformance pre-check engine. IBM AI Builders Challenge July 2026.
Some scripts used to convert AD scripts to many formats
DescribeAT - accessible audio description app (React/Vite). Public open-source release.
Echo Labs — independent third-party profile of a public API surface, by API Evangelist. Echo Labs is a San Francisco-based company building AI-powered media accessibility for higher education. Its platform audits, captions, and audio describes entire institutional video libraries within 24 hours, producing ADA / Title II and WCAG 2.1 AA compliant c
Add a description, image, and links to the audio-description topic page so that developers can more easily learn about it.
To associate your repository with the audio-description topic, visit your repo's landing page and select "manage topics."