Aster describes what a lecture video shows but never says, so blind and low-vision students can follow it
-
Updated
Aug 14, 2026 - JavaScript
Aster describes what a lecture video shows but never says, so blind and low-vision students can follow it
Detecção de cenas para que deficientes visuais consigam saber tudo que tem ao seu redor e suas respectivas posições no ambiente.
A system for generating music playlists based on the results of audio content analysis. The MusAV dataset is used as a music audio collection, music descriptors are extracted using Essentia, and a simple user interface is used to generate playlists based on these descriptors.
Script and manifest files to create an HLS program containing the Elephants Dream video with captions, subtitles, and audio description
Omni Describer — AI-powered audio description for video, images and PDFs. Accessible descriptions for blind and visually impaired users.
"DANTE-AD: Dual-Vision Attention Network for Long-Term Audio Description" CVPR Workshop AI4CC 2025
Sync fan-made audio description tracks with your video files. Accessibility-first, one command.
Automated Audio Description (AD) sync toolkit with acoustic cross-correlation, DTW alignment, ITU-R BS.775 2.0 downmixing, and commercial excision.
One video, two audio tracks, two output devices - so people who need different languages, or audio description, can watch the same screen together.
Audio description for Fire TV, generated for films that do not have it — and never spoken over the dialogue
Датасет включает более 200 произведений искусства с аннотациями и профессионально подготовленными тифлокомментариями на русском языке.
This Python script replaces the audio in a video file (MP4) with a provided audio file (MP3 or WAV).
Automatyczne tworzenie audiodeskrypcji do filmów z użyciem Google Gemini 2.5 i Google AI TTS
Create audio descriptions for videos using ai
Audio description and rich captions for video that has none — on Fire TV. Amazon Nova describes, Polly speaks, Fire TV's own audio-track selector delivers. Hackathon entry (Fire TV · AWS Builder · Open Source).
Demonstration WebVTT tracks that are copyright © to third-party publishers.
Local, explainable film accessibility conformance pre-check engine. IBM AI Builders Challenge July 2026.
Multi-agent framework that generates live audio descriptions fitted into dialogue gaps, and measures them against WCAG 2.2 AAA, Section 508, and ADA Title III.
DescribeAT Admin Portal - content management for audio descriptions (React/Vite). Public open-source release.
To associate your repository with the audio-description topic, visit your repo's landing page and select "manage topics."