2,357 tools found
a generative art project by Devi Parikh on BrainDrops.
AIArtists.org showcases leading artists using Artificial Intelligence, tools to make AI Art, and a timeline of AI Art History.
Derrick Schultz's GitHub
artworks dealing with computer vision technologies
synthesize high-quality personalized speech with only a 3-second samples
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams
AI Video Generation Platform [#avatar]
"A multi-voice TTS system trained with an emphasis on quality"
Microsoft's cloud cognitive services
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
a foundational dataset by Meta for research on video learning and multimodal perception
paid service for transcription
an open dataset with 30 trillion tokens for training Large Language Models
An Optimized Speech-to-Text Pipeline for the Whisper Model
Latest Papers and Datasets on Multimodal Large Language Models, and Their Evaluation.
"AI voice generator and realistic text to speech online"
General Corpus of Contemporary Brazilian Portuguese with provenance and typology information - Corpus Geral do Português Brasileiro Contemporâneo
a single API, enabling developers to reason over their spoken data with a few lines of code
Creating a Farming Game in 5 Days. Part 1
accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention