Generate audiobooks from EPUBs, PDFs and text with captions
Image inpainting tool powered by SOTA AI Model
OCR software, free and offline
Faster Whisper transcription with CTranslate2
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Comprehensive Gradio WebUI for audio processing
Use Microsoft Edge's online text-to-speech service from Python
Robust Speech Recognition via Large-Scale Weak Supervision
Open source healthcare AI
A TTS that fits in your CPU (and pocket)
Cut videos with a text editor
PDF to Markdown with vision models
Python library and CLI tool to interface with Google Translate
Contexts Optical Compression
Translate the video from one language to another and embed dubbing
Stable Diffusion web UI
1 min voice data can also be used to train a good TTS model
A modular voice assistant application for experimenting
Unlimited, private and free Speech-To-Text program
AsrTools: Smart Voice-to-Text Tool
EPUB to audiobook converter, optimized for Audiobookshelf
Visual Causal Flow
A theme for Sublime Text 3 by Mattia Astorino
Essential nodes that are weirdly missing from ComfyUI core
Implementation of Imagen, Google's Text-to-Image Neural Network