Open-weight, large-scale hybrid-attention reasoning model
Project to compile PDFium library to multiple platforms
Phi-3.5 for Mac: Locally-run Vision and Language Models
Python app to work with pictures and associated metadata
State-of-the-art diffusion models for image and audio generation
Visual Automation IDE — automate anything you see on screen
Open-source metadata collector based on ODD Specification
Reference implementation of the Transformer architecture optimized
Clone a voice in 5 seconds to generate arbitrary speech in real-time
Keras Temporal Convolutional Network