Python inference and LoRA trainer package for the LTX-2 audio–video
Official inference repo for FLUX.2 models
Native and Compact Structured Latents for 3D Generation
Genome modeling and design across all domains of life
The repository provides code for running inference with SAM 2
Long-form streaming TTS system for multi-speaker dialogue generation
Project Lyra: Open Generative 3D World Models
PyTorch code and models for VJEPA2 self-supervised learning from video
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Visual Causal Flow
From Images to High-Fidelity 3D Assets
Official Python inference and LoRA trainer package
High-Resolution Image Synthesis with Latent Diffusion Models
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Multi-modal large language model designed for audio understanding
An experimental version of DeepSeek model
Qwen's most powerful open-source image generation model
Qwen2.5-VL is the multimodal large language model series
Chat with your SQL database
A Python toolbox for scalable outlier detection
Infinite Worlds with Versatile Interactions
A Multi-Modal World Model for Reconstructing, Generating, Simulation
An Open-source Framework for Data-centric Language Agents
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System
Blender Model Context Protocol Integration