DeepSeek-OCR
π
16
Extract text from images and convert to markdown
ultra-fast video model, LTX 0.9.8 13B distilled
A unified multimodal understanding and generation model.
Edit images with sketches, colors, and text prompts
text-to-3D & image-to-3D
Engage in multimedia chat with LLMs and ML models
Generate audio from text descriptions