JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published Jul 26 • 125
DiffGI: Differentiable Geometry Images for High-Fidelity Thin-Shell 3D Generation Paper • 2607.13365 • Published Jul 15 • 22
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Paper • 2607.11562 • Published Jul 13 • 80
Running on CPU Upgrade Agents 2.48k Omni Image Editor 🖼 2.48k Image edit, text to image, image upscale, remove watermark
HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive Image-Text-to-Text • 35B • Updated Apr 17 • 2.07M • 3.56k
Running on Zero Agents Featured 116 Qwen Image Edit Inpaint ✒ 116 inapint with Qwen Image Edit for super precise edits
Running on Zero MCP Featured 411 Qwen Edit Any Pose 🕺 411 Edit any pose with Qwen Edit 2511 Any Pose LoRA
Runtime error Agents 21 Qwen3-TTS Demo 🎙 21 Transform text into natural-sounding speech with custom voices
Running Agents Featured 374 Qwen2.5 Omni 7B Demo 🏆 374 Chat with text, audio, images, and video, get spoken replies
Running Agents 161 Qwen3.5 Omni Offline Demo 🌍 161 Chat with a multimodal AI using text, audio, images or video
Running on Zero Agents Featured 643 Video Background Removal 📽 643 Remove/Change background of video.
Running on Zero Agents Featured 2.66k Qwen Image Multiple Angles 3D Camera 🎥 2.66k Edit image camera angle with custom 3D controls
Running Agents Featured 167 daVinci-MagiHuman 🎬 167 Generate short videos from an image and text prompt