InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency Paper • 2508.18265 • Published Aug 25, 2025 • 222
Running Agents 539 HiDream I1 Dev 🚀 539 Generate images from text prompts with customizable aspect ratio
Running on Zero Agents Featured 2.03k Chat With Janus-Pro-7B 🌍 2.03k A unified multimodal understanding and generation model.
Running on Zero Agents Featured 1.8k Joy Caption Alpha Two 👁 1.8k Generate detailed captions or prompts for any image
Generating Fine-Grained Human Motions Using ChatGPT-Refined Descriptions Paper • 2312.02772 • Published Dec 5, 2023 • 7