DEPICT: Scoring Text-to-Image Alignment by Answer Agreement Paper • 2610.03617 • Published 10 days ago • 4
Running 227 The ultimate guide to multi-harness RL 🔀 227 Train open models with RL inside real agent harnesses
bottlecapai/ThinkingCap-Qwen3.6-27B Image-Text-to-Text • 27B • Updated 4 days ago • 371 • • 706
Latent Beam Diffusion Models for Generating Visual Sequences Paper • 2503.20429 • Published Sep 20, 2025
Early Estimation of Language to Latent Alignment in Diffusion Models Paper • 2512.08505 • Published Jun 28
AMALIA Technical Report: A Fully Open Source Large Language Model for European Portuguese Paper • 2603.26511 • Published Mar 27
view article Article We’re open-sourcing our text-to-image model and the process behind it Photoroom • Nov 12, 2025 • 102
Contrastive Sequential-Diffusion Learning: Non-linear and Multi-Scene Instructional Video Synthesis Paper • 2407.11814 • Published Dec 6, 2024
Generating Coherent Sequences of Visual Illustrations for Real-World Manual Tasks Paper • 2405.10122 • Published May 16, 2024 • 1
TWIZ-v2: The Wizard of Multimodal Conversational-Stimulus Paper • 2310.02118 • Published Oct 3, 2023