Get Started
Topic · 26 recaps

Image Generation

Text-to-image synthesis and editing — diffusion, autoregressive, and hybrid approaches, plus controllability techniques like ControlNet, LoRA, and editing pipelines.
PostsQuestions
Home
Topics
Search
Library
Sort
Newest
Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation
Diffusion · Sep 8
WorldSculpt: Generating Compositional Worlds from Grounded Videos
Image Generation · Sep 7
Agentic Visual Generation: From Generative Models to Agentic Control
Agents · Sep 6
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes
Diffusion · Sep 3 · 7:30
Puffin-World: Scaling a Unified Multimodal Model with Native 3D World States
Image Generation · Sep 3
Editable Visual Design
Code Generation · Sep 3
Uncovering Understanding-Generation Synergy in Native Unified Multimodal Models: From Representation, Task to System
Image Generation · Sep 1
A Glance Is All You Need: Single-Pass Fine-Grained Image Captioning with SimLoss
Image Generation · Sep 1
ZipTok3D: High-Fidelity 3D Tokenization with Compact Token Prefixes
Image Generation · Sep 1
On the Design Fundamentals of Pixel Text Representation Learning
Image Generation · Sep 1