Get Started
Home
Topics
Search
Library
Topic · 15 recaps

Image Generation

Text-to-image synthesis and editing — diffusion, autoregressive, and hybrid approaches, plus controllability techniques like ControlNet, LoRA, and editing pipelines.
Sort
Newest
Video Generative Models as Geometry Learner
Diffusion · Aug 28
Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision
Image Generation · Aug 17
ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation
Agents · Aug 5
Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing
Diffusion · Aug 3
Meshy T2: Fast Native Mesh Generation with Flow Matching
Diffusion · Jul 28
Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text
Evaluation · Jul 23
Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers
Image Generation · Jul 21
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing
Diffusion · Jul 21
OvisOCR2 Technical Report
Image Generation · Jul 15
Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation
Image Generation · Jul 13