Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
102 tokens/sec
GPT-4o
59 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
50 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Blind Inverse Problem Solving Made Easy by Text-to-Image Latent Diffusion (2412.00557v1)

Published 30 Nov 2024 in cs.CV, cs.AI, and cs.LG

Abstract: Blind inverse problems, where both the target data and forward operator are unknown, are crucial to many computer vision applications. Existing methods often depend on restrictive assumptions such as additional training, operator linearity, or narrow image distributions, thus limiting their generalizability. In this work, we present LADiBI, a training-free framework that uses large-scale text-to-image diffusion models to solve blind inverse problems with minimal assumptions. By leveraging natural language prompts, LADiBI jointly models priors for both the target image and operator, allowing for flexible adaptation across a variety of tasks. Additionally, we propose a novel posterior sampling approach that combines effective operator initialization with iterative refinement, enabling LADiBI to operate without predefined operator forms. Our experiments show that LADiBI is capable of solving a broad range of image restoration tasks, including both linear and nonlinear problems, on diverse target image distributions.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (6)
  1. Michail Dontas (1 paper)
  2. Yutong He (43 papers)
  3. Naoki Murata (29 papers)
  4. Yuki Mitsufuji (127 papers)
  5. J. Zico Kolter (151 papers)
  6. Ruslan Salakhutdinov (248 papers)