GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration
Summary
This paper introduces GGT-100K, a dataset of 103,707 image pairs for real-world image restoration, generated by using multimodal foundation models like Nano-Banana-2 to produce high-quality targets from low-quality inputs. Experiments show the dataset improves the generalization of various image restoration models.
View Cached Full Text
Cached at: 06/01/26, 03:17 AM
Paper page - GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration
Source: https://huggingface.co/papers/2605.31039
Abstract
Generative multimodal foundation models are used to create high-quality training data for image restoration, improving model generalization across diverse real-world scenarios.
Real-worldimage restoration(IR) is bottlenecked by the scarcity of high-quality paired training data.Synthetic datasetsare abundant but often fail to modelreal-world degradations, while real-world paired datasets are expensive and difficult to capture. As a result, IR models trained on these datasets show limited generalization in real-world scenarios. In this work, we proposeGenerative Ground Truth(GGT) by usinggenerative multimodal foundation models(MFMs) to produce high-quality (HQ) targets from real-world low-quality (LQ) images. We first conduct a systematic evaluation of nine state-of-the-art MFMs, includingNano-Banana-2and GPT-Image-2, on images of various scenes and degradation types. The results demonstrate thatNano-Banana-2withVLM-based adaptive promptingshows the highest capability to synthesize perceptually realistic and content-faithful HQ targets, which can serve as the GGT for the LQ input. We then employNano-Banana-2to build a GGT synthesis pipeline, which involvesmulti-stage quality controlto ensure data reliability, and construct GGT-100K, anLQ-HQ paired datasetcomprising 103,707 training pairs and covering diverse scenes and complexreal-world degradations. A test set of 500 image pairs is also established. Extensive experiments show that GGT-100K consistently improves the real-world generalization of a wide range of IR models, with particularly strong benefits for finetuning generative models for IR tasks. Our results suggest that MFMs can serve as practical tools for restoration-oriented data generation, and GGT-100K is a useful resource to expand the generalization boundaries of real-world IR models.
View arXiv pageView PDFProject pageGitHub9Add to collection
Get this paper in your agent:
hf papers read 2605\.31039
Don’t have the latest CLI?curl \-LsSf https://hf\.co/cli/install\.sh \| bash
Models citing this paper0
No model linking this paper
Cite arxiv.org/abs/2605.31039 in a model README.md to link it from this page.
Datasets citing this paper1
#### VCLab-PolyU/GGT-100K Updatedabout 2 hours ago • 98
Spaces citing this paper0
No Space linking this paper
Cite arxiv.org/abs/2605.31039 in a Space README.md to link it from this page.
Collections including this paper0
No Collection including this paper
Add this paper to acollectionto link it from this page.
Similar Articles
Our agent found a clean rule across 120 CIKs, published it to four repos, and it was false at 494 companies
An LLM-assisted agent published a false rule about SEC 8-K filing timestamps after confirming it only against the 120-company sample that generated it; the author retracted the claim and argues that agent outputs must be validated against data outside the hypothesis-generation loop.
@johnschulman2: Bullish on this direction. Having a metric for explanation quality makes it possible to hillclimb, and counterfactual s…
John Schulman highlights research by Adam Karvonen and colleagues on using counterfactual simulatability as a metric to improve AI explanation quality. They developed a dataset and pipeline that trains models to generate better post-hoc explanations of their own behavior, showing generalization to held-out evaluations.
A Large Open Multi-Energy Corpus of Soil Compaction Tests, with Machine-Learning Baselines
This paper introduces a large open corpus of soil compaction tests and establishes machine-learning baselines for predicting compaction parameters, emphasizing physics-constrained models for practical screening.
PACE: Towards Surfacing Hidden Conflicts in User Requests
The paper introduces PACE, a dataset for evaluating whether AI models can identify hidden conflicts in user requests by retrieving implicit knowledge base facts, and proposes PaceMaker, a multi-agent framework to enhance conflict-aware decision-making.
Large Language Models in Resolving Contextual Knowledge Conflicts
This paper introduces a taxonomy and dataset for contextual knowledge conflicts in large language models, experiments with seven LLMs, and proposes a steering method to improve conflict resolution in reasoning and summarization tasks.