@drfeifei: I’m very excited by this new benchmark dataset for visual generation that is suitable for the modern era of large scale…
Summary
Introducing GPIC (Giant Permissive Image Corpus), a large-scale dataset of 100M VLM-captioned image-text pairs for training and 1M pairs for benchmarking, fully permissive for research and commercial use.
View Cached Full Text
Cached at: 05/29/26, 09:56 PM
I’m very excited by this new benchmark dataset for visual generation that is suitable for the modern era of large scale generative models!🤩
Keshigeyan Chandrasegaran (@keshigeyan): 1/ Introducing GPIC: a Giant Permissive Image Corpus and benchmark for visual generation!
🚀100M VLM-captioned image-text pairs for training 📊1M image-text pairs for benchmarking 🖼️~28 trillion pixels 🤗Centrally Hosted ✅Fully permissive for research + commercial use
Dataset,
Similar Articles
@jcjohnss: GPIC should be the new standard benchmark for generative modeling. Training 1 epoch on GPIC is the same cost as 100 epo…
GPIC is a new large-scale image-text dataset and benchmark for generative modeling, claimed to be much more efficient than ImageNet and a better proxy for real-world problems, with fully permissive licensing for research and commercial use.
PixVerve: Advancing Native UHR Image Generation to 100MP with a Large-Scale High-Quality Dataset
This paper introduces PixVerve-95K, a large-scale open-source dataset of 95K ultra-high-resolution (100MP) images with annotations, and PixVerve-Bench, a benchmark for evaluating native 100MP text-to-image generation, extending existing T2I models to unprecedented resolutions.
SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation
SciIR introduces a large-scale dataset (SciIR-82k) and benchmark (SciIR-Bench) to enhance scientific reasoning in text-to-image models, with fine-tuning on Qwen leading to a significant performance improvement.
The new ChatGPT images model is the new standard in photorealistic image generation
OpenAI has released a new ChatGPT image model that sets a new benchmark for photorealistic image generation.
DataComp-VLM: Improved Open Datasets for Vision-Language Models
This paper introduces DataComp-VLM (DCVLM), a comprehensive benchmark for evaluating data curation strategies for vision-language models. The authors find that data mixing, rather than filtering, significantly improves performance, and their resulting DCVLM-Baseline dataset achieves state-of-the-art results on 33 downstream tasks.