image-generation

Tag

Cards List
#image-generation

SenseNova U1.5 Lite preview just dropped

Reddit r/LocalLLaMA · 2026-07-31

SenseNova released a preview of its U1.5 Lite model, showing benchmark gains in image generation and editing, with native 4K output and improved Chinese/English text rendering, though acknowledged weaknesses remain.

0 favorites 0 likes
#image-generation

@NousResearch: We have opened free FLUX 3 Preview usage to all Nous Portal users, including the free tier. No subscription required to…

X AI KOLs Following · 2026-07-31 Cached

Nous Research has opened free access to FLUX 3 Preview for all Nous Portal users, including free tier, and is running a short film contest with prizes.

0 favorites 0 likes
#image-generation

lodestones/Kroma

Hugging Face Models Trending · 2026-07-31 Cached

Kroma v0.1 is a LoRA fine-tune for Krea 2, released as a ComfyUI-compatible safetensors file with rank 256 and weight deltas for RMSNorm/modulation tensors, under an MIT license.

0 favorites 0 likes
#image-generation

A Frozen Pixel-Space Diffusion Model Can Guide Itself with Its Own Samples

Hugging Face Daily Papers · 2026-07-31 Cached

This paper introduces Synthetic Self-Guidance (SSG), a method that attaches a lightweight prediction head to a frozen pretrained pixel-space diffusion model, using the discrepancy between intermediate and final predictions as self-guidance during sampling. It shows that model-generated samples suffice for training the head, improving FID by over 50% on several variants without classifier-free guidance and enhancing strong baselines with CFG.

0 favorites 0 likes
#image-generation

@LinusEkenstam: Midjourney V8.2 is here.

X AI KOLs Following · 2026-07-30 Cached

Midjourney V8.2 has been released, marking a new version of the popular AI image generation model.

0 favorites 0 likes
#image-generation

@Google: Now you can transform any place with Nano Banana 2’s image generation capabilities right in @GoogleEarth. Visualize wha…

X AI KOLs Timeline · 2026-07-30 Cached

Google integrates Nano Banana 2's image generation into Google Earth, enabling users to visualize historical scenes, reimagine spaces, and brainstorm real estate plans by typing prompts.

0 favorites 0 likes
#image-generation

Parallel Decoding for Video Generation (10 minute read)

TLDR AI · 2026-07-30 Cached

NVIDIA introduces Parallel Decoding Distillation (PDD) for accelerating image and video generation, enabling high-quality outputs with fewer neural function evaluations on models like LTX-2.3 and Wan2.1-14B.

0 favorites 0 likes
#image-generation

@gudanglifehack: Microsoft unveiled two new in-house AI models MAI Image 2.5 Pro delivers a significant boost in image generation qualit…

X AI KOLs Timeline · 2026-07-28 Cached

Microsoft unveiled MAI Image 2.5 Pro, a new in-house AI model with improved image generation quality, complex prompt handling, text rendering, and natural language editing.

0 favorites 0 likes
#image-generation

Parallel Decoding Distillation for Fast Image and Video Generation

Hugging Face Daily Papers · 2026-07-28 Cached

Parallel Decoding Distillation (PDD) is a trajectory-based distillation method that accelerates image and video generation by predicting multiple denoising steps per network evaluation, achieving state-of-the-art performance with 4-8 NFEs on models like LTX-2.3, Wan14B, and Qwen-Image.

0 favorites 0 likes
#image-generation

Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation

Hugging Face Daily Papers · 2026-07-27 Cached

This paper identifies a failure mode in classifier-free guidance distillation called Negative Branch Asymmetry, where errors in the positive and negative CFG branches cancel out, and proposes Positive-Direction Matching to supervise branches separately for more robust distilled models.

0 favorites 0 likes
#image-generation

Comfy-Org/Mage-Flow

Hugging Face Models Trending · 2026-07-24 Cached

Repackaged model files for ComfyUI from the Microsoft Mage-Flow model, including diffusion models, text encoder, and VAE.

0 favorites 0 likes
#image-generation

Spectral Prior for Reducing Exposure Bias in Diffusion Models

Hugging Face Daily Papers · 2026-07-24 Cached

This paper proposes Spectral Alignment (SPA), a lightweight guidance-based method that reduces exposure bias in diffusion models by calibrating the power spectrum of intermediate predictions, showing consistent improvements across pixel-space, latent, and flow-matching models with minimal computational overhead.

0 favorites 0 likes
#image-generation

Microsoft's New MAI-Image and MAI-Voice (2 minute read)

TLDR AI · 2026-07-24 Cached

Microsoft announces public preview of MAI-Image-2.5-Pro and MAI-Voice-2-Flash, their latest purpose-built generative AI models for image and voice, now available on Azure AI and powering Microsoft products like Bing Image Creator.

0 favorites 0 likes
#image-generation

@DanKornas: Visual AI workflows get hard to manage when every model and parameter lives behind a different script. ComfyUI is a nod…

X AI KOLs Timeline · 2026-07-23 Cached

ComfyUI is an open-source, node-based AI creation engine that lets builders and visual professionals design complex generation workflows for image, video, audio, and 3D without coding, with partial re-execution and broad model support.

0 favorites 0 likes
#image-generation

Black Forest Lab's Flux 3: Omni-modality for image, video, audio & action prediction

Reddit r/singularity · 2026-07-23

Black Forest Lab's Flux 3 is a new omni-modal AI model capable of generating and predicting images, video, audio, and actions.

0 favorites 0 likes
#image-generation

BFL Introduces FLUX 3 - multi-modal model for Image, Video and Audio

Reddit r/ArtificialInteligence · 2026-07-23

BFL has introduced FLUX 3, a multi-modal AI model capable of generating images, videos, and audio.

0 favorites 0 likes
#image-generation

@MrLarus: GPT-Image2 turns 'negative space' into a set of Eastern art posters—sophisticated and restrained! The focus is on the main object, embossing, and spatial layers, making the composition clean from afar yet detailed up close. I tried four themes: 1. Beyond the Mountains — ore fissures and distant peaks 2. Where the Wind Passes — double-layer papercraft and bamboo shadows 3. Returning Lantern — lacquer lantern…

X AI KOLs Timeline · 2026-07-23 Cached

User MrLarus created two series of art posters using GPT-Image2—Eastern negative space style and instrumental echo style—emphasizing their sophistication and detailed craftsmanship.

0 favorites 0 likes
#image-generation

Oxygen-TryOn: Fashion-Native Foundation Model for Any-item Virtual Try-On

Hugging Face Daily Papers · 2026-07-23 Cached

Oxygen-TryOn is a unified foundation model for any-item virtual try-on, achieving state-of-the-art consistency and realism across single and multi-item try-on tasks through a dedicated data engine and three-stage training pipeline.

0 favorites 0 likes
#image-generation

Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text

Hugging Face Daily Papers · 2026-07-23 Cached

Proposes ProVisE, a benchmark-agnostic framework to evaluate spatial cognition in image-generation models using pixel-space outputs, and introduces SpatialGen-Bench for unified evaluation across 14 spatial subtasks.

0 favorites 0 likes
#image-generation

Are AI labs pelicanmaxxing?

Simon Willison's Blog · 2026-07-22 Cached

Dylan Castillo conducted a rigorous investigation to determine if AI labs have been secretly training models to draw pelicans riding bicycles. Testing multiple models with various animal-vehicle combinations, he found no evidence of 'pelicanmaxxing'.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback