image-editing

Tag

Cards List
#image-editing

@elonmusk: Grok Imagine image editing is greatly improved

X AI KOLs Timeline · 17h ago Cached

Elon Musk highlights that Grok Imagine's image editing capabilities have been greatly improved, with precise segment editing tools that let users target specific parts of an image instead of regenerating the whole thing.

0 favorites 0 likes
#image-editing

@jefffhj: Try Imagine Image 2.0!

X AI KOLs Following · yesterday Cached

Grok announces Imagine Image 2.0, a next-generation image model with precision editing, crisp text rendering, and improved factuality for real-world use.

0 favorites 0 likes
#image-editing

@heyshrutimishra: - Ultra-wide panoramic scrolls. - Product posters. - Commercial photography. SenseNova U1.5-Lite-Preview generates all …

X AI KOLs Following · 2d ago Cached

SenseNova U1.5-Lite-Preview is an open-source 8B MoT multimodal model that natively generates and edits ultra-wide panoramas, product posters, and commercial photography at 4K, with improved material rendering and fewer artifacts.

0 favorites 0 likes
#image-editing

Perspec 1.0

Hacker News Top · 4d ago Cached

Adrian Sieber announces the 1.0 release of Perspec, a desktop app for correcting the perspective of images, useful for photos of documents and receipts.

0 favorites 0 likes
#image-editing

UniWorld-Design: From Pixel Generation to Layer-Native Design

Hugging Face Daily Papers · 5d ago Cached

UniWorld-Design is a framework that redefines image generation using semantic RGBA layers as atomic units, comprising T2RGBA for generating layered assets and I2L for decomposing images into editable layers, achieving state-of-the-art results on the Crello benchmark.

0 favorites 0 likes
#image-editing

Evaluation-Verification Reward for Consistent Multi-Reference Image Editing

Hugging Face Daily Papers · 2026-07-31 Cached

This paper introduces a Multi-dimensional Evaluation-Verification Reward (EVR) for reinforcement learning fine-tuning of multi-reference image editing models, improving visual consistency and harmony.

0 favorites 0 likes
#image-editing

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

Hugging Face Daily Papers · 2026-07-30 Cached

Introduces MPIE-Bench, a 2,500-sample benchmark for multi-person interaction image editing, along with MPIE-Eval, a mesh-based evaluation method that tracks human judgment more closely than VLM checklists across ten editors.

0 favorites 0 likes
#image-editing

Mage (GitHub Repo)

TLDR AI · 2026-07-22 Cached

Microsoft releases Mage, a family of lightweight 4B-parameter multimodal models for visual understanding and generation, including Mage-VL for image/video understanding and Mage-Flow for text-to-image generation and editing, designed for research and deployment on modest hardware.

0 favorites 0 likes
#image-editing

microsoft/Mage-Flow-Edit-Turbo

Hugging Face Models Trending · 2026-07-21 Cached

Microsoft releases Mage-Flow-Edit-Turbo, a compact 4B-scale generative model for efficient text-to-image generation and instruction-based image editing, achieving state-of-the-art competitive quality through co-designed tokenizer and backbone.

0 favorites 0 likes
#image-editing

microsoft/Mage-Flow

Hugging Face Models Trending · 2026-07-21 Cached

Microsoft releases Mage-Flow, a compact 4B-parameter foundation model for efficient native-resolution text-to-image generation and instruction-based image editing, achieving competitive quality against much larger models.

0 favorites 0 likes
#image-editing

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing

Hugging Face Daily Papers · 2026-07-21 Cached

Mage-Flow is a compact 4B-parameter generative stack for efficient text-to-image generation and instruction-based image editing, featuring a co-designed lightweight tokenizer (Mage-VAE) and a native-resolution multimodal diffusion transformer trained with rectified flow matching. It achieves competitive performance while enabling high-resolution generation at 0.59s on a single A100 GPU.

0 favorites 0 likes
#image-editing

@hao520: I've been using GIMP for over twenty years. Even though smartphones nowadays can accomplish many image processing tasks that previously could only be done on a computer (and even do them better), there are still many things I'm used to doing with GIMP. Figuring out how to break down a goal into tasks on different layers, dividing and conquering — it's still a lot of fun.

X AI KOLs Following · 2026-07-16 Cached

GIMP 3.0, its first major release in seven years, fixes long-standing issues like the confusing floating selection mechanism and introduces non-destructive editing for most GEGL-based effects, significantly improving the user experience.

0 favorites 0 likes
#image-editing

Mojave Paint

Product Hunt · 2026-07-14

Mojave Paint is a tool for direct manipulation of static images on the Mac platform.

0 favorites 0 likes
#image-editing

Boogu-Image-0.1: Boosting Open-Source Unified Multimodal Understanding and Generation

Hugging Face Daily Papers · 2026-07-14 Cached

Boogu-Image-0.1 is an open-source family of unified multimodal understanding and generation models that achieves competitive performance in text-to-image generation, fast inference, instruction-based editing, and bilingual text rendering, with low training cost of approximately $400K.

0 favorites 0 likes
#image-editing

Let RGB Be the Language of Vision

Hugging Face Daily Papers · 2026-07-14 Cached

This paper introduces RINO (RGB In and RGB Out), a unified framework that represents diverse visual information (masks, depth, etc.) as RGB images and converts visual tasks into RGB-to-RGB image editing, enabling a single model to perform zero-shot transfer across tasks.

0 favorites 0 likes
#image-editing

@LinusEkenstam: Seedream 5.0 Pro allows for hyper controllable edits. 1. Use Seedream 5.0 Pro (2k) 2. Upload a reference photo 3. Paste…

X AI KOLs Following · 2026-07-12 Cached

Seedream 5.0 Pro is a highly controllable AI image editing model, accessible via BytePlusGlobal API and Lumina platform, enabling precise edits for humans, items, and animals.

0 favorites 0 likes
#image-editing

CtrlVTON: Controllable Virtual Try-On via Visual-Instance-Prompt Segmentation

Hugging Face Daily Papers · 2026-07-10 Cached

This paper introduces VIP-SAM for instance-level garment segmentation and CtrlVTON, a controllable virtual try-on framework that treats try-on as an image editing problem, allowing precise control over garment layout, style, and placement. Both methods achieve state-of-the-art results on their respective tasks.

0 favorites 0 likes
#image-editing

@xingbugengming: I don't have the GPT 5.6-Sol model, but I spent an hour using up five hours' worth of quota to test the frontend capabilities of GPT 5.6-Terra. The results were very impressive, especially when handling image matting—it's amazing. I asked it to create a webpage: an interactive card of Guan Yu from the Three Kingdoms, with mouse drag...

X AI KOLs Timeline · 2026-07-10 Cached

The user tested the frontend capabilities of GPT 5.6-Terra, creating an interactive Guan Yu card webpage from the Three Kingdoms. The results were outstanding, particularly in image matting and depth-of-field perspective.

0 favorites 0 likes
#image-editing

conradlocke/krea2-identity-edit

Hugging Face Models Trending · 2026-07-07 Cached

A community fine-tune of Krea 2 Raw that enables instruction-based, identity-preserving image editing. It edits images while preserving details like faces, using a custom ComfyUI node pack.

0 favorites 0 likes
#image-editing

@AdinaYakup: Boogu-Image-0.1 New unified image generation + editing model family - 10B Base/Edit/Turbo - Apache 2.0 - Fast Turbo inf…

X AI KOLs Following · 2026-07-03 Cached

Boogu-Image-0.1 is a new unified image generation and editing model family with 10B parameters, available under Apache 2.0 license. It features fast turbo inference in 4 steps, trained on 10x less data, and supports Chinese and English.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback