SenseNova U1 dropped an infographic-specific finetune

Reddit r/LocalLLaMA Models

Summary

SenseNova U1 releases an infographic-specific finetune of its U1-8B-MoT base model, achieving significant benchmark improvements in infographic accuracy, chart understanding, and text rendering.

it's the same U1-8B-MoT base with an extended MT (multi-task) training phase focused on structured visual output. the benchmark jumps are significant: IGenBench I-ACC (infographic accuracy) : 4.2πŸ‘‰17.0 (4x) Chart Understanding: 51.3πŸ‘‰69.5Text Rendering: 39.8πŸ‘‰46.6Overall Aesthetic: 53.8πŸ‘‰53.3 Repo: https://github.com/OpenSenseNova/SenseNova-U1github (infographic model docs): https://github.com/OpenSenseNova/SenseNova-U1/blob/main/docs/u1\_infographic\_model.md
Original Article

Similar Articles

SenseNova U1.5 Lite preview just dropped

Reddit r/LocalLLaMA

SenseNova released a preview of its U1.5 Lite model, showing benchmark gains in image generation and editing, with native 4K output and improved Chinese/English text rendering, though acknowledged weaknesses remain.

sensenova/SenseNova-U1.5-8B-MoT

Hugging Face Models Trending

SenseNova-U1.5-8B-MoT is a native unified multimodal model for enhanced visual creation, featuring improvements in image generation quality, text rendering, and precise control.

SenseNova-U1.5: Towards Native Unified Visual Intelligence

Hugging Face Daily Papers

SenseNova-U1.5 is an 8B native unified multimodal model that performs visual understanding, reasoning, and generation without encoders or VAEs, achieving high fidelity and instruction following through patch reconstruction, curated data, and expert optimization.