Tag
SenseTime releases SenseNova-Vision-7B-MoT, a fully open-sourced unified multimodal model that handles multiple vision tasks using natural language instructions, supporting detection, OCR, depth, segmentation, and more.
SenseTime released SenseNova-Vision, a 7B parameter model that unifies computer vision tasks as generation, with open weights, instruction corpus, benchmark, paper, and demo.
SenseNova-U1-8b-MoT-Infographic-V2 is an open-source state-of-the-art model released by SenseTime for infographic design and image editing tasks.
SenseNova U1 is a unified model that handles understanding, reasoning, and generation of text and images in the same architecture, enabling tasks like planning infographics end-to-end.