swiss-ai/Apertus-v1.5 70B/8B

Reddit r/LocalLLaMA Models

Summary

Swiss AI releases Apertus 1.5, a family of fully open 8B and 70B multilingual multimodal language models supporting up to 262k context length, audio/image understanding, reasoning mode, and improved instruction-following and tool use.

https://huggingface.co/swiss-ai/Apertus-v1.5-70B https://huggingface.co/swiss-ai/Apertus-v1.5-8B Apertus 1.5 is a family of 8B and 70B parameter language models designed to advance the state of multilingual, multimodal, fully open, and transparent AI. The models support a wide range of languages, handle contexts of up to 262,144 tokens, and it uses only fully open training data whilst delivering performance comparable to other models of similar size. The released models are the result of continued pretraining of Apertus 1.0, adding a multimodal mix of 4T tokens to the 8B model and 2T tokens to the 70B model. Apertus 1.5 thus uses the same architecture as the original release, a decoder-only transformer with the xIELU activation function trained with the AdEMAMix optimizer. Our improved post-training recipe enhances the models' instruction-following and tool-use capabilities and, for the first time, allows developers to enable a thinking mode to improve the models' performance on reasoning tasks. As a first in the Apertus family, the Apertus 1.5 models support multimodal inputs. The model takes images, audio, and text as input and generates text. This enables many new exciting use cases for our developers. Key Features Fully Open Model: Open weights + open data + full training details including all data and training recipes. Massively Multilingual: Supporting a large variety of languages. Responsible Development: Apertus is trained while respecting opt-out consent of data owners (even retroactively) where possible and with methods to prevent memorization of training data. Native Audio & Image Understanding: Apertus 1.5 introduces multimodal support for processing audio and image inputs, enabling more intuitive and versatile interaction beyond text. Reasoning: The models can be switched to thinking mode to reason on the input before generating responses. Long Context: Apertus 1.5 by default supports a context length up to 262,144 tokens, a four-fold increase from our initial Apertus 1.0 release. Improved Instruction-Following: Significant improvements in instruction adherence ensure more predictable and accurate responses to user prompts. Improved Tool Use: Apertus 1.5 has been trained for better tool integration, allowing for more effective use of external tools and APIs. The technical report with further details along with benchmark results, training pipelines, and intermediate checkpoints will be published in the coming weeks.
Original Article

Similar Articles

Apertus – Open Foundation Model for Sovereign AI

Hacker News Top

Apertus is a fully open foundation model for sovereign AI, developed by the Swiss AI Initiative. It is open weights, open data, open science, compliant with EU AI Act, and competitive with top open models at 8B and 70B parameters, supporting over 1000 languages.

AIDC-AI/Ovis2.6-80B-A3B · Hugging Face

Reddit r/LocalLLaMA

Ovis2.6-80B-A3B is a new Multimodal Large Language Model released by AIDC-AI, featuring a Mixture-of-Experts architecture with 80B total parameters but only 3B active during inference. It offers enhanced long-context processing, high-resolution understanding, and active visual reasoning capabilities.

CohereLabs/command-a-plus-05-2026-w4a4

Hugging Face Models Trending

CohereLabs releases Command A+, an open-source 25B active parameter model optimized for agentic, multilingual, and reasoning tasks, with vision support and Apache 2.0 license.