open-model

Tag

Cards List
#open-model

@pilvar222: We ran DeepSeek v4 Pro 0813 on our cybersecurity benchmark, it outperformed EVERY (!) other model at finding vulnerabil…

X AI KOLs Timeline · 2026-08-13 Cached

DeepSeek v4 Pro 0813 outperforms all other models on a cybersecurity vulnerability-finding benchmark, achieving 87.5% CVE rediscovery at pass@3, though with lower precision and run consistency.

0 favorites 0 likes
#open-model

Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA

Hugging Face Daily Papers · 2026-08-10 Cached

This paper introduces Macaron-V1, an open continual learning agent-model family using Mixture-of-LoRA to compose specialist adapters on frozen base models, with recursive self-improvement and model-harness co-design.

0 favorites 0 likes
#open-model

@UnTalNixon_exe: Forget about GPUs and million-dollar clusters. They just made the world's largest open model (Kimi K3 – 2.78 trillion p…

X AI KOLs Timeline · 2026-08-04 Cached

A new tool called kimi-k3-in-c runs the 2.78T-parameter Kimi K3 open model on a single CPU with as little as 8.24 GB RAM, streaming experts from disk and achieving deterministic output at 10-32 seconds per token.

0 favorites 0 likes
#open-model

NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use

NVIDIA Blog · 2026-08-04 Cached

NVIDIA released Alpamayo 2 Super, an open reasoning model for autonomous vehicles, now available for commercial use under the OpenMDW-1.1 license, delivering frontier-scale reasoning for robotaxis and AV development.

0 favorites 0 likes
#open-model

Teaching an Open Model to Do Science (12 minute read)

TLDR AI · 2026-07-31 Cached

Arcee AI, Loka, AWS, and Prime Intellect post-trained an open model using reinforcement learning to improve scientific tool use and biological reasoning, achieving notable gains on drug tool and Gene Ontology benchmarks.

0 favorites 0 likes
#open-model

On Kimi K3: Its Capabilities And Related Discontents (70 minute read)

TLDR AI · 2026-07-21 Cached

Kimi K3 is a 2.8T parameter open model from Moonshot AI, showing strong benchmark performance but likely over-optimized and lagging behind top closed models by months. It is distilled from Claude and its release may precede an IPO.

0 favorites 0 likes
#open-model

Introducing Cosmos 3 Edge

Hugging Face Blog · 2026-07-20 Cached

NVIDIA released Cosmos 3 Edge, a 4-billion-parameter open world model for edge devices that helps robots and vision AI agents understand surroundings, reason in real time, and generate actions. It achieves best-in-class throughput and accuracy among similar-sized models.

0 favorites 0 likes
#open-model

Google announces Gemma 4 optimized for the Pixel 10's TPU (2 minute read)

TLDR AI · 2026-07-15 Cached

Google announced Gemma 4 E2B optimized for the Pixel 10's TPU, enabling on-device multimodal AI capabilities like offline chat, image recognition, and audio transcription.

0 favorites 0 likes
#open-model

An open model predicting a robot's actions from a control signal. The corner panels are the action and hand pose it was given, everything else is imagined. Is this a world model, or just a video generator?

Reddit r/singularity · 2026-07-12

An open model that predicts a robot's actions from a control signal, raising questions about whether it constitutes a world model or just a video generator.

0 favorites 0 likes
#open-model

@rohanpaul_ai: A paper on open model adoption and finds Chinese models, led by Qwen, now dominate. China passed the U.S. in open model…

X AI KOLs Following · 2026-07-04 Cached

This paper analyzes open language model adoption, finding that Chinese models led by Qwen now dominate downloads, surpassing US models by March 2026. Qwen's lead comes from a diverse range of model sizes, while DeepSeek leads in very large models, and some US models still show strong momentum.

0 favorites 0 likes
#open-model

@sophiamyang: Introducing Leanstral 1.5 A 119B (6B active) open model for formal proof engineering in Lean 4: 100% on miniF2F 587/672…

X AI KOLs Following · 2026-07-03 Cached

Introducing Leanstral 1.5, a 119B parameter (6B active) open model for formal proof engineering in Lean 4, achieving 100% on miniF2F, state-of-the-art scores on PutnamBench and FATE benchmarks, and discovering previously unknown bugs in open-source repositories.

0 favorites 0 likes
#open-model

@lqiao: Cursor . GLM5.2 . Fireworks

X AI KOLs Following · 2026-06-25 Cached

GLM 5.2, an open AI model, is now available in the Cursor coding tool via a partnership with Fireworks.

0 favorites 0 likes
#open-model

NVIDIA has released Nemotron-TwoTower-30B-A3B-Base-BF16, an unusual diffusion-based language model built from the Nemotron 3 Nano 30B-A3B backbone.

Reddit r/LocalLLaMA · 2026-06-25 Cached

NVIDIA released Nemotron-TwoTower-30B-A3B-Base-BF16, a diffusion-based language model that uses block-wise autoregressive diffusion to generate text by iterative denoising of token blocks, achieving 2.42× the generation throughput of the autoregressive baseline while retaining 98.7% of benchmark quality.

0 favorites 0 likes
#open-model

GLM-5.2 Raises the Bar for Open Models (14 minute read)

TLDR AI · 2026-06-23 Cached

GLM-5.2 is a new open-source AI model that sets a high bar for open models, though it still trails proprietary frontier models and lacks some features like vision.

0 favorites 0 likes
#open-model

"How NVIDIA Built Nemotron 3 Open Model" by "Caleb Writes Code" x "Joey Conway"

Reddit r/LocalLLaMA · 2026-06-11 Cached

NVIDIA released the Nemotron 3 open model, offering three sizes: Nano, Super, and Ultra. It optimizes hardware efficiency through architectural innovations such as hybrid Mamba Transformer, latent MoE, and multi-token prediction, and adopts the Open MDW 1.1 open license.

0 favorites 0 likes
#open-model

NVIDIA Accelerates Google DeepMind’s DiffusionGemma for Local AI

NVIDIA Blog · 2026-06-10 Cached

NVIDIA optimizes Google DeepMind's DiffusionGemma, an open model that generates text in parallel 256-token blocks, achieving up to 4x faster performance on local RTX GPUs, DGX Spark, and DGX Station systems.

0 favorites 0 likes
#open-model

DiffusionGemma: 4x Faster Text Generation

Hacker News Top · 2026-06-10 Cached

Google introduces DiffusionGemma, an experimental 26B MoE open model that achieves up to 4x faster text generation on GPUs using text diffusion, targeting speed-critical interactive local workflows.

0 favorites 0 likes
#open-model

@ModelScope2022: Nex-N2 is now open source!An agentic model series from Nex AGI built for coding, tool use, deep research, and long-hori…

X AI KOLs Timeline · 2026-06-08 Cached

Nex AGI releases Nex-N2, an open-source agentic model series for coding, tool use, deep research, and long-horizon workflows, with state-of-the-art benchmarks and Apache 2.0 license.

0 favorites 0 likes
#open-model

NVIDIA just released a 32B open reasoning model for robotaxis

Reddit r/artificial · 2026-06-01

NVIDIA announces Alpamayo 2 Super, a 32B open reasoning model for Level 4 robotaxis, featuring 360-degree perception, meta-actions, and a full stack including AlpaGym simulation and OmniDreams scenario generation.

0 favorites 0 likes
#open-model

Why is nobody talking about Tencent’s Hy3 Preview?

Reddit r/singularity · 2026-06-01

The article observes that Tencent's Hy3 Preview open model performs surprisingly well in evaluations, narrowing the gap with top closed models, yet remains underdiscussed compared to Western AI labs.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback