MOOSE-Star (ICML 2026): 7B model + 108K-paper dataset for scientific hypothesis discovery
Summary
MOOSE-Star presents a 7B model fine-tuned from DeepSeek-R1-Distill-Qwen-7B for scientific hypothesis discovery, along with a dataset of 108K NCBI papers. The model achieves state-of-the-art inspiration retrieval accuracy, outperforming larger models like GPT-5.4 and Gemini-3 Pro.
Similar Articles
@ModelScope2022: SciJudge-30B and 4B learn to predict which scientific work will carry stronger citation impact. License: Apache-2.0 30B…
ModelScope releases SciJudge-30B and SciJudge-4B, models trained on millions of arXiv papers to predict citation impact, achieving state-of-the-art accuracy surpassing larger models like GPT-5.2 and Gemini 3 Pro.
deepseek-ai/DeepSeek-V4-Flash-DSpark
DeepSeek releases V4 series of Mixture-of-Experts language models (Pro 1.6T/49B activated, Flash 284B/13B activated) supporting one-million-token context with hybrid attention and speculative decoding, claiming best open-source model performance.
deepseek-ai/DeepSeek-V4-Pro
DeepSeek releases V4-Pro and V4-Flash, Mixture-of-Experts models supporting million-token context with hybrid attention and Muon optimizer.
@NielsRogge: That's right, a 3.9GB file getting 99.2% on Math 500 A dataset once introduced by @OpenAI in the paper "Let's verify st…
PrismML's Bonsai-27B model achieves 99.2% on the Math 500 dataset, outperforming DeepSeek-R1 and Kimi K2.
@AlphaSignalAI: A 4B model can now anticipate scientific breakthroughs before scientists do. Researchers often build breakthroughs by c…
A new paper introduces GIANTS-4B, a 4-billion-parameter model trained with reinforcement learning to predict scientific insights by combining ideas from foundational papers, achieving higher similarity and citation potential than larger models like Gemini 3 Pro.