reasoning-llm

Tag

Cards List
#reasoning-llm

Granite 4.2 LLMs: How They're Built

Hugging Face Blog · yesterday Cached

Granite 4.2 is IBM's new family of reasoning LLMs available in 3B, 8B, and 30B sizes, featuring thinking modes, tool calling, and trained with a multi-stage reinforcement learning pipeline under the Apache 2.0 license.

0 favorites 0 likes
#reasoning-llm

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

arXiv cs.CL · 2026-06-30 Cached

MIThinker proposes a lightweight reasoning model for motivational interviewing counseling agents, trained via supervised fine-tuning and reinforcement learning to generate therapeutic thoughts, achieving MI competency comparable to state-of-the-art systems with lower computation.

0 favorites 0 likes
#reasoning-llm

A specialized reasoning large language model for accelerating rare disease diagnosis: a randomized AI physician assistance trial

arXiv cs.AI · 2026-06-24 Cached

This paper presents RaDaR, a 32B open-source reasoning LLM trained on public and synthetic rare disease cases, which outperforms larger models like DeepSeek-R1 in diagnosis benchmarks and improves physician accuracy by 21.44 percentage points in a randomized trial.

0 favorites 0 likes
← Back to home

Submit Feedback