nemotron

Tag

Cards List
#nemotron

Can someone explain what "controlling reasoning with system prompt" means?

Reddit r/artificial ↗ · 2026-09-14 Cached

This article details the release of NVIDIA's Nemotron-3-Nano-4B-GGUF model, a small language model designed for both reasoning and non-reasoning tasks with reasoning controllable via system prompts.

0 favorites 0 likes
#nemotron

Palantir Foundry and cuOpt drive NVIDIA supply chain allocation

Reddit r/artificial ↗ · 2026-09-12 Cached

NVIDIA uses Palantir Foundry and cuOpt to automate supply chain allocation decisions, training Nemotron 3.5 Lightning on unstructured operational data to improve decision accuracy.

0 favorites 0 likes
#nemotron

An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics

Hugging Face Daily Papers ↗ · 2026-09-09 Cached

The paper presents an open-method using post-trained Nemotron 3 Ultra checkpoints to achieve gold-medal performance on IMO 2026 through iterative verification and refinement in natural language without external tools.

0 favorites 0 likes
#nemotron

@AIatMeta: As a test of our progress to advance the frontier of AI research, in June we entered the next generation of our autonom…

X AI KOLs Timeline ↗ · 2026-09-05

Meta's autonomous AI research system AIRA₃ placed 8th out of approximately 4,000 teams to win gold in a NVIDIA Kaggle competition to fine-tune a 30B Nemotron model, outperforming human competitors with access to the same tools.

0 favorites 0 likes
#nemotron

model: add NVIDIA Nemotron-3-Puzzle-75B-A9B (NemotronHPuzzle) support by YanissAmz · Pull Request #25444 · ggml-org/llama.cpp

Reddit r/LocalLLaMA ↗ · 2026-09-03 Cached

A pull request adds support for the NVIDIA Nemotron-3-Puzzle-75B-A9B model in the llama.cpp inference tool.

0 favorites 0 likes
#nemotron

nvidia/Nemotron-3-Diarization

Hugging Face Models Trending ↗ · 2026-09-01 Cached

Nemotron-3-Diarization is an open-weight speaker diarization model by NVIDIA for real-world audio analysis, supporting streaming and offline inference up to eight speakers with commercial use permitted.

0 favorites 0 likes
#nemotron

Nemotron-3.5-Lightning at 11.77 GiB, a 16 GB option for a model that didn't have one

Reddit r/LocalLLaMA ↗ · 2026-08-29

ShimQuant enables running Nemotron-3.5-Lightning on 16 GB GPUs with a 11.77 GiB quantized file, providing a usable option below previous 18 GiB limits.

0 favorites 0 likes
#nemotron

@PyTorch: Use PyTorch-native libraries within the NVIDIA NeMo Framework to customize models to hit your exacting requirements for…

X AI KOLs Timeline ↗ · 2026-08-27 Cached

NVIDIA demonstrates how quantization-aware distillation (QAD) using NVIDIA Model Optimizer improves the Nemotron 3.5 Lightning model, reducing memory usage and increasing throughput while preserving accuracy for agentic benchmarks.

0 favorites 0 likes
#nemotron

Nvidia Poolside deal to compete with Chinese Open Weights

Reddit r/LocalLLaMA ↗ · 2026-08-23

Nvidia has made a substantial investment and licensing deal with Poolside, involving a $1 billion investment and $6 billion payment, with over 100 engineers joining Nvidia to work on Nemotron.

0 favorites 0 likes
#nemotron

@modal: Vice President of Applied Deep Learning Research at @NVIDIAAI, @ctnzr, is joining us on stage at Runtime. Bryan leads t…

X AI KOLs Following ↗ · 2026-08-21 Cached

Bryan Catanzaro, NVIDIA's VP of Applied Deep Learning Research, will speak at Runtime about the future of open AI models, drawing on his work with the Nemotron team and past contributions to cuDNN, DLSS, and Megatron.

0 favorites 0 likes
#nemotron

@no_stp_on_snek: Nvidia's Nemotron 3.5 Lightning Behavioral Analysis It caught an auth bug a tech lead and four approvers had already si…

X AI KOLs Timeline ↗ · 2026-08-12 Cached

A behavioral audit of Nvidia's Nemotron 3.5 Lightning finds strong integrity, catching a subtle auth bug in reviewed code and passing safety probes, while noting a blind spot in action-based honesty.

0 favorites 0 likes
#nemotron

Tested Nemotron 3.5 Lightning locally on coding, Hermes Agent and agentic work

Reddit r/LocalLLaMA ↗ · 2026-08-12

User tested Nemotron 3.5 Lightning locally with llama.cpp and quants, finding good speed and agentic tool-calling but below-expectation coding output for its size.

0 favorites 0 likes
#nemotron

@FinanceYF5: NVIDIA releases Nemotron 3.5 Lightning with 30B total parameters, only 3B active, supports 1M token context, commercially usable. Local deployment friendly: BF16 and smaller NVFP4 versions, can run on a single H100 or DGX Spark...

X AI KOLs Following ↗ · 2026-08-12 Cached

NVIDIA releases Nemotron 3.5 Lightning model, with 30B total parameters and only 3B active, supports 1M token context, commercially usable, local deployment friendly, output speed up to 4x faster.

0 favorites 0 likes
#nemotron

Nvidia's Switchyard router reshuffles AI models mid-task, cutting task costs to a third in its own tests (8 minute read)

TLDR AI ↗ · 2026-08-12 Cached

Nvidia released Nemotron 3.5 Lightning, a 30B open mixture-of-experts model, and NeMo Switchyard, an open-source routing library that dynamically assigns each step of an AI agent workflow to the most suitable model. Nvidia claims the combination can cut agent task costs to about a third while maintaining frontier-level performance.

0 favorites 0 likes
#nemotron

@heyshrutimishra: NVIDIA just dropped Nemotron 3.5 Lightning 30 billion parameters. Only 3 billion active. Built for the execution layer …

X AI KOLs Following ↗ · 2026-08-11 Cached

NVIDIA released Nemotron 3.5 Lightning, a 30B-parameter MoE model with only 3B active parameters, optimized for agent execution tasks. It claims faster, cheaper tool calls and agent execution while staying fully open-source under OpenMDW-1.1.

0 favorites 0 likes
#nemotron

NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AI

NVIDIA Blog ↗ · 2026-08-11 Cached

NVIDIA announced Nemotron 3.5 Lightning, a 30B mixture-of-experts open model optimized for high-volume agentic AI workloads, alongside NeMo Switchyard, an open-source library for intelligent model routing across heterogeneous model ecosystems.

0 favorites 0 likes
#nemotron

@nvidia: Every business has different data, workflows and standards. Its AI should reflect that. Watch to see how NVIDIA Nemotro…

X AI KOLs Timeline ↗ · 2026-08-06 Cached

NVIDIA highlights how its Nemotron open models let teams build specialized, trustworthy AI tailored to their business data and workflows.

0 favorites 0 likes
#nemotron

@googledevs: Over 5,000 Kagglers. One @nvidia challenge. Five key takeaways. Learn how developers fine-tuned reasoning models using …

X AI KOLs Timeline ↗ · 2026-08-05 Cached

NVIDIA shares five lessons from over 5,000 Kagglers who fine-tuned reasoning models using LoRA adapters and synthetic chain-of-thought data in the Nemotron Model Reasoning Challenge, focusing on verifiable data, token budget, and infrastructure.

0 favorites 0 likes
#nemotron

Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains

Hugging Face Daily Papers ↗ · 2026-08-05 Cached

This paper presents an end-to-end adaptation of NVIDIA's Nemotron retrieval stack for Modern Greek, including a new benchmark HERA and models fine-tuned for retrieval, reranking, and grounded generation across specialist domains.

0 favorites 0 likes
#nemotron

nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4

Hugging Face Models Trending ↗ · 2026-08-04 Cached

NVIDIA released Nemotron 3.5 Lightning 30B-A3B-NVFP4, a hybrid MoE LLM with 3B active parameters, up to 1M context, and speculative decoding support for efficient single-GPU inference.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback