nemotron

Tag

Cards List
#nemotron

@nvidia: Every business has different data, workflows and standards. Its AI should reflect that. Watch to see how NVIDIA Nemotro…

X AI KOLs Timeline · 4d ago Cached

NVIDIA highlights how its Nemotron open models let teams build specialized, trustworthy AI tailored to their business data and workflows.

0 favorites 0 likes
#nemotron

@googledevs: Over 5,000 Kagglers. One @nvidia challenge. Five key takeaways. Learn how developers fine-tuned reasoning models using …

X AI KOLs Timeline · 5d ago Cached

NVIDIA shares five lessons from over 5,000 Kagglers who fine-tuned reasoning models using LoRA adapters and synthetic chain-of-thought data in the Nemotron Model Reasoning Challenge, focusing on verifiable data, token budget, and infrastructure.

0 favorites 0 likes
#nemotron

Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains

Hugging Face Daily Papers · 6d ago Cached

This paper presents an end-to-end adaptation of NVIDIA's Nemotron retrieval stack for Modern Greek, including a new benchmark HERA and models fine-tuned for retrieval, reranking, and grounded generation across specialist domains.

0 favorites 0 likes
#nemotron

@nvidia: Translation queues and custom integrations can slow global launches. Our Digital Marketing team built an AI-powered loc…

X AI KOLs Timeline · 2026-07-27 Cached

NVIDIA's Digital Marketing team built an AI-powered localization platform using NVIDIA Nemotron Speech, achieving ~70% reduction in translation turnaround time and ~25% cost savings, processing over 11 million words across 20,000 files.

0 favorites 0 likes
#nemotron

From a Multilingual Streaming ASR Backbone to Kenyan-Language Systems: Data-Centric Adaptation of Nemotron 3.5 for Kikuyu, Dholuo, and Kalenjin

arXiv cs.CL · 2026-07-22 Cached

This paper presents an engineering study adapting NVIDIA Nemotron 3.5 ASR Streaming 0.6B to Kikuyu, Dholuo, and Kalenjin, achieving 42.97% and 33.98% WER on internal sets for Kikuyu and Dholuo, respectively, through data-centric techniques including corpus auditing, normalization, and streaming evaluation.

0 favorites 0 likes
#nemotron

P40's + MI50's + RPC on 550B Nemotron Ultra Q3_S

Reddit r/LocalLLaMA · 2026-07-22

User shares benchmark results running the 550B Nemotron Ultra model across two machines using RPC, achieving impressive throughput on older AMD MI50 and Nvidia P40 GPUs.

0 favorites 0 likes
#nemotron

@HuggingApps: NVIDIA Nemotron just dropped an audio-native model that hears the world, not just words transcription, translation, sou…

X AI KOLs Following · 2026-07-20 Cached

NVIDIA released Nemotron, an audio-native model capable of transcription, translation, sound recognition, audio Q&A, TTS, and full speech-to-speech, with open weights in 2B and 30B sizes.

0 favorites 0 likes
#nemotron

NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval

Hugging Face Blog · 2026-07-16 Cached

NVIDIA releases Nemotron 3 Embed, a collection of open embedding models that top the RTEB leaderboard, featuring an 8B flagship model and efficient 1B variants for production-scale retrieval.

0 favorites 0 likes
#nemotron

NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B on 2x3090s

Reddit r/LocalLLaMA · 2026-07-16

A detailed guide on running the quantized NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B model on two RTX 3090s using vLLM with full 262K context, achieving high inference speeds without CPU offloading.

0 favorites 0 likes
#nemotron

@LangChain: In 7 minutes, Partner Engineer Srimanth Tangedipalli shows you how to run Deep Agents Code inside a governed NemoClaw O…

X AI KOLs Following · 2026-07-13 Cached

LangChain shows a 7-minute tutorial by Partner Engineer Srimanth Tangedipalli on running Deep Agents Code inside a governed NemoClaw OpenShell Sandbox with NVIDIA Nemotron 3 Ultra via Baseten.

0 favorites 0 likes
#nemotron

I got Nemotron Puzzle 75B running smoothly on a 64GB M2 Max

Reddit r/LocalLLaMA · 2026-07-12

Successfully ran the 75B Nemotron Puzzle model locally on a 64GB M2 Max Mac, demonstrating large model inference on consumer hardware.

0 favorites 0 likes
#nemotron

@akshay_pachaar: NVIDIA might just have solved the biggest tradeoff in LLMs. Every LLM makes you pick between speed and quality. Autoreg…

X AI KOLs Timeline · 2026-07-11 Cached

NVIDIA introduces TwoTower, a method that decouples context representation and denoising in diffusion language models, achieving 2.42x throughput while retaining 98.7% of autoregressive quality on a 30B MoE backbone.

0 favorites 0 likes
#nemotron

@fujikanaeda: Today marks the end of my last week at Nvidia. I joined with the rest of the excellent Gretel team when we were acquire…

X AI KOLs Timeline · 2026-07-10 Cached

Fuji Kanaeda announces departure from Nvidia after a year, highlighting contributions to synthetic data generation (NeMo Data Designer) and Nemotron LLM builds, praising the team's work on open-source AI.

0 favorites 0 likes
#nemotron

Data for Agents

Hugging Face Blog · 2026-07-08 Cached

NVIDIA discusses the importance of open and synthetic data for building robust AI agents, highlighting their Nemotron open datasets for training, reasoning, and tool-use.

0 favorites 0 likes
#nemotron

NVIDIA Nemotron Achieves Benchmark-Leading Performance With LangChain Deep Agents Harness

NVIDIA Blog · 2026-07-08 Cached

NVIDIA Nemotron 3 Ultra achieves benchmark-leading performance with LangChain Deep Agents harness, offering higher accuracy at lower cost than closed models without retraining.

0 favorites 0 likes
#nemotron

@NVIDIAAI: You're welcome

X AI KOLs Timeline · 2026-07-08 Cached

NVIDIA AI releases a 75B MoE model (9.3B active) compressed from Nemotron-3-Super-120B using the Iterative Puzzle framework, with 1M token context support.

0 favorites 0 likes
#nemotron

nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-BF16 · Hugging Face

Reddit r/LocalLLaMA · 2026-07-07 Cached

NVIDIA releases Nemotron-Labs-3-Puzzle-75B-A9B, a compressed hybrid MoE LLM derived from Nemotron-3-Super, achieving approximately 2× higher server throughput and improved concurrency while maintaining strong accuracy across reasoning, coding, and long-context benchmarks.

0 favorites 0 likes
#nemotron

nvidia/Nemotron-Labs-Audex-30B-A3B · Hugging Face

Reddit r/LocalLLaMA · 2026-07-07 Cached

NVIDIA released Nemotron-Labs-Audex-30B-A3B, a unified audio-text LLM built on a 30B MoE backbone with 3B activated parameters, offering strong performance on audio understanding, speech recognition/translation, and generation while preserving text reasoning and alignment capabilities.

0 favorites 0 likes
#nemotron

How Open Models Are Driving AI Research

NVIDIA Blog · 2026-07-06 Cached

NVIDIA highlights how open frontier models and AI infrastructure are driving AI research, as reflected in accepted papers at ICML 2026, with contributions spanning robotics, life sciences, and synthetic data.

0 favorites 0 likes
#nemotron

@MaziyarPanahi: We got 755 tokens per second! That's OpenMed privacy-filter v2 (nemotron, MLX 8-bit) reading a 13,000-token clinical fi…

X AI KOLs Timeline · 2026-07-04 Cached

OpenMed privacy-filter v2 using nemotron and MLX 8-bit achieves 755 tokens per second on a Mac, redacting 1,152 PII identifiers across 22 categories from a 13,000-token clinical file without data leaving the machine.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback