Tag
NVIDIA NeMo Automodel integrates with Hugging Face Diffusers to enable scalable distributed fine-tuning of diffusion models for image and video generation, supporting models like FLUX.1-dev, Wan 2.1, and HunyuanVideo.
NVIDIA's blog post describes an end-to-end workflow using PyTorch-native NeMo AutoModel for pretraining a transaction foundation model. The workflow uses GPU-accelerated data processing and tokenization, decoder-only model pretraining, and embedding extraction to improve fraud classification performance by over 40% on the IBM TabFormer dataset.
Nous Research integrated NVIDIA's official Agent Skills catalog into the Hermes Skills Hub, enabling agents to use CUDA-X, Omniverse, Physical AI, and NeMo tools.
NVIDIA has officially published a set of Skills for AI agents, covering video analysis, voice agents, LLM training, model acceleration, RAG, secure environments, logistics optimization, and CUDA programming.