tensorrt

Tag

Cards List
#tensorrt

Whisper Live - A nearly-live implementation of Open AI's Whisper, free & open-source

Reddit r/artificial · 2d ago Cached

WhisperLive is an open-source real-time transcription tool using OpenAI's Whisper, supporting multiple backends like faster-whisper and TensorRT for live speech-to-text.

0 favorites 0 likes
#tensorrt

TurboOCR v3 — high-speed document OCR server (C++/CUDA), ~520 img/s on RTX 5090

Reddit r/LocalLLaMA · 2026-06-30

TurboOCR v3 is a self-hosted, high-speed OCR server that achieves ~520 images per second on an RTX 5090 using PP-OCRv6 models, with new structured parsing for tables and formulas.

0 favorites 0 likes
#tensorrt

@PyTorch: Bridging the gap between model optimization and production deployment This tutorial walks through a typical end-to-end …

X AI KOLs Following · 2026-06-16 Cached

This tutorial from NVIDIA walks through the end-to-end workflow of converting an FP8-quantized PyTorch model into a TensorRT inference engine for production deployment, covering ONNX export and performance profiling.

0 favorites 0 likes
#tensorrt

DEMON: Diffusion Engine for Musical Orchestrated Noise

Hugging Face Daily Papers · 2026-05-27 Cached

DEMON presents a real-time diffusion engine that enables live musical performance by controlling the denoising process, achieving up to 12.3 decoder completions per second on a single RTX 5090. It introduces heterogeneous scheduling, shared mutable state, per-frame blending, and windowed VAE decode for responsive control.

0 favorites 0 likes
#tensorrt

@nicos_ai: NVIDIA has just officially published the Skills they use for their AI agents. Right now they have Skills for: → analyzi…

X AI KOLs Timeline · 2026-05-24 Cached

NVIDIA has officially published a set of Skills for AI agents, covering video analysis, voice agents, LLM training, model acceleration, RAG, secure environments, logistics optimization, and CUDA programming.

1 favorites 1 likes
← Back to home

Submit Feedback