Tag
A web tool that aggregates all 24/7 YouTube live camera streams onto a single interactive globe, allowing users to explore live feeds from around the world.
A developer built a real-time slop detection tool with jev that operates during scrolling to identify low-quality content in browsers.
Jev is a new AI model from TypeSafe AI, designed for high-speed decision-making. As a general-purpose classifier, it offers extremely fast response speeds and low cost, making it ideal for real-time tasks like content filtering, gaming, and trading.
A new AI model named Jev, developed by the co-creator of ChatGPT, is capable of playing the game Subway Surfers in real-time, demonstrating decision-making capabilities.
NetEase Youdao AI is open-sourcing Confucius4-R2T2, a 1.7B frontier real-time streaming ASR model designed for voice agents, featuring low latency, high accuracy, and configurable decoding chunks.
ABot-Recon is an open-source real-time streaming 3D reconstruction model from Amap CV Lab that uses KV cache context and per-frame estimation to generate updatable 3D point cloud maps, achieving 24.45 FPS on H100 GPU and deployable on edge devices like AX650N NPU.
The article highlights the impressive capabilities of the Gemini 3.8 live AI model, showcased by a real-time insurance claim agent that supports voice and multilingual interactions, with an open-source implementation.
Google has released new Gemini models, including Gemini 3.8 Live and 3.5 Transcribe, to enhance real-time voice applications with improved conversational AI and multilingual transcription for developers.
The author describes challenges in team collaboration with AI coding agents, such as conflicting decisions and context drift, and seeks advice from the community.
A browser extension called ChessInsights AI has been developed for real-time chessboard detection and analysis using 100% client-side computer vision, ensuring privacy by processing all data locally without server uploads.
GPT-Live-1's API allows developers to integrate voice agents that can listen and speak simultaneously, providing a natural and fluent conversational experience for applications.
LynnReal-Omni is a unified multimodal video diffusion framework that integrates agentic visual controls with high-fidelity generation and real-time acceleration for stable, controllable video creation.
WalShadow is an open-source tool that replicates Postgres data to ClickHouse with sub-second latency by directly consuming physical WAL, enabling real-time analytics with high throughput and schema evolution support.
The browser has introduced a new feature that allows it to fill the whole screen and be adjusted in real time.
RelateAnything is a lightweight, real-time open-vocabulary relation prediction model that accepts arbitrary predicate vocabularies and region sources, trained on a large geometrically verified dataset and evaluated on new cross-dataset benchmarks, showing significant performance gains over comparable methods.
Phoenix-4.5 is introduced as the fastest and most expressive real-time human rendering model, featuring enhanced facial animation and upper body movement, claimed to be the closest AI to passing the Turing test face to face.
Confucius4-R2T2 is a low-latency, high-accuracy real-time speech recognition model developed by NetEase Youdao, featuring configurable chunking and stable output for applications like live captioning.
Vidu S2 presents real-time interactive AI models for avatar and video editing, enabling high-resolution spatial video generation with dynamic reference updates and superior performance over baselines.
GetCandidly's state model measures gaps in conversation resolution in real time and improves it by inserting prompts that mirror user wording, as highlighted by LangChain.
ByteDance is developing a real-time spatial video model under founder Zhang Yiming's oversight, aiming to generate interactive virtual worlds for XR applications with a potential launch soon.