computer-vision

Tag

Cards List
#computer-vision

Hackers reveal how Flock cameras really track cars and people

Ars Technica ↗ · 2026-09-17 Cached

Hackers hacked Flock Safety cameras, copying data to reveal that the system tracks both vehicles and people in detail, raising privacy concerns and exposing surveillance capabilities.

0 favorites 0 likes
#computer-vision

102 raw inventory photos → 22 live eBay listings. Now I’m trying to use less AI

Reddit r/artificial ↗ · 2026-09-17

The author tested an AI system to turn raw inventory photos into eBay listings, finding that ensuring correct reasoning and handling marketplace quirks are more challenging than initial item identification.

0 favorites 0 likes
#computer-vision

Training-Adaptive Convolutional Sparse Coding via Information Bottleneck for Robust Visual Representation

Hugging Face Daily Papers ↗ · 2026-09-17 Cached

This paper proposes a training-adaptive convolutional sparse coding framework that leverages information bottleneck principles for robust visual representation, achieving improved performance on CIFAR and ImageNet under input perturbations.

0 favorites 0 likes
#computer-vision

FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations

Hugging Face Daily Papers ↗ · 2026-09-17 Cached

FAMOS is a feed-forward model that predicts movable-part segmentation and joint parameters from sparse point clouds using a Multi-state Articulation Transformer and a procedural data generator, showing consistent improvements over baselines in experiments.

0 favorites 0 likes
#computer-vision

@StabilityAI: Color consistency is a persistent friction point in production. Our interactive research team just returned from The 19…

X AI KOLs Timeline ↗ · 2026-09-15

Stability AI's research team presented new work on color consistency for AI-generated images at the 19th European Conference on Computer Vision, addressing production challenges in ensuring color matching across shots.

0 favorites 0 likes
#computer-vision

@HuggingPapers: NVIDIA just released FoundationPose on Hugging Face A unified foundation model for 6-DoF object pose estimation and tra…

X AI KOLs Timeline ↗ · 2026-09-15 Cached

NVIDIA released FoundationPose on Hugging Face, a unified foundation model for 6-DoF object pose estimation and tracking that works on novel objects without fine-tuning.

0 favorites 0 likes
#computer-vision

Discovering and Preserving Category Correlation Knowledge via Adaptive Reciprocal Knowledge Distillation

arXiv cs.LG ↗ · 2026-09-15 Cached

This paper proposes adaptive reciprocal knowledge distillation (AR-KD), a novel method that improves knowledge transfer from teacher to student models by simplifying the teacher's output distribution through relational alignment, achieving up to 7.13% accuracy gain on CIFAR-100 and ImageNet-1k datasets.

0 favorites 0 likes
#computer-vision

TAPe+ML: A Compact Structured Representation for Multi-Task Computer Vision

Hugging Face Daily Papers ↗ · 2026-09-15 Cached

The paper introduces TAPe+MLv3, a compact computer vision system using structured representation for multi-task tasks, achieving competitive performance on benchmarks like COCO with fewer than 100,000 parameters.

0 favorites 0 likes
#computer-vision

Adversarial Fashion Makes a Statement on AI Panopticon

Hacker News Top ↗ · 2026-09-14 Cached

Adversarial fashion is emerging as a creative response to AI-powered surveillance, using colorful patterns and designs to evade object detection systems, with projects and products like noRecognition and Cap_able already available.

0 favorites 0 likes
#computer-vision

[P] Built a 100% Client-Side Vision Pipeline for Real-Time Chessboard & Multi-Board Detection (Chrome/Firefox Extension) [P]

Reddit r/MachineLearning ↗ · 2026-09-14

A browser extension called ChessInsights AI has been developed for real-time chessboard detection and analysis using 100% client-side computer vision, ensuring privacy by processing all data locally without server uploads.

0 favorites 0 likes
#computer-vision

@sts_3d: Sentradel is hiring! See https://sentradel.com for more. If you DM me, please include details on relevant projects / ex…

X AI KOLs Following ↗ · 2026-09-13 Cached

Sentradel is hiring and describes their autonomous counter-drone systems that detect, track, and engage small drones cost-effectively using vision and thermal sensing.

0 favorites 0 likes
#computer-vision

@oliviscusAI: this tool can track perfect 3D motion. rtmlib is a lightweight pose estimation library covering full body, hands, face,…

X AI KOLs Following ↗ · 2026-09-13 Cached

rtmlib is a lightweight, open-source pose estimation library that supports full-body, hand, face, and animal pose tracking, built on rtmpose and vitpose models, with a built-in Gradio web UI.

0 favorites 0 likes
#computer-vision

RelateAnything: Real-Time Open-Vocabulary Relation Prediction From Any Inputs

Hugging Face Daily Papers ↗ · 2026-09-11 Cached

RelateAnything is a lightweight, real-time open-vocabulary relation prediction model that accepts arbitrary predicate vocabularies and region sources, trained on a large geometrically verified dataset and evaluated on new cross-dataset benchmarks, showing significant performance gains over comparable methods.

0 favorites 0 likes
#computer-vision

@Tesla: Photon count reconstruction FTW

X AI KOLs Following ↗ · 2026-09-10 Cached

Tesla's Full Self-Driving (FSD) system effectively handles sunlight glare scenarios, as demonstrated by a user's positive experience shared in a tweet, highlighting its advanced photon count reconstruction capabilities.

0 favorites 0 likes
#computer-vision

Show HN: MultiMatte, a Promptable Image Background Removal Model

Hacker News Top ↗ · 2026-09-10 Cached

MultiMatte is a promptable image background removal model fine-tuned from SAM 3 using LoRA, achieving higher accuracy on benchmarks by outputting alpha mattes for better handling of fuzzy boundaries.

0 favorites 0 likes
#computer-vision

@skalskip92: basketball AI (95% local AI + 5% GPT-6 Astra) - detect ball and players - track players - re-identify players across pl…

X AI KOLs Timeline ↗ · 2026-09-10 Cached

A basketball AI system that combines local AI with GPT-6 Astra to detect and track players, perform OCR for player numbers, and map trajectories for sports analytics.

0 favorites 0 likes
#computer-vision

@Michael_J_Black: Today at ECCV: Towards in-the-wild Egocentric 3D Hand-Object Pose Estimation, ExHall Poster: #70, PM CEST, Poster Sessi…

X AI KOLs Timeline ↗ · 2026-09-10 Cached

This paper presents EPIC-Contact, an in-the-wild dataset for 3D hand-object pose estimation, and HOPformer, a transformer model that jointly predicts hand and object poses from a single RGB image.

0 favorites 0 likes
#computer-vision

Feature Recovery for Object Understanding After Irreversible Fire Damage

Hugging Face Daily Papers ↗ · 2026-09-10 Cached

This paper introduces TRACE, a benchmark for post-fire object understanding, and proposes a Feature Recovery Module (FRM) to restore degraded features and improve detection and vision-language tasks under severe physical damage.

0 favorites 0 likes
#computer-vision

FreeFlow: A Bias-free Hierarchical Transformer for Optical Flow Estimation

Hugging Face Daily Papers ↗ · 2026-09-10 Cached

FreeFlow is a hierarchical transformer for optical flow estimation that eliminates task-specific inductive biases and achieves state-of-the-art accuracy on benchmarks like Sintel and KITTI-2015.

0 favorites 0 likes
#computer-vision

World in World: Explore the World with World Models

Hugging Face Daily Papers ↗ · 2026-09-10 Cached

The paper presents World in World, a training-free interface that enables flexible camera and time control in frozen autoregressive video world models by using correspondence-guided queries and evidence-wise attention guidance.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback