image

Tag

Cards List
#image

KVAE: Family of Tokenizers for Multimodal Generative Models

Hugging Face Daily Papers · 2026-08-06 Cached

This paper introduces KVAE, a family of tokenizers for audio, image, and video designed for text-conditioned generative models, claiming competitive or superior reconstruction and generation quality compared to existing open-source tokenizers. The code and training details are publicly released.

0 favorites 0 likes
#image

Show HN: I built 184 free browser tools – PDF, image, dev, AI tasks, no upload

Hacker News Top · 2026-06-17 Cached

Brevio provides 184 free browser-based tools for PDF, image, development, and AI tasks, all without requiring uploads or signup.

0 favorites 0 likes
#image

MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image

Hugging Face Daily Papers · 2026-05-11 Cached

Introduces MulTaBench, a benchmark of 40 datasets for multimodal tabular learning with text and image modalities, demonstrating that task-specific embedding tuning improves performance over frozen pretrained embeddings, particularly when modalities provide complementary predictive signals.

0 favorites 0 likes
← Back to home

Submit Feedback