zero-shot

Tag

Cards List
#zero-shot

Audio8/Audio8-TTS-Preview-0.1b

Hugging Face Models Trending ↗ · 2026-08-19 Cached

Audio8 TTS Preview 0.1b is a compact zero-shot text-to-speech model with approximately 170M parameters for the main model, supporting voice cloning and multiple languages.

0 favorites 0 likes
#zero-shot

Roboflow Playground: Try and Compare 30 Computer Vision Models

Hacker News Top ↗ · 2026-08-17 Cached

Roboflow Playground is a new developer tool that enables side-by-side comparison of over 30 computer vision models across tasks like object detection and classification, streamlining the evaluation process.

0 favorites 0 likes
#zero-shot

TinyCast: Probabilistic Zero-Shot Forecasting with Computed Periodicity

Hugging Face Daily Papers ↗ · 2026-08-16 Cached

TinyCast is a compact zero-shot time series foundation model with computed periodicity, enabling efficient probabilistic forecasting on edge devices like Cortex-M7.

0 favorites 0 likes
#zero-shot

Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Text-to-Image Retrieval

arXiv cs.AI ↗ · 2026-08-13 Cached

This paper compares natively multimodal embedding models (Gemini Embedding 2, Amazon Nova 2) against frontier LLMs (GPT-4.1, Claude Sonnet 4.6) for hard-negative text-to-image retrieval, finding comparable accuracy but much lower latency for embedding-based ranking.

0 favorites 0 likes
#zero-shot

The Field Knows: Cross-Dimensional Geometry from Navigation to Black Holes

arXiv cs.AI ↗ · 2026-08-11 Cached

This paper introduces a continuous metric field framework trained by a single causal contrastive loss that unifies geometric structure discovery from robot navigation to black hole emergence, demonstrating zero-shot generalization across dimensions.

0 favorites 0 likes
#zero-shot

Once a Response, Always a Response: Detecting LLM-generated Text via Latent Prompt Restoration

arXiv cs.CL ↗ · 2026-08-07 Cached

This paper proposes EchoPrompt, a training-free detector for LLM-generated text that restores a latent prompt dependency by prepending a generic prefix and measuring likelihood gain differences between instruction-tuned and base models, achieving state-of-the-art zero-shot detection performance.

0 favorites 0 likes
#zero-shot

Different Perturbations, Different Mechanisms: Understanding Continued Pre-training for Zero-Shot Dialect Robustness

arXiv cs.CL ↗ · 2026-08-07 Cached

This paper systematically studies perturbation-based continued pre-training (CPT) for improving zero-shot dialect robustness in multilingual LLMs, comparing six training conditions across German, Italian, and Arabic. It finds that character-noised CPT is the most effective general strategy and reveals that different perturbation methods induce distinct robustness mechanisms.

0 favorites 0 likes
#zero-shot

HyperODE: Zero-Shot Surrogate for Simulation and Inference of Dynamical Systems

arXiv cs.LG ↗ · 2026-08-04 Cached

Introduces HyperODE, a zero-shot surrogate that maps ODE structures to hypergraphs, enabling simulation and parameter inference across entire families of dynamical systems without retraining.

0 favorites 0 likes
#zero-shot

DE-NER : Zero-shot Named Entity Recognition via Dialogue Elicitation of Large Language Models

arXiv cs.CL ↗ · 2026-08-04 Cached

Introduces DE-NER, a dialogue elicitation framework for zero-shot named entity recognition that uses self-play between questioner and roleplayer LLMs to clarify entity boundaries, achieving an average 3.75% F1 improvement over baselines.

0 favorites 0 likes
#zero-shot

Can Zero-Shot LLMs Predict Child Malnutrition? A Fairness and Temporal Robustness Study

arXiv cs.CL ↗ · 2026-08-03 Cached

A study evaluating zero-shot GPT-4o-mini for predicting child stunting from Bangladesh Demographic and Health Survey data, comparing against a random forest baseline and assessing fairness across demographic groups and temporal robustness. Results show comparable balanced accuracy but notable fairness disparities across residence and wealth categories.

0 favorites 0 likes
#zero-shot

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks

Hugging Face Daily Papers ↗ · 2026-08-03 Cached

This paper introduces SwanTale, a unified multi-speaker speech and audio generation model supporting both zero-shot and instruct tasks, along with SwanData-Caption for data annotation and SwanVAE for high-quality multi-audio-modality generation.

0 favorites 0 likes
#zero-shot

Selecting Open-Weight Language Models for Zero-Shot Intent Classification: A Systematic Evaluation of 41 Models

arXiv cs.CL ↗ · 2026-07-31 Cached

This paper systematically evaluates 41 open-weight language models (135M–9B) for zero-shot intent classification across 8 datasets, analyzing accuracy, calibration, robustness, and deployment efficiency. It finds instruction-tuned 3B models can beat 7B base models and that some benchmarks like SNIPS are saturated.

0 favorites 0 likes
#zero-shot

Knowledge before Reasoning: EC-Reason-Bench, a Training-Free Diagnostic Benchmark for LLM Enzyme Classification

arXiv cs.CL ↗ · 2026-07-30 Cached

This paper introduces EC-Reason-Bench, a training-free diagnostic benchmark to analyze why general LLMs fail on enzyme EC number prediction. It finds that external knowledge is decisive and must precede reasoning, and that reasoning over evidence acts as an arbiter of conflicting nearest neighbors rather than a source of new knowledge.

0 favorites 0 likes
#zero-shot

Reasoning with Memory: A Temporal Granularity-Adaptive Framework for Training-Free Long Video Understanding

arXiv cs.AI ↗ · 2026-07-29 Cached

ReMem introduces a dual-level memory-augmented keyframe selection framework for training-free long video understanding, achieving state-of-the-art zero-shot performance on multiple benchmarks.

0 favorites 0 likes
#zero-shot

OVEarth-Bench: Evaluating Category Breadth and Query Diversity for Open-Vocabulary Earth Observation

Hugging Face Daily Papers ↗ · 2026-07-29 Cached

Introduces OVEarth-Bench, a benchmark for open-vocabulary Earth observation that broadens category coverage and query diversity, revealing that current methods remain limited and MLLM-based approaches perform best.

0 favorites 0 likes
#zero-shot

Zero-Shot Mission-Level Evaluation for Aerial MLLM Agents

arXiv cs.CL ↗ · 2026-07-27 Cached

MissionBench is a new benchmark for evaluating multimodal large language models (MLLMs) on long-horizon embodied tasks in aerial 3D environments, revealing that even the best models succeed on fewer than 35% of missions compared to 84.4% human performance.

0 favorites 0 likes
#zero-shot

DWT-Fusion: A Signal-Based Framework for Training-Free LLM-Generated Text Detection

arXiv cs.CL ↗ · 2026-07-27 Cached

Introduces DWT-Fusion, a training-free framework using discrete wavelet analysis of token log-probabilities for detecting LLM-generated text, achieving strong AUROC results on multiple datasets.

0 favorites 0 likes
#zero-shot

A Knowledge-Injection Framework for Zero-Shot Adaptation of LLMs to Delirium Prediction

arXiv cs.CL ↗ · 2026-07-24 Cached

Presents a lightweight knowledge-injection framework for zero-shot ICU delirium prediction that augments structured EHR data summaries with external clinical knowledge at inference time, improving AUROC by up to 8.57 percentage points on LLaMA models without fine-tuning.

0 favorites 0 likes
#zero-shot

A Graph Neural Network approach to zero-shot Digital Twins

arXiv cs.LG ↗ · 2026-07-24 Cached

This paper presents a novel framework for zero-shot Digital Twins that integrates real-time visual perception with a geometry-agnostic, physics-informed Graph Neural Network. The approach uses a Thermodynamics-Informed GNN to enforce energy conservation and entropy production, achieving physically accurate simulations on unseen geometries without retraining.

0 favorites 0 likes
#zero-shot

Large Language Models for Citation Function Classification

arXiv cs.CL ↗ · 2026-07-21 Cached

This paper presents a comprehensive evaluation of five large language models for citation function classification, achieving new state-of-the-art results on the ACL-ARC dataset with a fine-tuned Falcon 7B model. It also introduces the AC3 dataset, which includes a seven-category annotation scheme distinguishing neutral acknowledgments from evaluative stances.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback