few-shot

Tag

Cards List
#few-shot

@skalskip92: Qwen3.8-Max can be prompted with positive and negative boxes and use them to generate new detections super useful when …

X AI KOLs Timeline · 2026-08-04 Cached

SkalskiP highlights Qwen3.8-Max, a vision-language model for object detection that can be prompted with positive and negative boxes to generate detections, achieving 60-80% mAP with single or multiple prompts and performing well on diverse image types.

0 favorites 0 likes
#few-shot

Dual-Path LLM Reasoning for Multimodal Few-Shot Knowledge Graph Completion

arXiv cs.CL · 2026-07-30 Cached

Proposes DuPLeR, a dual-path LLM reasoning framework for multimodal few-shot knowledge graph completion, combining LLM-derived type priors with factual structures to improve inductive KGC under data scarcity.

0 favorites 0 likes
#few-shot

Cross-Domain Off-Policy Evaluation and Learning for Contextual Bandits

arXiv cs.LG · 2026-07-27 Cached

This paper introduces cross-domain off-policy evaluation and learning (OPE/L) for contextual bandits, allowing the use of logged data from multiple source domains to improve policy evaluation and learning in target domains with challenging conditions like few-shot data, deterministic logging policies, and new actions.

0 favorites 0 likes
#few-shot

DriveDNA: A Large-Scale Multimodal Naturalistic Driving Dataset and Benchmark for Driving Style Identification

Hugging Face Daily Papers · 2026-07-26 Cached

Introduces DriveDNA, a large-scale multimodal naturalistic driving dataset with 4,121 drives from 465 drivers across 115 vehicle models, and a benchmark for driving style identification through tasks like few-shot driver re-identification and personalized behavior prediction.

0 favorites 0 likes
#few-shot

What Transfers Under Source Shift? Definitions, Examples, and Fine-Tuning for Climate Disclosure Classification

arXiv cs.CL · 2026-07-21 Cached

This paper studies how LLM adaptation strategies (definitions, examples, fine-tuning) transfer under source shift in climate disclosure classification, finding that simpler strategies like definitions transfer more consistently than complex ones.

0 favorites 0 likes
#few-shot

Large Language Models for Citation Function Classification

arXiv cs.CL · 2026-07-21 Cached

This paper presents a comprehensive evaluation of five large language models for citation function classification, achieving new state-of-the-art results on the ACL-ARC dataset with a fine-tuned Falcon 7B model. It also introduces the AC3 dataset, which includes a seven-category annotation scheme distinguishing neutral acknowledgments from evaluative stances.

0 favorites 0 likes
#few-shot

@FinanceYF5: Oh my god... Fable 5 is back, and it's insanely powerful. Someone asked Fable to make a game called 'Super Smart Racing'... With just 4 prompts and $173 worth of tokens, Fable 5 created this game. (Prompts below)

X AI KOLs Timeline · 2026-07-02 Cached

Fable 5 model only used 4 prompts and $173 worth of tokens to create a game called 'Super Smart Racing', demonstrating its extremely strong generative capabilities.

0 favorites 0 likes
#few-shot

When Reranking Hurts: Uncertainty-Based Gating for Few-Shot Reranking

arXiv cs.CL · 2026-07-01 Cached

This paper challenges the assumption that reranking always improves few-shot selection for LLMs, proposing a training-free gated reranking approach that uses model uncertainty to decide when to rerank, reducing computational costs by 15-80% while slightly improving performance.

0 favorites 0 likes
#few-shot

Comparing BERT Sentence-Pair Classification and Few-Shot LLM Prompting for Detecting Threat and Solution Framing in German Climate News

arXiv cs.CL · 2026-06-26 Cached

This paper compares fine-tuned BERT (gbert-large) with few-shot LLM prompting (Llama 4 Maverick) for detecting threat and solution framing in German climate news sentences. BERT achieves higher F1 scores (0.83 vs 0.78), and an ablation study shows that providing preceding sentence context improves performance.

0 favorites 0 likes
#few-shot

AnySimLite: A Lightweight Few-Shot Similarity Encoder for On-Device Speech-Adjacent Classification

arXiv cs.CL · 2026-06-26 Cached

Introduces AnySimLite, a lightweight similarity encoder for on-device speech-adjacent classification tasks, achieving state-of-the-art or competitive performance while using less than 1/250th the model size of the qLLaMA-LoRA-7B baseline.

0 favorites 0 likes
#few-shot

Do LLMs Reliably Identify Correct Information Units in Aphasic Discourse?

arXiv cs.AI · 2026-06-16 Cached

This study investigates whether instruction-tuned LLMs (Llama-3.1-8B, Qwen2.5-7B, Mistral-7B, Phi-3-mini) can reliably classify Correct Information Units in aphasic discourse transcripts. Few-shot prompting yields competitive F1 scores (0.776–0.817) for three models, but performance varies by severity and human agreement remains insufficient for fully autonomous use.

0 favorites 0 likes
#few-shot

Few-Shot Biomedical Relation Extraction with Large Language Models: A Viable Alternative to Supervised Learning?

arXiv cs.CL · 2026-06-16 Cached

This paper investigates few-shot biomedical relation extraction using prompt-based learning with LLMs, comparing pairwise classification and joint generation approaches. The best model achieves micro-F1 of 0.44, outperforming previous few-shot results but remaining below supervised baselines, while macro-F1 surpasses the supervised baseline on rare relation types.

0 favorites 0 likes
#few-shot

PrintGuard 2.0 — ShuffleNetV2 + few-shot prototypical network, TFLite via LiteRT, ≈5 MB, runs unmodified in the browser (Pyodide) and on CPython [P]

Reddit r/MachineLearning · 2026-06-15

PrintGuard 2.0 is a major rewrite of a few-shot FDM fault detector using a ShuffleNetV2 backbone and prototypical network, now with a single Python engine that runs unmodified on both CPython and Pyodide in the browser via a platform abstraction layer, enabling per-printer sensitivity tuning and fair inference scheduling.

0 favorites 0 likes
#few-shot

Beyond the Golden Teacher: Enhancing Graph Learning through LLM-GNN Co-teaching

arXiv cs.LG · 2026-06-11 Cached

This paper proposes LLM-GNN Co-Teaching, a bidirectional framework for few-shot graph learning on text-attributed graphs. The LLM and GNN exchange confident pseudo-labels and use round-based preference optimization (RPL-PO) to mutually improve, outperforming prior methods on benchmarks.

0 favorites 0 likes
#few-shot

From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models

arXiv cs.LG · 2026-06-02 Cached

Proposes Demo2Reward, a test-time prompt optimization technique for VLM reward models using a few expert demonstrations, significantly reducing false positives and improving policy learning in robotics without additional model training.

0 favorites 0 likes
#few-shot

LLMs for Cardiovascular Risk Prediction from Structured Clinical Data

arXiv cs.CL · 2026-06-02 Cached

This paper presents a hybrid framework that combines structured clinical data with LLM-generated narratives for coronary artery disease prediction, achieving high fidelity in variable extraction and comparing ML models with LLM-based zero-shot and few-shot classification.

0 favorites 0 likes
#few-shot

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning

arXiv cs.AI · 2026-06-01 Cached

GraphARC is a new benchmark for abstract reasoning on graph-structured data, extending the ARC paradigm to graphs. Evaluations of state-of-the-art language models reveal a comprehension-execution gap and performance degradation on larger instances, highlighting scaling challenges.

0 favorites 0 likes
#few-shot

ACIL: Auto Chain of Thoughts for In-Context Learning

arXiv cs.CL · 2026-05-19 Cached

This paper introduces ACIL, an automatic Chain-of-Thought framework to enhance In-Context Learning by generating and pruning reasoning chains, improving LLM performance on complex tasks.

0 favorites 0 likes
#few-shot

Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries

arXiv cs.CL · 2026-05-18 Cached

This paper explores using few-shot prompted LLMs for actionable triage categorization of online patient inquiries into self-care, schedule-visit, urgent-clinician-review, or emergency-referral. The best model (Claude Haiku 4.5 with 12-shot prompting) achieves macro-F1 of 0.475, surpassing supervised baselines, but the authors conclude that LLMs can support triage prioritization and selective human review, not autonomous deployment.

0 favorites 0 likes
#few-shot

FFAvatar: Few-Shot, Feed-Forward, and Generalizable Avatar Reconstruction

Hugging Face Daily Papers · 2026-05-14 Cached

FFAvatar proposes a feed-forward framework for reconstructing high-quality, animatable 3D Gaussian head avatars from few unposed images in seconds, achieving a 5.5 PSNR improvement over state-of-the-art on the NeRSemble benchmark.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback