text-classification

Tag

Cards List
#text-classification

@interfaze_ai: Jev, now open source: Lev A 4B open source System One model based on Qwen backbone The best performance for it's small …

X AI KOLs Timeline ↗ · 3d ago Cached

Lev is a 4B open-source AI model based on Qwen, designed for efficient typed decision-making in a single forward pass with zero output tokens, achieving strong benchmarks for its size.

0 favorites 0 likes
#text-classification

An Explainable DistilBERT-BiLSTM-Attention Framework for Binary and Multi-Class Hate Speech Detection

arXiv cs.CL ↗ · 4d ago Cached

This paper proposes an explainable hate speech detection framework integrating DistilBERT embeddings, BiLSTM, and an attention mechanism, achieving high F1-scores on benchmark datasets for both binary and multi-class classification.

0 favorites 0 likes
#text-classification

JEV broke down 724 live ads from 37 brands in 40 seconds for $0.09 of tokens

Reddit r/artificial ↗ · 5d ago

Jev, a structured text classification model, demonstrated the ability to process 724 live ads from 37 brands in 40 seconds for just $0.09 in tokens, achieving 216ms median latency per record through parallel processing and integration with Gemini.

0 favorites 0 likes
#text-classification

fastino/GLiNER2.5-Decide

Hugging Face Models Trending ↗ · 5d ago Cached

GLiNER2.5-Decide is a 340M parameter English classification model for operational decisions, supporting multiple label sets in a single forward pass without prompt templates.

0 favorites 0 likes
#text-classification

@Zefan_Cai: Spot the wrong target before you hit Confirm. Try the Open-Jev-27B-v1.1 demo: click between two examples and see how th…

X AI KOLs Following ↗ · 6d ago Cached

Open-Jev-27B-v1.1 is an open-source AI model with a LoRA adapter, achieving 85.28% accuracy on JevBench and featuring interactive demos for various tasks.

0 favorites 0 likes
#text-classification

Building Trustworthy Mental Health Benchmarks on Bluesky: A Validation-Aware Weak-Supervision Framework

arXiv cs.AI ↗ · 6d ago Cached

This paper introduces a validation-aware weak-supervision framework for building and evaluating suicidal ideation and mental health disclosure benchmarks on the decentralized social media platform Bluesky, leveraging AI models like Llama-3-8B and transformer-based classifiers to analyze performance and label validity.

0 favorites 0 likes
#text-classification

Blaming Across the Aisle: Political Contrasting and Blame Attribution in the Danish Parliament

Hugging Face Daily Papers ↗ · 2026-09-22 Cached

This study examines blame attribution in the Danish Parliament from 1997 to 2026 using the BlameBERT classifier, revealing ideological asymmetries and a banana-shaped trajectory in political discourse.

0 favorites 0 likes
#text-classification

Beyond Accuracy: Centroid-Guided Contrastive Loss for Structured Fraudulent Job Posting Detection

arXiv cs.AI ↗ · 2026-09-21 Cached

This paper proposes Centroid-Guided Contrastive Loss (CGCL) for structured fraudulent job posting detection, unifying classification and clustering in latent space to achieve state-of-the-art performance.

0 favorites 0 likes
#text-classification

FakeSpotter: A content and strategy agnostic Viral Misinformation Detection Tool

arXiv cs.CL ↗ · 2026-09-18 Cached

FakeSpotter is a content- and strategy-agnostic tool designed to estimate the viral misinformation risk of textual content by measuring structural fingerprints, using repeated LLM assessments and logistic regression classifiers, with reported macro F1 scores of 0.788 for short texts and 0.793 for long texts on a labeled corpus.

0 favorites 0 likes
#text-classification

The record is part of the task: matched-record evaluation of text classifiers across maintenance, safety and recall reporting

arXiv cs.LG ↗ · 2026-09-16 Cached

This paper proposes matched-record evaluation for text classifiers in industrial settings, demonstrating that record selection significantly impacts performance metrics across maintenance, safety, and recall systems.

0 favorites 0 likes
#text-classification

LLMs or Naive Bayes? Old Gems or New Ways

arXiv cs.LG ↗ · 2026-09-15 Cached

The paper benchmarks Naive Bayes against large language models for text classification, finding that NB remains competitive with LLMs when labeled data is available, offering higher throughput and lower energy consumption for resource-constrained environments.

0 favorites 0 likes
#text-classification

LLM-Enhanced Dual-Branch Learning for Large-Scale Multi-Label Text Classification

arXiv cs.CL ↗ · 2026-09-14 Cached

The paper proposes DualMLC, a dual-branch framework that combines autoregressive and bidirectional language models for large-scale multi-label text classification, achieving state-of-the-art results on benchmarks.

0 favorites 0 likes
#text-classification

RAPID: Reliability-Aware Pair Importance Distillation

arXiv cs.AI ↗ · 2026-09-10 Cached

RAPID introduces a reliability-aware pair importance distillation method to improve knowledge distillation efficiency and performance in text classification tasks.

0 favorites 0 likes
#text-classification

Train Smarter, Not Harder: Switching Signal-Guided Training in Active Learning

Hugging Face Daily Papers ↗ · 2026-09-06 Cached

This paper introduces HybridAL, a method for active learning that dynamically switches between retraining and fine-tuning based on a stabilization signal, improving efficiency and performance in text classification tasks.

0 favorites 0 likes
#text-classification

The Signal in the Noise: An Auditable Reliability Layer for Biomedical Text Classification

arXiv cs.AI ↗ · 2026-09-01 Cached

The paper presents an auditable reliability layer for biomedical text classification that uses deterministic spell-correction to address OCR artifacts, improving classifier performance while ensuring safety through abstention under uncertainty.

0 favorites 0 likes
#text-classification

MIL-BERT: Classification of Arbitrarily Large Text with Performance and Explanatory Guarantees

arXiv cs.CL ↗ · 2026-08-24 Cached

The paper introduces MIL-BERT, an algorithm for classifying arbitrarily long texts by selecting relevant excerpts, achieving state-of-the-art results on multiple datasets with performance and explanatory guarantees.

0 favorites 0 likes
#text-classification

J-Miner: Recovering Executable Decision Knowledge from Language-Model Classifiers

arXiv cs.LG ↗ · 2026-08-19 Cached

J-Miner recovers executable decision knowledge from fine-tuned language-model classifiers by mining named concepts and learning decision rules, enabling inspection and transfer to lightweight models with high fidelity.

0 favorites 0 likes
#text-classification

Discrete Diffusion Language Models Are Training-Free Multi-Label Classifiers

arXiv cs.LG ↗ · 2026-08-18 Cached

The paper proposes dLLM-SetScore, a training-free framework using discrete masked-diffusion language models for multi-label text classification, achieving competitive performance with minimal validation data.

0 favorites 0 likes
#text-classification

Geometric Filtering of LLM-Generated Samples for Few-Shot Text Classification

arXiv cs.CL ↗ · 2026-08-17 Cached

This paper proposes a geometric filtering framework that selects high-quality LLM-generated samples by evaluating their Euclidean distance to real class examples in an embedding space, improving few-shot text classification performance by +2.61 percentage points over SMOTE.

0 favorites 0 likes
#text-classification

Detection of Self-Introductions in Legislative Testimony

arXiv cs.CL ↗ · 2026-08-11 Cached

This paper presents a machine learning pipeline for detecting self-introductions in legislative committee testimony, using features like bag-of-words and BERT probabilities. XGBoost achieves the best F1 score of 0.9747, improving further with BERT-augmented features.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback