model-extraction

Tag

Cards List
#model-extraction

Curvature Cryptanalysis of Smooth Transformer Feed-Forward Networks

arXiv cs.LG · 2026-09-01 Cached

The paper introduces a curvature-based cryptanalysis method to extract hidden feed-forward network structures in transformers using black-box queries, achieving high-fidelity functional model substitutes.

0 favorites 0 likes
#model-extraction

Defending against Model Extraction for GNNs with Model Reprogramming

arXiv cs.LG · 2026-08-13 Cached

This paper proposes GraphRP, a proactive defense framework using model reprogramming to protect GNNs from model extraction attacks, with a structure-aware gating mechanism that preserves benign utility while degrading adversarial queries.

0 favorites 0 likes
#model-extraction

Cryptanalytic Extraction of Isolated Bias-Free GLU Feed-Forward Blocks by Antipodal Separation

arXiv cs.LG · 2026-08-10 Cached

A research paper introducing a multi-stage forward-query method to cryptanalytically extract isolated bias-free GLU feed-forward block weights, demonstrating sub-percent recovery accuracy on Qwen, Llama, and Gemma components while noting end-to-end model API attacks remain unsolved.

0 favorites 0 likes
#model-extraction

ADS-C: Antidistillation Sampling for Classification

arXiv cs.LG · 2026-07-20 Cached

This paper introduces ADS-C, an antidistillation defense for classification that provably preserves top-1 accuracy while degrading student model performance by up to 29.7 percentage points, achieving zero utility cost for the teacher.

0 favorites 0 likes
#model-extraction

Hidden Thoughts Are Not Secret: Reasoning Trace Exposure in LLMs

arXiv cs.AI · 2026-06-02 Cached

This paper introduces Reasoning Exposure Prompting (REP), a method that uses shadow-model demonstrations in code-like formats to elicit hidden reasoning traces from LLMs, showing that interface-level trace hiding is insufficient to prevent extraction of useful reasoning signals.

0 favorites 0 likes
#model-extraction

Can Subgraph Explanations Be Weaponized to Steal Graph Neural Networks?

arXiv cs.LG · 2026-06-01 Cached

This paper presents the first model extraction attack on graph classification under strict black-box constraints, exploiting subgraph explanations to estimate decision boundaries. The findings reveal that mandated explainability interfaces create exploitable security vulnerabilities in Graph Neural Network services.

0 favorites 0 likes
#model-extraction

Did Google hide the best version of Gemma 4 e4b in Android? The extracted model beats Unsloth and everything else I've tried.

Reddit r/LocalLLaMA · 2026-04-21

A user reports that the 3.6 GB Gemma 4 e4b model extracted from Google AI Edge Gallery on Android outperforms larger 3.7 GB Unsloth versions and community ports, raising questions about hidden optimizations.

0 favorites 0 likes
← Back to home

Submit Feedback