abliteration

Tag

Cards List
#abliteration

"Uncensored" LLMs are measurably more optimistic than their base models

Reddit r/LocalLLaMA · 7h ago

A study on uncensored LLMs (Gemma and Qwen) shows that removing censorship makes them more optimistic in stock market predictions, but not more accurate. The effect varies by model family.

0 favorites 0 likes
#abliteration

23 Gemma4-E4B models compared with abliterlitics: the most downloaded one is also the most broken

Reddit r/LocalLLaMA · 3d ago

A comparison of 23 Gemma 4 E4B models on HuggingFace shows that the most downloaded model, OBLITERATUS, is completely broken, while the more surgical 'heretic' variants perform best.

0 favorites 0 likes
#abliteration

We released an abliterated + fine-tuned GLM-5.2. High scores on adversarial benchmarks while keeping coding performance.

Reddit r/artificial · 4d ago

Released an abliterated and fine-tuned version of GLM-5.2 (abliterated-model-large) that achieves high scores on adversarial and agent benchmarks while maintaining coding performance. The model is available via API with zero data retention and no built-in policy.

0 favorites 0 likes
#abliteration

I created a super harmful model ! :D (by tweaking it's J-Space!!!)

Reddit r/LocalLLaMA · 2026-07-11

The author created a tool based on Anthropic's Jacobian-Lens to manually tweak a model's Jacobian Space, producing an uncensored model called Nikusui-v1, released with GGUF quantizations.

0 favorites 0 likes
#abliteration

Norm-preserving abliteration on Qwen3.6-35B-A3B: 0% refusal, benchmarks intact, open source dataset

Reddit r/LocalLLaMA · 2026-06-30

Norm-preserving abliteration technique applied to Qwen3.6-35B-A3B achieves 0% refusal rate while maintaining benchmark performance, with open source dataset released.

0 favorites 0 likes
#abliteration

@support_huihui: New GGUF: huihui-ai/Huihui-Qwythos-9B-Claude-Mythos-5-1M-abliterated-GGUF This is an uncensored version of empero-ai/Qw…

X AI KOLs Timeline · 2026-06-25 Cached

A new uncensored GGUF quantized version of the Qwythos-9B-Claude-Mythos-5-1M model, created using abliteration, is released on Hugging Face.

0 favorites 0 likes
#abliteration

huihui-ai/Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliterated

Hugging Face Models Trending · 2026-06-21 Cached

An uncensored version of the gemma-4-12B-coder model created using abliteration to remove refusals, intended for research and experimental use.

0 favorites 0 likes
#abliteration

@SpaceTimeViking: Qwen3.6 27B getting some love on the new AEON ULTIMATE VLLM image @NVIDIAAI DGX SPARK OPTIMIZED! https://github.com/AEO…

X AI KOLs Timeline · 2026-06-18 Cached

AEON-7 releases a fully uncensored, capability-enhanced abliteration of Qwen3.6-27B, optimized for NVIDIA DGX Spark with NVFP4 quantization and DFlash speculative decoding for improved performance.

0 favorites 0 likes
#abliteration

@0x0SojalSec: Fully automatic censorship removal for Any LLM models, Built a tool that removes LLM censorship in 45 minutes flat. You…

X AI KOLs Timeline · 2026-06-17 Cached

Heretic is a fully automatic tool that removes censorship from transformer-based LLMs via directional ablation/abliteration, achieving results comparable to manual methods in under an hour with minimal human effort.

0 favorites 0 likes
#abliteration

@support_huihui: New Model: huihui-ai/Huihui-Nex-N2-mini-abliterated This is an uncensored version of nex-agi/Nex-N2-mini created with a…

X AI KOLs Timeline · 2026-06-16 Cached

Huihui AI released an uncensored version of the Nex-N2-mini model created using abliteration, a technique to remove refusals from LLMs. The model lacks safety filtering and is intended for research use only.

0 favorites 0 likes
#abliteration

@elder_plinius: OBLITERATION ALERT GOOGLE: PWNED GEMMA-4-12B: OBLITERATED ‍ 0.0% REFUSAL RATE — NO CAPABILITY LOSS! https://huggingface…

X AI KOLs Following · 2026-06-08 Cached

A novel two-pass ablation technique (ASPA) applied to Gemma-4-12B achieves zero refusal rate with zero capability loss, using source-tethering to recover benchmark performance.

0 favorites 0 likes
#abliteration

huihui-ai/Huihui-gemma-4-12B-it-abliterated

Hugging Face Models Trending · 2026-06-06 Cached

This model is an uncensored version of Google's Gemma 4 12B it model, created using abliteration to remove refusals. It is available on Hugging Face and Ollama, with warnings about sensitive outputs.

0 favorites 0 likes
#abliteration

OBLITERATUS/Gemma-4-12B-OBLITERATED

Hugging Face Models Trending · 2026-06-05 Cached

OBLITERATUS releases Gemma-4-12B-OBLITERATED, the first abliterated model achieving zero refusal without benchmark regression, using a novel two-pass surgery pipeline for alignment research.

0 favorites 0 likes
#abliteration

How does the new abliteration tool Apostate compare with others? - Abliterlitics

Reddit r/LocalLLaMA · 2026-06-03

A detailed comparison of three abliteration tools—Apostate, Heretic, and Huihui—applied to Qwen 2.5 7B, showing they all effectively remove refusal behaviors with minimal performance degradation.

0 favorites 0 likes
#abliteration

These AI models are free, private, and will never say 'no'

Reddit r/artificial · 2026-05-31 Cached

The article discusses the growing accessibility of open-weight AI models whose safety guardrails can be easily removed, allowing them to answer harmful requests without refusal, raising significant concerns about misuse and national security.

0 favorites 0 likes
#abliteration

13 abliterated Gemma 4 E2B variants, 44 GPU hours, Benchmark and Comparison - Abliterlitics

Reddit r/LocalLLaMA · 2026-05-31

A detailed comparison of 13 abliterated variants of Google's Gemma 4 E2B model, evaluating safety removal and capability preservation. It finds that surgical abliteration can preserve or even improve reasoning, while aggressive methods cause significant performance drops.

0 favorites 0 likes
#abliteration

⚠️ Meta's AI safety filters were stripped in less than 10 minutes

Reddit r/ArtificialInteligence · 2026-05-27

A joint test by the Financial Times and AI safety group Alice reveals that safety filters on Meta's Llama 3.3 and Google's Gemma 4 models can be removed in under 10 minutes using a free tool called Heretic, highlighting the difficulty of regulating open-source AI safety.

0 favorites 0 likes
#abliteration

@dealignai: Qwen3.6-27b and 35b MXFP4 MXFP8 CRACK is out now with MTP. Enjoy uncensored speediness! 35b mxfp4: https://huggingface.…

X AI KOLs Timeline · 2026-05-24 Cached

DealignAI releases CRACK-abliterated and MXFP4/MXFP8 quantized versions of Qwen3.6-27B and 35B models, preserving MTP for faster speculative decoding on Apple Silicon.

0 favorites 0 likes
#abliteration

@support_huihui: New MTP-GGUF: huihui-ai/Huihui-Qwen3.6-27B-abliterated-MTP-GGUF This is an uncensored version of Qwen/Qwen3.6-27B creat…

X AI KOLs Timeline · 2026-05-19 Cached

An uncensored GGUF version of Qwen3.6-27B, created via abliteration, is now available on Hugging Face from huihui-ai.

0 favorites 0 likes
#abliteration

85 GPU-hours comparing 5 abliteration methods on Qwen3.6-27B: benchmarks, safety, weight forensics - Abliterlitics

Reddit r/LocalLLaMA · 2026-05-17

This post presents Abliterlitics, an open-source toolkit for analyzing abliteration techniques, and compares five abliteration variants of Qwen3.6-27B using 85 GPU-hours of benchmarks, safety evaluations, and weight forensics. Heretic and Huihui show best capability preservation while all achieve near-complete safety removal.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback