safetensors

Tag

Cards List
#safetensors

Deepseek V4.1 Flash is 748B, not 552B

Reddit r/LocalLLaMA · 5d ago

The article clarifies the parameter count of the Deepseek V4.1 Flash AI model, detailing its components like FFN experts and vision encoder, and highlights the substantial hardware requirements for deployment.

0 favorites 0 likes
#safetensors

Vacuum 16T

Reddit r/LocalLLaMA · 2026-08-02

A deliberately empty 16.5-trillion-parameter model uploaded to Hugging Face exposes that parameter counts are computed from safetensors headers alone, and that Xet's content-defined deduplication reduces upload bandwidth enormously while storage quota still bills the full logical size.

0 favorites 0 likes
#safetensors

@ariG23498: I have always admired @stevhliu's work. I consider his technical writeups to be among the best there is. In the latest …

X AI KOLs Timeline · 2026-07-08 Cached

A thread highlighting a technical blog series on how Hugging Face's transformers library loads models efficiently, covering meta device, safetensors, CUDA caching, and more.

0 favorites 0 likes
#safetensors

GitHub - kallewoof/tftf: Transforming Transformers -- ultra light-weight pipeline for enormous transformer model manipulation with minimal overhead

Reddit r/LocalLLaMA · 2026-07-06 Cached

tftf is a lightweight, streaming pipeline for manipulating HuggingFace safetensors models, enabling FP8 dequantisation, LoRA merging, and other operations without loading the full model into memory, minimizing RAM and VRAM overhead.

0 favorites 0 likes
#safetensors

Comfy-Org/Boogu-Image

Hugging Face Models Trending · 2026-06-17 Cached

Comfy-Org has repackaged Boogu-Image model files for ComfyUI, including base, edit, and turbo variants with different quantization formats, plus a LoRA and text encoder.

0 favorites 0 likes
#safetensors

@LucSGeorges: perf packed release: safetensors 0.8.0 is out Main takeaways: - direct copy into metal MTLBuffers + dlpack for 0-copy h…

X AI KOLs Following · 2026-06-09 Cached

safetensors 0.8.0 release brings major performance improvements: direct copy into Metal MTLBuffers with dlpack for 2-3x faster loading and OOM fix on macOS, plus GIL-free serialization for faster multi-file saves.

0 favorites 0 likes
#safetensors

Qwen3.6-35B-A3B-Uncensored-Genesis-APEX-MTP

Reddit r/LocalLLaMA · 2026-05-24

A fine-tuned uncensored version of the Qwen model (Qwen3.6-35B-A3B) with MTP support and APEX quantization, tested stable at 200k context and recommended for use in LM Studio.

0 favorites 0 likes
#safetensors

HF flagged safetensors as unsafe? wtf?

Reddit r/LocalLLaMA · 2026-05-21

Hugging Face flagged a safetensors file as unsafe, confusing users who question the policy.

0 favorites 0 likes
#safetensors

MiniMax M2.7 ultra uncensored heretic is Out Now with 4/100 Refusals, Available in Safetensors and GGUFs Formats!

Reddit r/LocalLLaMA · 2026-05-15

The MiniMax M2.7 model has been fine-tuned into an uncensored variant, 'ultra uncensored heretic', with very low refusal rates (4/100). Available in Safetensors and GGUF formats on HuggingFace.

0 favorites 0 likes
#safetensors

Qwen3.6 35B A3B uncensored heretic Native MTP Preserved is Out Now With KLD 0.0015, 10/100 Refusals and the Full 19 MTPs Preserved and Retained, Available in Safetensors, GGUFs. NVFP4, NVFP4 GGUFs and GPTQ-Int4 Formats

Reddit r/LocalLLaMA · 2026-05-09

Community release of Qwen3.6 35B A3B uncensored variant with full 19 MTP tensors preserved, available in multiple formats including Safetensors, GGUF, NVFP4 and GPTQ-Int4.

0 favorites 0 likes
#safetensors

Safetensors is Joining the PyTorch Foundation

Hugging Face Blog · 2026-04-08 Cached

Safetensors has officially joined the PyTorch Foundation under the Linux Foundation to establish a vendor-neutral governance structure, while maintaining its status as the default model format for the Hugging Face Hub.

0 favorites 0 likes
← Back to home

Submit Feedback