model-conversion

Tag

Cards List
#model-conversion

Run any model, on any backend (Website)

TLDR AI · 2026-09-10 Cached

ZeroModels is a tool that enables running any AI model on multiple backends (JAX, PyTorch, TensorFlow) using Keras 3, with weights converted from original checkpoints and no runtime dependency on transformers or torch.

0 favorites 0 likes
#model-conversion

Kijai/MiniMax-H3_comfy

Hugging Face Models Trending · 2026-08-07 Cached

A Hugging Face repository hosting MiniMax-H3 models converted for ComfyUI usage, along with a Lightx2v distill LoRA for faster inference.

0 favorites 0 likes
#model-conversion

Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't

arXiv cs.LG · 2026-08-05 Cached

This paper studies what transfers between transformer models of different sizes in the same family (Pythia), showing that representations align while weights don't, and that conversion works best via initialization rather than direct weight projection.

0 favorites 0 likes
#model-conversion

Beyond KV Reconstruction: Functional Reconstruction for MLA Draft Models in Speculative Decoding

arXiv cs.LG · 2026-07-31 Cached

This paper proposes functional reconstruction for converting MHA/GQA checkpoints into MLA draft models for speculative decoding, directly optimizing attention modules to preserve token acceptance. It reports consistent improvements across 192 configurations involving Llama/Qwen models and multiple conversion methods.

0 favorites 0 likes
#model-conversion

@Italianclownz: Converted Qwen 3.6 35b a3b to ROCmfp4 and this is flying. Used the mtp version bc this ROCmfp4 can also incorporate the…

X AI KOLs Timeline · 2026-05-24 Cached

Converted the Qwen 3.6 35b a3b model to ROCmfp4 format, leveraging MTP benefits for improved performance on AMD hardware.

0 favorites 0 likes
#model-conversion

Extracted MTP tensor GGUFs - smaller donor models for grafting.

Reddit r/LocalLLaMA · 2026-05-07

The author provides extracted GGUF files containing only MTP tensors for Qwen3.6 models, allowing users to graft tensors with a significantly reduced download size compared to full model files.

0 favorites 0 likes
← Back to home

Submit Feedback