lora

Tag

Cards List
#lora

One Rate Is Not Enough: Adaptive Anisotropic Learning Rates for LoRA Fine-Tuning

arXiv cs.LG ↗ · 2026-09-10 Cached

This paper introduces an adaptive anisotropic learning-rate model for LoRA fine-tuning to address within-module heterogeneity, improving performance and rank capacity utilization across benchmarks.

0 favorites 0 likes
#lora

Async GRPO with LoRA across HF Jobs: a bucket, a proxy, and no NCCL

Hugging Face Blog ↗ · 2026-09-10 Cached

This article describes a method to train LoRA adapters using AsyncGRPOTrainer and sync them via Storage Buckets across separate Hugging Face Jobs, eliminating the need for NCCL communication.

0 favorites 0 likes
#lora

X-AuT: Progressive Audio-Encoder Compression for Speech LLMs with Cross-Scale Distillation

Hugging Face Daily Papers ↗ · 2026-09-10 Cached

X-AuT is a progressive framework for compressing audio-encoder layers in speech large language models, reducing inference cost while restoring accuracy via techniques like cross-scale distillation and LoRA adaptation.

0 favorites 0 likes
#lora

@emilyzsh: 𝓑𝓤𝓕𝓞 i trained a bufo lora to automate fast and style-consistent bufo creation in our slack -- opening up inference…

X AI KOLs Following ↗ · 2026-09-07 Cached

A user trained a LoRA model to automate the creation of style-consistent 'bufos' in Slack, opening up inference for public use with a dedicated website.

0 favorites 0 likes
#lora

Alissonerdx/Minimax-H3-ComfyUI

Hugging Face Models Trending ↗ · 2026-09-06 Cached

This repository provides LoRAs for the MiniMax H3 model, designed to run in ComfyUI for video enhancement, such as sharpening videos while maintaining photorealism.

0 favorites 0 likes
#lora

Routing Is Not Enough: Diagnosing Intra-Adapter Subspace Contention in MoE+LoRA Fine-Tuning

arXiv cs.LG ↗ · 2026-09-04 Cached

This paper diagnoses intra-adapter contention in MoE+LoRA fine-tuning and introduces SpawnLoRA to dynamically add sub-adapters, reducing negative transfer across domains.

0 favorites 0 likes
#lora

Can a 4B local model actually feel like an AI assistant?

Reddit r/LocalLLaMA ↗ · 2026-09-03

A developer is experimenting with building an AI assistant called Arcon using a 4B local model with LoRA, incorporating persistent memory and personality features, and seeks community feedback.

0 favorites 0 likes
#lora

Federated LoRA Adaptation of BiomedCLIP Across Four International Chest X-Ray Cohorts

arXiv cs.LG ↗ · 2026-09-03 Cached

This paper benchmarks federated LoRA adaptation of BiomedCLIP for chest X-ray classification across four international cohorts, demonstrating improved performance over unadapted models and approaching centralized training results.

0 favorites 0 likes
#lora

TalkFa: A Unified Benchmark for Farsi Dialogue Generation and Understanding

arXiv cs.CL ↗ · 2026-09-03 Cached

TalkFa is a unified benchmark for Farsi dialogue generation and understanding, consisting of three datasets validated through experiments with LLMs and human evaluation.

0 favorites 0 likes
#lora

SpeakPay: Domain-Adaptive LoRA Fine-Tuning of Whisper for Low-Resource Nepali Financial Speech Recognition

arXiv cs.CL ↗ · 2026-09-03 Cached

This paper introduces SpeakPay and a Nepali financial speech dataset, showing that LoRA fine-tuning of Whisper reduces Word Error Rate by 67.2% and improves transaction success rates for low-resource language accessibility.

0 favorites 0 likes
#lora

Breaking the Structural Identity: Personalized Federated LoRA Fine-tuning under Rank Heterogeneity

arXiv cs.LG ↗ · 2026-09-02 Cached

FedRoRA is a novel framework for personalized federated LoRA fine-tuning that addresses rank heterogeneity and data heterogeneity in federated learning by decoupling adaptation into shared global directions and personalized magnitudes.

0 favorites 0 likes
#lora

RW-LoRA: Communication-Efficient Decentralized LoRA Fine-Tuning via Random Walks

arXiv cs.LG ↗ · 2026-09-02 Cached

This paper introduces RW-LoRA, a decentralized LoRA fine-tuning method using random walks to reduce communication and computation costs while achieving competitive performance on NLP tasks.

0 favorites 0 likes
#lora

Made my first fine tune!

Reddit r/ArtificialInteligence ↗ · 2026-09-01 Cached

RustEAI is a local AI-powered Rust coder for Apple Silicon Macs, based on a LoRA fine-tune of Qwen2.5-Coder-0.5B, providing one-shot code generation without cloud dependency.

0 favorites 0 likes
#lora

Normalized Low-Rank Adaptation

Hugging Face Daily Papers ↗ · 2026-08-31 Cached

Normalized Low-Rank Adaptation (NoRA) stabilizes LoRA training by normalizing down-projection matrices, accelerating convergence and improving performance without extra parameters or inference cost.

0 favorites 0 likes
#lora

Acquire, Repair, Preserve: A Diagnosis-Guided Post-Training Recipe for Small-Model Dialogue Game Agents

Hugging Face Daily Papers ↗ · 2026-08-28 Cached

This paper investigates failures in a 2B model for dialogue games and introduces a diagnosis-guided post-training recipe using SFT, DPO, and LoRA to boost performance while maintaining general capabilities.

0 favorites 0 likes
#lora

@techNmak: Everyone is fine-tuning LLMs. Almost nobody understands what is actually being updated inside the model. Here are 5 tec…

X AI KOLs Timeline ↗ · 2026-08-26 Cached

This article explains five parameter-efficient fine-tuning techniques for large language models, such as LoRA and VeRA, detailing how each method adapts model weights with minimal updates.

0 favorites 0 likes
#lora

alibaba-pai/MiniMax-H3-Acc-LoRAs

Hugging Face Models Trending ↗ · 2026-08-26 Cached

Alibaba PAI releases LoRA checkpoints for accelerating MiniMax-H3 video generation using Parallel Decoding Distillation, enabling efficient inference in 8 steps.

0 favorites 0 likes
#lora

FCPRAG: Fusion-Controller Parametric Retrieval-Augmented Generation for Stable Multi-Passage LoRA Injection

arXiv cs.CL ↗ · 2026-08-25 Cached

FCPRAG proposes a fusion-controller framework for parametric retrieval-augmented generation that improves stability and performance in multi-passage LoRA injection, showing consistent gains over baselines in experiments.

0 favorites 0 likes
#lora

Efficient Adaptation of LLMs for Hate Speech Detection in Low-Resource Languages: A Comparative Study on Roman Urdu

arXiv cs.AI ↗ · 2026-08-20 Cached

The paper evaluates Large Language Models for hate speech detection in Roman Urdu, a low-resource language, demonstrating that Parameter-Efficient Fine-Tuning with LoRA significantly improves classification performance compared to zero-shot inference.

0 favorites 0 likes
#lora

@HuggingApps: one prompt in, a shot plan out sbgrid is a Krea 2 Turbo LoRA that creates an 8-panel storyboard as a single image, from…

X AI KOLs Timeline ↗ · 2026-08-16 Cached

sbgrid is a Krea 2 Turbo LoRA that creates an 8-panel storyboard as a single image from one prompt, hosted as a Hugging Face Space.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback