training

Tag

Cards List
#training

Where Should Optimizer State Live? Tiered State Allocation for Memory-Efficient Mixture-of-Experts Training

Hugging Face Daily Papers ↗ · 2026-07-21 Cached

SkewAdam is a novel optimizer for mixture-of-experts models that tier allocates optimizer state across backbone, experts, and router, reducing memory footprint to 2.6% of AdamW while achieving better validation perplexity in controlled comparisons.

0 favorites 0 likes
#training

@a1zhang: Transformers struggle to generalize to tasks they were not explicitly trained on. Instead, we propose in 2026 that it i…

X AI KOLs Following ↗ · 2026-07-20 Cached

The article proposes that Transformers can generalize to new tasks through a well-designed harness that induces composition, without needing intrinsic model generalization. It shows RLMs can generalize from short tasks to 8-32x longer tasks and across domains.

0 favorites 0 likes
#training

@rasbt: How can an LLM switch between low-, medium-, and high-effort reasoning? And how does an LLM learn to reason more or les…

X AI KOLs Timeline ↗ · 2026-07-18 Cached

An article explaining how LLMs can switch between low, medium, and high effort reasoning during inference and training.

0 favorites 0 likes
#training

The industry keeps getting agentic security wrong, so I developed a free platform to teach what actually works

Reddit r/AI_Agents ↗ · 2026-07-15

The article introduces Tantalus, a free platform for learning how to secure agentic AI systems against indirect prompt injection and data exfiltration through realistic challenges.

0 favorites 0 likes
#training

@QuixiAI: @MicrosoftAI never published a BitNet trainer. I fixed that bug. https://github.com/QuixiAI/bitnet trained my own BitNe…

X AI KOLs Following ↗ · 2026-07-15 Cached

A user released a BitNet trainer that Microsoft never published, along with custom kernels for training and inference, while also highlighting Microsoft's bitnet.cpp inference framework for fast 1-bit LLM inference on CPUs and GPUs.

0 favorites 0 likes
#training

Velo 3.0

Product Hunt ↗ · 2026-07-14

Velo 3.0 is an AI video infrastructure platform designed to help explain, train, and sell faster.

0 favorites 0 likes
#training

@samsja19: We are also releasing prime-rl 0.7.0 which has full support for verifiers v1 and bring your own harness for training. W…

X AI KOLs Following ↗ · 2026-07-14 Cached

Prime Intellect released verifiers v1 and prime-rl 0.7.0, an RL training tool with full support for verifiers, multiple algorithms like GRPO and OPD, and performance improvements.

0 favorites 0 likes
#training

@Raman_bansal_: If you’ve trained an LLM, you’ve may have seen doom loop, in which a LLM endlessly repeats the same token or sentence, …

X AI KOLs Timeline ↗ · 2026-07-13 Cached

A Substack article explains the 'doom loop' problem in LLMs where models repeat tokens endlessly, and introduces Final Token Preference Optimization (FTPO) from Liquid AI as a method to detect and fix such loops during fine-tuning.

0 favorites 0 likes
#training

A practical recipe for building agent trajectory datasets

Reddit r/AI_Agents ↗ · 2026-07-13

A practical guide for building structured agent trajectory datasets for training tool-using agents, emphasizing the importance of designing trajectories with six key parts and treating them as data assets rather than logs.

0 favorites 0 likes
#training

@h100envy: Prime Intellect engineers explained how they train reasoning models over the open internet in 30 minutes - better than …

X AI KOLs Timeline ↗ · 2026-07-12 Cached

Prime Intellect engineers demonstrated a method to train reasoning models in 30 minutes using distributed RL over the open internet, utilizing Prime-RL, LLM judges, and multi-cloud GPUs, enabling open models to compete with closed labs without owning data centers.

0 favorites 0 likes
#training

@PrajwalTomar_: Fable 5 leaves your subscription tonight whatever you were saving it for, today is the day. lock the f*ck in. this one …

X AI KOLs Following ↗ · 2026-07-12 Cached

Fable 5 is leaving its subscription service tonight; users are urged to train a replacement model before midnight to retain its capabilities.

0 favorites 0 likes
#training

@FinanceYF5: Chubby shares his judgments on the "very near future." First, what was previously almost just a rumor has now been confirmed: GPT-5.6 actually finished training two months ago and has been opened for early access to some users. The obvious question is: why hasn't it been publicly released sooner?

X AI KOLs Timeline ↗ · 2026-07-11 Cached

According to Chubby, GPT-5.6 finished training two months ago and has been opened for early access to some users, but has not been publicly released.

0 favorites 0 likes
#training

AI harness with Computer Use and frontier models that perform - not a file/browser usage discussion - not an MCP discussion - but a model and training discussion only

Reddit r/AI_Agents ↗ · 2026-07-10

Discusses an AI harness that integrates computer use with frontier models, focusing on model capabilities and training rather than file/browser or MCP topics.

0 favorites 0 likes
#training

@VukRosic99: NVFP4 end-to-end training diverges. Current recipes patch around it with Hadamard transforms, stochastic rounding, high…

X AI KOLs Timeline ↗ · 2026-07-10 Cached

Four Over Six (4/6) introduces adaptive block scaling for NVFP4 quantization, reducing quantization error with minimal overhead, improving both training and post-training quantization for large language models.

0 favorites 0 likes
#training

LiST: Lipschitz Scaling Training for Robust and Calibrated Neural Networks

arXiv cs.LG ↗ · 2026-07-10 Cached

Introduces LiST, a training paradigm that uses Lipschitz constraints to achieve robust and calibrated neural networks, selecting optimal operating points on the accuracy-robustness Pareto front. Demonstrates competitive performance on CIFAR and Tiny-ImageNet.

0 favorites 0 likes
#training

@Huahuazo: Normally when using PyTorch's high-level API, everything works smoothly, but the moment you step away from the framework and write raw operators, you're quickly exposed — formulas swirl in your head but get stuck when turned into code; you can talk a good game, but freeze when you actually code. There's an open-source coding platform on GitHub called TorchCode, designed to fix this, turning "writing deep learning operators by hand" into a LeetCode-style practice...

X AI KOLs Timeline ↗ · 2026-07-09 Cached

TorchCode is an open-source coding platform that turns manual deep learning operator implementation into a LeetCode-style experience. It includes 40 high-frequency interview questions, provides automated evaluation and hints, and supports one-click Docker deployment and online use via Hugging Face Spaces.

0 favorites 0 likes
#training

@QCXINT_: Most people stop after building a chatbot. Production AI systems are a completely different game. This open-source comp…

X AI KOLs Timeline ↗ · 2026-07-09 Cached

An open-source companion to the LLM Engineer's Handbook that provides a complete blueprint for building production-ready LLM systems, covering synthetic data generation, training (including DPO), RAG, deployment on AWS, evaluation, and monitoring.

0 favorites 0 likes
#training

@yunta_tsai: If the first thing of the new hire did was not springing up GPU for training but obsessing with labeling, then you hire…

X AI KOLs Timeline ↗ · 2026-07-08

A tweet emphasizing that a new hire who focuses on data labeling rather than immediately setting up GPU training is a good sign for a machine learning role.

0 favorites 0 likes
#training

Is AI trained to lie?

Reddit r/ArtificialInteligence ↗ · 2026-07-07

An exploration of whether AI systems are trained to be deceptive, raising concerns about AI safety and ethics.

0 favorites 0 likes
#training

TorchJD: Training with multiple losses in PyTorch [P]

Reddit r/MachineLearning ↗ · 2026-07-07

TorchJD is a library for training models with multiple losses in PyTorch, implementing both scalarization and Jacobian descent methods. It has been accepted into the PyTorch ecosystem and aims to become the go-to library for multi-loss training.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback