training

Tag

Cards List
#training

The $28,000 Course for an AI Job Nobody Quite Understands Yet

Reddit r/ArtificialInteligence ↗ · 2026-08-22

This article discusses a high-priced course aimed at preparing individuals for emerging AI jobs that are not yet fully defined, highlighting the growing demand and uncertainty in the AI job market.

0 favorites 0 likes
#training

@percyliang: Marin 535B-A23B started training this week! As usual, the whole process is open. Voyage plan: pretraining (80%) + midtr…

X AI KOLs Following ↗ · 2026-08-21 Cached

Training of the Marin 535B-A23B AI model has begun with an open process, involving pretraining and midtraining on 18.75T tokens using GB200 NVL72 hardware over about 3 months.

0 favorites 0 likes
#training

@modal: Apply to attend.

X AI KOLs Following ↗ · 2026-08-21 Cached

Runtime by Modal is a conference event for AI and ML engineers to discuss running AI in production, featuring talks on inference, training, and agents.

0 favorites 0 likes
#training

@mattpocockuk: FYI Folks who are wanting to expense my course, but not sure how to convince your boss I have a boss letter you can cop…

X AI KOLs Timeline ↗ · 2026-08-21 Cached

Matt Pocock shares a template letter for developers to convince their bosses to invest in his AI Hero course, which teaches a structured engineering process for using AI coding assistants effectively.

0 favorites 0 likes
#training

@timodonnell: Want to watch a 535B parameter (23B active) LLM get trained live? Follow along here https://wandb.ai/marin-community/ma…

X AI KOLs Timeline ↗ · 2026-08-20

A tweet announces the live training of a 535B parameter (23B active) large language model, with links to follow the process on Weights & Biases and GitHub.

0 favorites 0 likes
#training

Same GRPO recipe on three from-scratch LLMs (353M/316M/672M) gave three different outcomes, with no clean relationship to scale [P]

Reddit r/MachineLearning ↗ · 2026-08-19

An experiment training three from-scratch LLMs with the same GRPO recipe yielded inconsistent results, with GRPO degrading performance in some models, particularly the middle-sized one, and no clear relationship to scale.

0 favorites 0 likes
#training

LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents

arXiv cs.AI ↗ · 2026-08-19 Cached

LEGO-RL presents a framework to bridge native coding-agent harnesses with scalable policy-gradient reinforcement learning, improving performance on benchmarks like SWE-bench.

0 favorites 0 likes
#training

what would you actually train a browser-agent model to be good at?

Reddit r/AI_Agents ↗ · 2026-08-18

The article discusses challenges in training browser-agent models for sequential decision-making, such as error recovery and memory, and seeks input on optimizing training objectives. The author mentions working on the 'mako' model at tinyfish and invites community feedback.

0 favorites 0 likes
#training

@gkcs_: Registrations for the next AI Engineering Cohort are on! Dates: 12th September - 1st November Time: Saturday & Sunday, …

X AI KOLs Timeline ↗ · 2026-08-15 Cached

Registrations are open for an 8-week AI Engineering Cohort focusing on RAG & Agents, aimed at software engineers upskilling into AI roles.

0 favorites 0 likes
#training

Training Under Challenge: Executable Certificates and Challenge-Closed Optimality for Neural Networks

arXiv cs.LG ↗ · 2026-08-14 Cached

Introduces 'Training Under Challenge', an executable-certificate framework that uses architecture-valid procedures to construct alternative models and estimate the empirical global-optimality gap of neural network checkpoints, with theoretical guarantees and experiments on ResNet-18 distillation and quantized denoising.

0 favorites 0 likes
#training

@0xRicker: ANDREJ KARPATHY JUST BUILT A LLM FROM SCRATCH AND OPEN-SOURCED ALL CODE And it still contains the full LLM pipeline: → …

X AI KOLs Timeline ↗ · 2026-08-13 Cached

Andrej Karpathy built an LLM from scratch and open-sourced all code, covering the full pipeline of dataset, training loop, and inference loop, continuing his work on simplifying AI education with micrograd and makemore.

0 favorites 0 likes
#training

@wu_taiqiang: How to maximize OPD performance? One important thing is warm-up. Then the student-sampled sequence is well defined in t…

X AI KOLs Following ↗ · 2026-08-13 Cached

The author discusses a paper that demystifies the warm-up process for OPD (likely on-policy distillation), explaining how warm-up enables well-defined student-sampled sequences and educational token-level dense rewards from the teacher.

0 favorites 0 likes
#training

Adaptive Supervised Anchoring for On-Policy Self-Distillation

arXiv cs.LG ↗ · 2026-08-11 Cached

This paper proposes an adaptive supervised anchoring framework for on-policy self-distillation, addressing the problem of rollout-conditioned signal degradation in language model training. The method separates rollout-conditioned distribution matching from canonical-context supervision, improving task acquisition while preserving general capabilities.

0 favorites 0 likes
#training

Making Knowledge Distillation Cheap Enough to Run at Scale

Hugging Face Blog ↗ · 2026-08-10 Cached

Multiverse Computing announces a paper on making LLM knowledge distillation cheaper via offline top-K logits and a fused chunked KL loss, cutting VRAM usage for distillation at scale.

0 favorites 0 likes
#training

Pathway's BDH(post-transformer arch) matches GPT2 scaling from 10M to 1B params trained from scratch. runs on Normal GPUs

Reddit r/LocalLLaMA ↗ · 2026-08-09

Pathway's BDH, a post-transformer architecture, reportedly matches GPT-2 scaling from 10M to 1B parameters while training from scratch on standard GPUs.

0 favorites 0 likes
#training

@samsja19: prime rl can now express and train multi agent systems, enabling usecase like adjentic judge, self play, user simulatio…

X AI KOLs Following ↗ · 2026-08-07 Cached

Prime RL now supports expressing and training multi-agent systems, enabling use cases like agentic judge, self-play, user simulation, and complex agent collaboration.

0 favorites 0 likes
#training

@teslaeurope: FSD Supervised has been trained & tested on 2.2 million km across 19 EU countries, handling edge cases most people woul…

X AI KOLs Following ↗ · 2026-08-07 Cached

Tesla Europe announces FSD Supervised has been trained and tested on 2.2 million km across 19 EU countries, handling rare edge cases.

0 favorites 0 likes
#training

ByteDance is at an early stage of training a model with as many as 10 trillion parameters

Reddit r/singularity ↗ · 2026-08-07

ByteDance is in the early stages of training a large language model with up to 10 trillion parameters, signaling a massive scale-up in AI development.

0 favorites 0 likes
#training

@gmkurtzer: I just listened to this video from @latkins who is CTO of @arcee_ai talking about lessons learns of training a large MO…

X AI KOLs Following ↗ · 2026-08-07 Cached

Gregory Kurtzer praises Lucas Atkins' talk on lessons learned from training a large sparse Mixture-of-Experts model and life at an AI lab startup.

0 favorites 0 likes
#training

ByteDance trains a 10-trillion-parameter AI model, aiming for global leadership (1 minute read)

TLDR AI ↗ · 2026-08-07 Cached

据金融时报报道,字节跳动正在训练一个估算参数量达10万亿级别的AI大模型,规模接近Anthropic的先进系统,旨在缩小与美国顶级AI实验室的差距。

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback