adaptive

Tag

Cards List
#adaptive

FinSMART: Financial Sentiment Analysis for Algorithmic Trading through Market-Aligned Reinforcement Learning

arXiv cs.CL · yesterday Cached

FinSMART introduces a market-aligned reinforcement learning framework for financial sentiment analysis, optimizing sentiment signals with realized market outcomes and achieving a 220% improvement in cumulative trading returns over the strongest baseline.

0 favorites 0 likes
#adaptive

MixQuant: Adaptive Mixed-Precision Quantization for Large Language Models

arXiv cs.LG · 4d ago Cached

MixQuant proposes an adaptive mixed-precision quantization framework for LLMs that handles variable memory budgets by marginalizing layer distortion over random upstream configurations, outperforming existing methods across multiple models and budgets.

0 favorites 0 likes
#adaptive

Adaptive Multi-Horizon Reinforcement Learning

arXiv cs.LG · 2026-07-24 Cached

This paper proposes a multi-horizon reinforcement learning approach that adaptively selects and combines temporal horizons, enabling robust adaptation to changing reward structures without manual discount factor tuning, with empirical validation in MiniGrid environments.

0 favorites 0 likes
#adaptive

@_markfenner: On today's episode of Devinmaxxing: pick your model from your pocket. DevinX now does full model selection for local se…

X AI KOLs Following · 2026-07-12 Cached

DevinX now supports full model selection for local sessions, including Sol, Fable 5, GLM, Kimi, and an adaptive cost-balancing option, plus reasoning effort control.

0 favorites 0 likes
#adaptive

@yingwww_: Warm take: Your world model should never stop learning Introducing AdaJEPA, an adaptive WM that plans, acts, and adapts…

X AI KOLs Following · 2026-07-05 Cached

AdaJEPA introduces an adaptive latent world model that continuously updates during test-time via closed-loop model predictive control, significantly improving planning success under distribution shift.

0 favorites 0 likes
#adaptive

Building an AI that remembers, adapts, and becomes more useful over time. A real partner not just an assistant or a tool.

Reddit r/AI_Agents · 2026-07-04

Describes the development of an AI that remembers, adapts, and becomes more useful over time, positioning it as a real partner rather than just an assistant or tool.

0 favorites 0 likes
#adaptive

@NFTCPS: HarnessX is pretty interesting: an agent architecture that can modify itself. Previously, architectural changes relied entirely on manual tuning. When a new model came out, Anthropic removed the planning steps from Claude Code, and Manus refactored its agents five times in six months, each time simplifying. What to change and when to change it — all decided by humans.

X AI KOLs Timeline · 2026-06-17 Cached

HarnessX introduces a framework for self-evolving AI agent harnesses that treats the runtime harness as a first-class object, enabling automatic adaptation via trace-driven reinforcement learning. It achieves average gains of +14.5% across five benchmarks, with larger improvements for weaker models.

0 favorites 0 likes
#adaptive

HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry

Hugging Face Daily Papers · 2026-06-12 Cached

HarnessX is a foundry for composable, adaptive, and evolvable AI agent harnesses that uses compositional primitives and trace-driven evolution to improve agent performance. Across five benchmarks, it achieves an average gain of +14.5% (up to +44.0%), demonstrating that runtime interface evolution is a complementary lever to model scaling.

0 favorites 0 likes
#adaptive

Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models

Hugging Face Daily Papers · 2026-06-10 Cached

SKIM is an adaptive multi-resolution soft token compression framework that compresses procedural skills for LLMs, maintaining task performance while reducing prefill cost and latency.

0 favorites 0 likes
#adaptive

AdaPLD: Adaptive Retrieval and Reuse for Efficient Model-Free Speculative Decoding

arXiv cs.CL · 2026-06-05 Cached

AdaPLD is a training-free method that improves model-free speculative decoding by using adaptive retrieval combining lexical and semantic similarity, and constructing branched reuse hypotheses to handle continuation uncertainty, achieving up to 3.10x decoding speedup.

0 favorites 0 likes
#adaptive

CosmicFish-HRM: Adaptive Reasoning via Hierarchical Recurrent Mechanisms in Compact Language Models

arXiv cs.LG · 2026-05-29 Cached

This paper presents CosmicFish-HRM, a compact 82.77M parameter language model with a hierarchical reasoning module that dynamically allocates reasoning compute during inference, learning when to halt based on input complexity.

0 favorites 0 likes
#adaptive

Consistently Informative Soft-Label Temperature for Knowledge Distillation

arXiv cs.LG · 2026-05-21 Cached

Proposes CIST, a method that assigns separate sample-wise adaptive temperatures to teacher and student in knowledge distillation, producing consistently informative soft labels and relaxing rigid logit-scale matching. Experiments on vision and language tasks show consistent improvements over standard KD.

0 favorites 0 likes
#adaptive

Not All Tokens Are Worth Caching: Learning Semantic-Aware Eviction for LLM Prefix Caches

arXiv cs.LG · 2026-05-20

A new semantic-adaptive eviction policy for LLM prefix caches that learns token reuse patterns across different token types, achieving 1.4x-2.7x TTFT improvement over existing policies.

0 favorites 0 likes
← Back to home

Submit Feedback