hindsight

Tag

Cards List
#hindsight

Built a customer-support AI agent with persistent memory using Hindsight

Reddit r/AI_Agents ↗ · yesterday

Built a customer-support AI agent named SupportMemory using Hindsight for persistent memory to recall and retain customer context, improving personalization in interactions.

0 favorites 0 likes
#hindsight

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning

Hugging Face Daily Papers ↗ · 2026-08-04 Cached

TurnSight introduces a turn-level hindsight self-distillation framework for tool-integrated reasoning, providing dense supervision via execution-conditioned hindsight and adaptive RL advantage modulation.

0 favorites 0 likes
#hindsight

From Outcomes to Actions: Leveraging Hindsight for Long-Horizon Language Agent Training

arXiv cs.LG ↗ · 2026-07-21 Cached

Introduces Hindsight Policy Optimization (HPO), a novel policy gradient method that uses an intent space and Wasserstein distance to reduce variance in long-horizon language agent training, showing improved stability over GRPO and PPO.

0 favorites 0 likes
#hindsight

What If Your AI Actually Remembered Every Meeting? I Built MeetMemory with Hindsight to Find Out.

Reddit r/AI_Agents ↗ · 2026-06-14

MeetMemory is a tool that gives AI permanent memory for meetings, built on Hindsight and Groq, allowing instant recall across all past conversations without manual search or note-taking.

0 favorites 0 likes
#hindsight

HERO: Hindsight-Enhanced Reflection from Environment Observations for Agentic Self-Distillation

arXiv cs.AI ↗ · 2026-06-11 Cached

HERO introduces a hindsight-enhanced self-distillation framework that uses environment observations as locally aligned feedback to improve multi-turn agent capabilities, outperforming existing methods on TauBench and WebShop, especially under limited turn budgets.

0 favorites 0 likes
#hindsight

HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents

Hugging Face Daily Papers ↗ · 2026-05-18 Cached

HINT-SD proposes a targeted self-distillation framework that selects failure-relevant actions from full trajectories to improve long-horizon LLM agent training, achieving up to 18.80% improvement and 2.26× speedup over dense feedback baselines.

0 favorites 0 likes
← Back to home

Submit Feedback