research-framework

Tag

Cards List
#research-framework

From Symbolic Perception to Logical Deduction: A Framework for Guiding Language Models in Geometric Reasoning

arXiv cs.AI · 4h ago Cached

This paper presents a framework that augments Large Language Models with geometric vision parsing and symbolic solving to match state-of-the-art multimodal models on complex geometry problems, using a new benchmark from 2025 Chinese Zhongkao exams for evaluation.

0 favorites 0 likes
#research-framework

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

Hugging Face Daily Papers · 3d ago Cached

This paper introduces Procedural Graph, a framework that organizes LLM agent actions into structured triplets for improved long-horizon tool use, with self-evolving topology to enhance performance.

0 favorites 0 likes
#research-framework

PragAlign: Feedback-Guided Pragmatic Alignment for Controlled Synthetic Dialogue Generation

arXiv cs.CL · 2026-09-03 Cached

PragAlign introduces a feedback-guided framework for controlled synthetic dialogue generation using an LLM-based evaluator to improve alignment with intent, emotion, coherence, and fluency, achieving 99.50% acceptance compared to 72.25% for one-shot generation.

0 favorites 0 likes
#research-framework

EvoUndo: Recoverability-Constrained Self-Evolution for LLM Agent Harnesses

Hugging Face Daily Papers · 2026-08-28 Cached

EvoUndo introduces a framework for evaluating and ensuring recoverability in self-modifying LLM agents, showing that reliable recovery requires co-designing verification, state grounding, and recovery language expressivity.

0 favorites 0 likes
#research-framework

Google bid $10M for Spirit’s data. Under one scenario, it needs just 0.236% economic uplift to break even

Reddit r/ArtificialInteligence · 2026-08-26

The article presents a framework for valuing proprietary data in AI systems, using Google's $10M bid for Spirit's data to calculate the economic uplift needed for break-even.

0 favorites 0 likes
#research-framework

@dair_ai: New research from NVIDIA. They just dropped a PyTorch-native training framework for agentic RL. (bookmark it) Paper sum…

X AI KOLs Following · 2026-07-27 Cached

NVIDIA released Molt, a PyTorch-native agentic RL framework designed for compactness and readability, with performance comparable to Megatron-based stacks. The framework is open-source and includes a paper.

0 favorites 0 likes
#research-framework

@AYi_AInotes: Two Hong Kong students achieved a 5x speedup on Karpathy's automated research framework. They didn't switch to a stronger model, add more compute, or even change much code. They just added another loop on top of the original loop. This might be the most useful paper of the year for ordinary Agent developers, bar none. Here's the breakdown:

X AI KOLs Timeline · 2026-07-10 Cached

Two Hong Kong students achieved a 5x speedup by adding another loop outside the original automated research framework, without needing a better model or more compute. It is considered one of the most useful papers for ordinary Agent developers.

0 favorites 0 likes
#research-framework

@Jason23818126: After earning +1.46 million in two years, someone open-sourced their research system. The project is called AI Berkshire, and the real trading results are directly released: 2024: +69.29% 2025: +66.38% The core is not "AI stock picking", but breaking down the investment frameworks of Buffett, Munger, Duan Yongping, Li Lu into...

X AI KOLs Timeline · 2026-07-05 Cached

This project open-sources an AI research system based on the frameworks of value investing masters like Buffett, Munger, Duan Yongping, and Li Lu. It uses Claude Code/Codex to enable multi-agent parallel analysis of financial statements, valuations, etc., and shows real trading returns of over 1.46 million yuan in two years, significantly outperforming major indices.

0 favorites 0 likes
#research-framework

@degenrsc: https://x.com/degenrsc/status/2064714047241736302

X AI KOLs Timeline · 2026-06-10 Cached

A detailed guide on building an agentic research framework using a multi-LLM system with persistent memory, allowing researchers to avoid re-explaining context across sessions by leveraging file-based identity, project docs, and memory indices.

0 favorites 0 likes
#research-framework

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing

arXiv cs.CL · 2026-06-05 Cached

This paper introduces PersuasionTrace, a framework for studying multi-turn persuasion in human-LLM interaction, using a Bayesian-network simulated target that models belief updates. The framework reveals that LLMs are persuasive across topics and modalities, and that the Bayesian target better matches human belief dynamics than vanilla LLM simulators.

0 favorites 0 likes
#research-framework

@yibie: This week's autoresearch ecosystem evidence scan: 9 new records, total count 383. AutoResearch-RL: A continuous RL research framework with http://prepare.py/train.py isolation, supporting LLM/hybrid strategy experiment scheduling l…

X AI KOLs Timeline · 2026-05-26 Cached

This week, 9 new records were added to the autoresearch ecosystem, bringing the total to 383, covering multiple open-source tools and projects such as the AutoResearch-RL reinforcement learning framework, lance-autoresearch database kernel optimization, and Clio prediction market backtesting framework.

0 favorites 0 likes
← Back to home

Submit Feedback