arxiv-paper

Tag

Cards List
#arxiv-paper

@askalphaxiv: GPT 5.6 Sol can one-shot convert an arXiv paper into an interactive Marimo notebook! Great for papers best understood h…

X AI KOLs Timeline · 2026-07-14 Cached

GPT 5.6 Sol can one-shot convert an arXiv paper into an interactive Marimo notebook, useful for hands-on understanding of papers in interpretability, inference engineering, and more.

0 favorites 0 likes
#arxiv-paper

FirstResearch: Auditable Question Formation for LLM Scientific Discovery Agents

arXiv cs.AI · 2026-07-08 Cached

FirstResearch introduces a structured framework for LLM scientific discovery agents that generates a Research Question Certificate containing primitive definitions, assumptions, mechanism, falsifiable hypothesis, and failure update rules, making the proposed research question inspectable before execution. Preliminary evaluations using LLM judges show that the certificate-centered approach outperforms baseline systems in audibility and score.

0 favorites 0 likes
#arxiv-paper

World-Model Collapse as a Phase Transition

arXiv cs.AI · 2026-07-01 Cached

研究语言智能体在长期任务中世界模型塌缩的相变现象,发现状态负载和依赖密度等参数在临界点附近导致模型突然崩溃,而非逐渐退化。

0 favorites 0 likes
#arxiv-paper

@omarsar0: If you build with MCPs, this one is worth reading. (bookmark it) The paper covers five recurring MCP server patterns ac…

X AI KOLs Following · 2026-06-30 Cached

This paper catalogues five recurring MCP server architectural patterns observed across fifteen independently developed servers, providing a taxonomy with context, problem, solution, and consequences. It also documents anti-patterns, cross-cutting concerns, and quantitative evaluations including inter-rater reliability and transport overhead.

0 favorites 0 likes
#arxiv-paper

S-GAI: Spectral Geometry-Aware Initialization for Sigmoidal MLPs -- From Dataset Geometry to Network Weights

arXiv cs.LG · 2026-06-30 Cached

S-GAI is a spectral geometry-aware initialization framework for one-hidden-layer sigmoidal MLPs that uses class-wise spectral geometry from image data to initialize weights, outperforming random initialization in terms of starting hidden state quality and achieving comparable final accuracy on benchmarks like MNIST and CIFAR-10.

0 favorites 0 likes
#arxiv-paper

@a1zhang: RLM arXiv paper update: depth>1 results, more comparisons, more training, and more error analysis! We add depth=2/3 exp…

X AI KOLs Following · 2026-05-12

This update to the RLM arXiv paper adds depth>1 experiments with recursive RLM calls, showing significant performance gains on OOLONG-Pairs and other benchmarks, along with new comparisons to OpenCode and Claude Code, additional training results on MRCRv2, and an expanded error analysis.

0 favorites 0 likes
#arxiv-paper

@oshaikh13: very cool idea @OpenAI I’m really excited about this research preview- learning from how people interact with their com…

X AI KOLs Following · 2026-04-20

An OpenAI research preview explores learning from how people interact with their computers beyond chat, accompanied by a new arxiv paper on the topic.

0 favorites 0 likes
← Back to home

Submit Feedback