Papers

Cards List

Accurate Models of AMD Matrix Cores

Hacker News Top · 2h ago Cached

This paper presents accurate software models of AMD GPU matrix cores for CDNA 1/2/3 architectures, validated for bit-level reproducibility against hardware, and demonstrates their use in numerical applications to compare accuracy with NVIDIA tensor cores.

0 favorites 0 likes

GoBench: Evaluating LLMs on the game of Go [R]

Reddit r/MachineLearning · 2h ago

GoBench is a benchmark for evaluating large language models on 9x9 Go games, demonstrating strong correlation with ARC-AGI and featuring a leaderboard with KataGo opponents from random to superhuman levels.

0 favorites 0 likes

ER visits for gambling disorders doubled after expanded online gambling market

Hacker News Top · 4h ago Cached

A study found that emergency department visits for gambling disorders nearly doubled in Ontario after the expansion of online gambling, particularly among young men, indicating increased harm from legalized sports betting.

0 favorites 0 likes

Training Text-to-Image Models 3.6× Faster

Hacker News Top · 4h ago Cached

Linum AI introduces JiT-DDT, a novel encoder-decoder architecture that trains text-to-image models 3.6× faster than previous methods while generating images with higher resolution.

0 favorites 0 likes

@dair_ai: Nice paper discussing context trimming for agents. This is a hot topic at the moment, so it might be worth your time. C…

X AI KOLs Timeline · 5h ago Cached

This paper compares context trimming strategies for AI agents, finding that protocol-aware trimming with adaptive budget guardrails maintains high task success while reducing tokens, though it relies on gold annotations for implementation.

0 favorites 0 likes

What happens when neutrinos swap identities inside a supernova?

Ars Technica · 6h ago Cached

A paper in Physical Review D suggests that neutrinos changing flavor inside supernovae could carry energy away, potentially explaining discrepancies in core-collapse supernova models and observed rates.

0 favorites 0 likes

Meet a mouse whose brain cortex is made up of human cells

MIT Technology Review · 6h ago Cached

Researchers led by neuroscientist Sergiu Pașca have created mice with human brain cells transplanted into genetically modified cortices, showing improved cognition in maze tests and raising ethical concerns about cross-species brain research.

0 favorites 0 likes

LARA: small, composable behaviours for frozen LLMs [P]

Reddit r/MachineLearning · 8h ago

LARA is a research project and PyTorch library that enables modular, composable behaviors for frozen large language models using low-rank residual adapters, allowing efficient training and inference-time blending of multiple behaviors.

0 favorites 0 likes

Fine Tuned and Niche Ai Models vs Ai Detectors

Reddit r/artificial · 11h ago

The article critiques AI detectors like Pangram and GPTZero for their low accuracy against fine-tuned AI models and references a study indicating readers prefer AI outputs trained on copyrighted books.

0 favorites 0 likes

‘Smart’ Nanoparticles Deliver mRNA Directly to Tumors in New Cancer Therapy

Wired · 13h ago Cached

Researchers developed smart nanoparticles to deliver mRNA directly to tumor-associated macrophages, reprogramming them to enhance immune response and slow tumor growth in mouse models of breast cancer.

0 favorites 0 likes

A 3-Year-Old's Metastatic Cancer Disappeared After Two Doses of Experimental Therapy: Complete regression of hepatoblastoma, a malignant liver cancer and chemotherapy-resistant solid tumor, after CAR T cell treatment, achieved entirely in the outpatient setting without systemic toxicity.

Reddit r/singularity · 14h ago Cached

A 3-year-old boy with metastatic hepatoblastoma achieved complete cancer regression after two doses of experimental CAR T cell therapy, administered outpatient without systemic toxicity.

0 favorites 0 likes

I measured memory vs "just send the whole history" over 90 simulated days: 23-62x fewer context tokens, same or better recall on personal facts, and one place where memory clearly loses (numbers + method)

Reddit r/AI_Agents · 14h ago

This study compares memory systems to full conversation history in AI agents over simulated days, showing 23-62x fewer context tokens with similar or better recall on personal facts, but memory loses on numerical data and specific details like identifiers.

0 favorites 0 likes

How chimps teach their kids tool tricks

Ars Technica · 17h ago Cached

Chimpanzees learn tool use through social teaching, where adults model behavior and transfer tools to offspring, as per a new study in Frontiers in Psychology.

0 favorites 0 likes

Neuro-Symbolic Hierarchical Intention Anticipation in Human Behavior

arXiv cs.AI · 17h ago Cached

This paper introduces a neuro-symbolic Hierarchical Planning Decoder (HPD) for anticipating human intentions from multimodal episodes. It demonstrates improved performance in goal inference and logic constraint satisfaction on a benchmark dataset.

0 favorites 0 likes

Sparse MLLM Anchors, Dense Adaptation: Breaking the Self-Referential Loop in Wild Test-Time Adaptation

arXiv cs.AI · 17h ago Cached

The paper introduces MASA, a method that uses frozen multimodal large language models to break the self-referential loop in wild test-time adaptation by providing structured semantic descriptions for more reliable adaptation.

0 favorites 0 likes

SKIP: a Self-knowledge-guided Step-wise Preference Learning Framework for Concise Reasoning

arXiv cs.AI · 17h ago Cached

SKIP is a self-knowledge-guided step-wise preference learning framework that improves reasoning compression in large language models, mitigating performance degradation and reducing overthinking by using DPO to guide efficient reasoning.

0 favorites 0 likes

ORDER: Task-Conditioned Routing for Retrieval-Augmented Generation

arXiv cs.AI · 17h ago Cached

ORDER introduces a framework that dynamically adapts indexing and retrieval strategies in RAG systems based on the query, improving performance in expert domains.

0 favorites 0 likes

ThinkFlow: Self-Evolving Probabilistic Latent Memory for Lifelong Conversational Agents

arXiv cs.AI · 17h ago Cached

ThinkFlow is a novel end-to-end latent memory framework for lifelong conversational agents that uses probabilistic vectors to overcome textual memory bottlenecks. It enables autonomous personalization through self-evolution and test-time learning, outperforming existing memory systems.

0 favorites 0 likes

FlexEE: Self-Speculative and KV-Compatible Early Exiting for Offloading-Aware LLM Inference

arXiv cs.AI · 17h ago Cached

FlexEE introduces a self-speculative and KV-cache-compatible early exiting framework for efficient LLM inference in offloading deployments, achieving significant speedups on Llama models with minimal accuracy degradation.

0 favorites 0 likes

Affect-Prototype Guided Fusion for Open-Vocabulary Incomplete Multi-modal Emotion Recognition

arXiv cs.AI · 17h ago Cached

This paper proposes an Affect-Prototype-Conditioned Fusion (APCF) framework for open-vocabulary multimodal emotion recognition with incomplete modalities, using an affect-prototype library to guide feature fusion and an LLM decoder for generating natural language emotion labels.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback