search-agent

Tag

Cards List
#search-agent

Introducing Toast 1

Hacker News Top ↗ · 2026-08-14 Cached

Mixedbread introduces Toast 1, a specialized search agent that matches frontier model quality while being up to 10x cheaper and 12x faster. It automates agentic search loops and achieves state-of-the-art results on benchmarks like OfficeQA Pro V2 and legal knowledge tasks.

0 favorites 0 likes
#search-agent

@samsja19: It was a pleasure to collaborate with the talented @mixedbreadai team to push a sota agentic search with our prime rl s…

X AI KOLs Timeline ↗ · 2026-08-13 Cached

Samsja19 highlights their collaboration with Mixedbread AI on Toast 1, a specialized agentic search model that achieves frontier quality at 12x faster speed and 1/10th the cost via reinforcement learning.

0 favorites 0 likes
#search-agent

I open sourced my hackathon search agent, but I’m still figuring out the best model for evaluation

Reddit r/AI_Agents ↗ · 2026-08-12

The author open-sourced hackathon-searcher, a tool that discovers, evaluates, and helps apply to hackathons, and is seeking community advice on how to design its evaluation layer using open-source models.

0 favorites 0 likes
#search-agent

TourSynbio-Search: A Large Language Model Driven Agent Framework for Unified Search Method for Protein Engineering

arXiv cs.AI ↗ · 2026-08-06 Cached

The paper presents TourSynbio-Search, an LLM-driven agent framework for unified protein engineering search across literature and biological databases, powered by the TourSynbio-7B multimodal model with dual PaperSearch and ProteinSearch components.

0 favorites 0 likes
#search-agent

@KaiZhang_CS: Check out one of the best open-source search agents trained by @jianxie_ !! glad to see early experience methods work o…

X AI KOLs Timeline ↗ · 2026-06-17 Cached

Yu Su's team trained a frontier Deep Research Agent on an academic budget using 8K synthetic samples and RL, releasing fully open training infrastructure and models from 2B to 35B parameters.

0 favorites 0 likes
#search-agent

Spent the weekend on the Apodex 4b, plus a quick look at the 35b mini

Reddit r/LocalLLaMA ↗ · 2026-06-12

The author tests the Apodex 4B-SFT and 35B mini models, finding the 4B-SFT surpasses other 4B models in multi-hop search tasks without hallucination, and notes the design philosophy of separating answer checking from generation.

0 favorites 0 likes
#search-agent

LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards

Hugging Face Daily Papers ↗ · 2026-05-29 Cached

LongTraceRL introduces tiered distractor construction and rubric reward design to improve long-context reasoning in language models using reinforcement learning. The method generates multi-hop questions via knowledge graph random walks and uses search agent trajectories to build challenging distractors, with a rubric reward providing entity-level process supervision.

0 favorites 0 likes
← Back to home

Submit Feedback