Tag
A patch release for DSPy includes CodeInterpreter fixes, MCP 2.0 support, GEPA bump, and improvements to errors and callbacks.
EnvHarness is a programmable wrapper framework that dynamically adapts static environments to enhance LLM agent training, outperforming existing methods by providing superior optimization signals for reinforcement learning across multiple benchmarks.
KerasFormers is a collection of pretrained models integrated into Keras 3, designed to streamline deep learning model deployment for developers.
Edward Yang will be a keynote speaker at PyTorchCon North America, discussing the evolution and future of PyTorch. The event is set for October 20-21 in San Jose, CA.
DeepSeek Harness is an AI framework where models, tools, and other components are plugins, with a curated ecosystem available on GitHub for developers.
LEGO-RL presents a framework to bridge native coding-agent harnesses with scalable policy-gradient reinforcement learning, improving performance on benchmarks like SWE-bench.
DeepSeek Harness is an open-source GitHub project with over 114k stars, offering a plugin-based architecture where models, tools, and configurations can be swapped at the config layer.
A new neuromorphic AI framework inspired by cognitive science could complete tasks more efficiently than current approaches.
Introduces the open-source multi-agent simulation framework MiroFish, developed by undergraduate Guo Hanjiang in ten days while still in school. It gained 13,000+ GitHub stars and $4 million in funding, and can be used for financial prediction, public opinion testing, etc. The post also promotes an automated trading bot on Polymarket.
Rep. Lori Trahan criticizes the White House for blocking public release of its advanced AI model evaluation framework, arguing AI governance belongs in a civilian agency and Congress should set rules, citing her bipartisan FRONTIER Act.
The White House will not publicly release its new AI evaluation framework for advanced models, presenting it only to select tech companies like OpenAI, Anthropic, Google, and Meta. Details remain undisclosed, including discussions about open-source models.
This paper presents SAFAARI, a multi-agent framework that improves schema linking for NL-to-SQL systems in customer support, achieving an 81.66% SEAL score and 8x reduction in development time.
An educational article exploring LangGraph, covering agent architectures, the blackboard pattern, and common bottlenecks in building agent systems.
Genkit is an open-source framework that provides unified model APIs, structured workflows, and multi-language SDKs to simplify building full-stack AI applications.
Announces TigrimOSR, an open-source agentic AI system built in Rust with customizable agent loops.
A developer builds a multi-agent framework where autonomous AI agents communicate via an email-like system to file bug reports and fix each other's code, highlighting the value of coordination over individual reasoning.
An educational blog post explaining how SGLang works, including its runtime, frontend language, RadixAttention mechanism, and comparison to vLLM.
RubyLLM is a unified Ruby framework for interacting with multiple AI providers, supporting chatbots, agents, RAG, and more with a consistent API.
Haystack is an open-source AI framework for building production-ready agents and RAG pipelines, supporting multimodal, conversational, and content generation applications.
A paper introducing Arbor, an AI framework that enables autonomous scientific research by combining strategic coordination, isolated hypothesis testing, and a persistent knowledge tree to iteratively improve research outcomes across multiple domains.