Tag
A mixture-of-agents paper (arxiv 2406.04692) shows that a committee of cheap open models can outperform GPT-4o on AlpacaEval 2.0 by leveraging decorrelated errors, and the author shares similar real-world findings where multiple cheap models catch more bugs than a single expensive model.
Mercor announces joining the OpenEnv committee alongside Meta, PyTorch, NVIDIA, PrimeIntellect, and Hugging Face to guide the open foundation for agentic environments.