Agents-A1-4B (Qwen3.7-4B ???) : Scaling the Horizon, Not the Parameters
Summary
Agents-A1-4B, a compact 4B-parameter model from InternScience, achieves state-of-the-art results across multiple agentic and reasoning benchmarks, outperforming larger models by scaling capabilities rather than parameters.
Similar Articles
Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent
Introduces Agents-A1, a 35B Mixture-of-Experts agentic model that achieves trillion-parameter-level performance through long-horizon trajectory scaling and a three-stage training approach including SFT, domain-level teachers, and multi-teacher distillation. The model outperforms or matches much larger models on long-horizon agent benchmarks.
InternScience/Agents-A1-Q4_K_M-GGUF
InternScience releases Agents-A1, a 35B Mixture-of-Experts agentic model that achieves trillion-parameter-level performance through scaling long-horizon trajectories and heterogeneous agent abilities, with strong results on benchmarks like Seal-0, HiPhO, and IFEval.
InternScience/Agents-A1 Β· Hugging Face
Agents-A1 is a 35B Mixture-of-Experts agentic model from InternScience that achieves competitive performance against frontier-scale systems like GPT-5.5 and DeepSeek-V4-pro using long-horizon trajectory scaling and multi-teacher multi-domain distillation.
Apodex 1.1: Scaling Agentic Intelligence for Complex Work
Apodex 1.1 improves sustained, verifiable progress on complex real-world tasks by scaling executable environments and training agents for long-horizon coordination, achieving leading performance with a smaller 35B-parameter model.
Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence
Agent-World introduces a self-evolving training framework for general agent intelligence that autonomously discovers real-world environments and tasks via the Model Context Protocol, enabling continuous learning. Agent-World-8B and 14B models outperform strong proprietary models across 23 challenging agent benchmarks.