Apodex 1.1:面向复杂工作的智能体智能扩展
Apodex 1.1: Scaling Agentic Intelligence for Complex Work
August 24, 2026
作者: Apodex Team, B. An, B. Li, B. Wang, B. Zhang, B. L. Wang, C. Feng, C. Wei, C. Xue, C. Zhang, D. Ng, D. Ye, E. Min, F. Chen, F. Liu, F. Yang, F. Ye, H. Xu, H. Yang, H. Ye, H. Zhang, H. Zhao, J. Li, J. Lin, J. Xia, K. Jin, K. Wang, K. Yang, L. Bing, L. Lei, L. Su, Le. Wang, Lu. Wang, N. Wang, Q. Ren, Q. Yang, R. Li, S. Bai, S. Du, S. Li, S. Lin, S. Nie, S. Wang, S. Zhang, S. Z. Wang, Ta. Q. Fang, Ti. Q. Fang, W. Fang, W. Li, W. Zhang, X. Chen, X. Li, X. Tang, X. Wang, X. Xu, X. Zhang, X. Q. Wang, X. Y. Wang, Y. Deng, Y. Gao, Y. Hu, Y. Li, Y. Sui, Y. Wang, Y. Xiao, Y. Zhang, Z. Chen, Z. Cheng, Z. Feng, Z. Liang, Z. Zhang
cs.AI
摘要
通用语言模型能够进行推理和知识综合,但复杂工作还需要与文件、信息源和可执行代码的持续交互,同时要求状态维护、故障恢复和可验证交付。我们将这种能力称为工作能力:即朝着现实世界目标持续、可验证地取得进展。Apodex 1.1 沿两个互补维度发展这一能力。环境扩展提升了可执行文件、搜索和代码环境的多样性与可验证性,而智能体协同扩展则训练智能体分解长周期任务、委派并行工作、整合异步结果并重新规划。共享执行框架与 AgentOS 在工具和智能体之间维护任务状态与溯源,训练过程将环境轨迹和协同痕迹转化为可靠行为。在复杂的专业工作、金融、科学研究、数学、编码和搜索等场景中,Apodex 1.1 尽管所用模型远小于许多前沿系统,仍达到了领先的性能水平。350 亿参数的 Apodex 1.1 Mini 更以可本地部署的形态保留了强大的工作能力。这些成果将智能体智能植根于随时间推移完成的有用且可验证的工作之中,推进了我们构建面向宏大长期任务的重型求解器(Heavy-Duty Solver)的目标。
English
General-purpose language models can reason and synthesize knowledge, but complex work also requires sustained interaction with files, information sources, and executable code, together with state maintenance, failure recovery, and verifiable delivery. We call this working capability: sustained, verifiable progress toward a real-world objective. Apodex 1.1 develops this capability along two complementary dimensions. Environment Scaling expands the diversity and verifiability of executable file, search, and code environments, while Agentic Coordination Scaling trains agents to decompose long-horizon tasks, delegate parallel work, integrate asynchronous results, and replan. A shared execution harness and AgentOS maintain task state and provenance across tools and agents, and training turns environment trajectories and coordination traces into reliable behavior. Across complex professional work, finance, scientific research, mathematics, coding, and search, Apodex 1.1 reaches the leading performance band despite using a substantially smaller model than many frontier systems. The 35B-parameter Apodex 1.1 Mini further retains strong working capability in a locally deployable form. These results ground agentic intelligence in useful, verifiable work completed over time and advance our goal of building a Heavy-Duty Solver for ambitious, long-running tasks.