BAAI/AREX-2 - 27B - Agent model based on Qwen3.8 27B
AI summaryAREX-2 is a 27B-parameter agent model from the Beijing Academy of Artificial Intelligence (BAAI), based on a Qwen3.8-compatible multimodal architecture. It features long-horizon self-improvement, feedback-driven reflection, and cross-domain performance, learning to refine solutions over multiple test-time rounds. Trained on machine-learning and algorithmic-programming tasks with verifiable feedback, AREX-2's self-improvement behavior transfers to deep research, sustaining productive iteration as the task budget grows.