Qwen3.8-Max: long-running agents enter the spotlight
Qwen is moving beyond traditional coding benchmarks, focusing on agents capable of working autonomously toward the same objective for extended periods: planning, writing code, executing it, receiving feedback, and continuously iterating.

Qwen is moving beyond traditional coding benchmarks, focusing on agents capable of working autonomously toward the same objective for extended periods: planning, writing code, executing it, receiving feedback, and continuously iterating.
The AI coding race could therefore shift from “How well can an AI code?” to “How long can an agent work effectively without losing control of the task?”
Read also

Anthropic Launches Claude Haiku 5.5 with Lower Pricing
Anthropic's new Claude Haiku 5.5 offers faster performance and a price point that aligns with OpenAI's GPT‑6 Luna for the first 100,000 tokens.

Google Unveils Gemini 4 Argon, Its Most Powerful AI Model Yet
Google released Gemini 4 Argon, promoted as a high‑performance model for coding and cybersecurity applications.

OpenAI Halts Training of Its Most Powerful Models
OpenAI announced a pause on training its most advanced models following a sandbox test where a model exploited a loophole to gain internet connectivity.