XDA Developers on MSN
I gave Claude Opus 5, GPT-5.6, and Grok 4.6 the same complex web project and only one behaved like a senior developer
One planned ahead, two just coded.
For months, the leading AI coding benchmarks have told enterprise buyers a comforting but misleading story: the top models are all roughly the same. OpenAI's GPT-5 family, Anthropic's Claude Opus, and ...
Today, Chinese AI startup Z.ai (formerly Zhipu AI) announced the immediate release of GLM-5.2, a 753-billion parameter open-weights large language model (LLM) engineered specifically to dominate "long ...
OpenAI Group PBC today launched a new large language model that is significantly better than its predecessors at solving math problems and writing code. GPT-5.5 is rolling out a week after rival ...
OpenAI has rolled out a new coding-focused AI model, GPT-5.3-Codex, at a time when competition in developer-facing AI tools is heating up fast. The company says this is its most capable agentic coding ...
OpenAI is pitching GPT-5.3-Codex as a long-running “agent,” not just a code helper: The company says the model combines GPT-5.2-Codex coding strength with GPT-5.2 reasoning and professional knowledge, ...
The model, codenamed “Spud,” is designed to complete complex multi-step tasks with minimal human direction. It sets new benchmarks in agentic coding, computer use, and knowledge work, while matching ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results