OpenAI has built an LLM super-hacker called GPT-Red that it uses as a sparring partner to help its other models boost their defenses against cyberattacks. Last week the company released the latest version of its flagship LLM, GPT-5.6. OpenAI says that training it against GPT-Red…
ELIZA is remembered as the world’s first AI star, a kindly therapist in chatbot form that gently probed users’ worries. Even its creator, Joseph Weizenbaum, was surprised by the warm reception given to his experiment in human-machine interaction. For some, it heralded an age of…
Shaped like dogs, stars, and the Mona Lisa, you could mistake these DNA structures for fun-shaped macaroni if they weren’t only nanometers wide. South Korean scientists made the constructions using a technique called DNA origami, which can bend genetic material into any form. Designing DNA…
This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. PsiQuantum has a plan to make a massive quantum computer out of light The machine that could change the world will…
There Is No Best AI Model—Only the Best Model for the Right Job
Every few months, a new benchmark appears claiming that one model has become the world’s “best.”
Today it’s GPT.
Tomorrow it’s Claude.
Next month it may be Gemini, DeepSeek, or Qwen.
The reality is much more interesting.
There is no universal winner.
The latest independent benchmark reports consistently show that each frontier model dominates different dimensions of intelligence, cost, reasoning, coding, multimodal understanding, or context handling.
arXiv:2606.06715v2 Announce Type: replace-cross Abstract: We ask whether topic sentiment has a causal effect on perceived political ideology, and whether the answer depends on who assigns the ideology label. Using articles from AllSides, paired with shared sentiment annotations from Llama-3.3-70b-versatile, we compare ideology labels from…
arXiv:2602.11198v2 Announce Type: replace-cross Abstract: Multi-agent frameworks (MAFs) promise to simplify LLM-driven software development, yet no principled metric captures how well AI coding assistants can generate correct, framework-specific code. We introduce textit{AI-assistability} ($mathcal{AI}$), a composite metric that quantifies a framework's amenability to AI-assisted development by…
arXiv:2606.30347v2 Announce Type: replace-cross Abstract: We present FFAvatar, a Transformer-based 3D Gaussian framework for fast construction of high-quality and animatable 4D head avatars from one or more reference portrait images. Unlike existing feed-forward approaches that require a fixed number of input views, FFAvatar supports incremental…