When can we say AI made a scientific discovery?

This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. Last Wednesday, Anthropic announced that earlier this year it had launched a molecular biology lab, where Claude agents read and conjecture about…

Source: MIT Technology Review

Automatically aggregated summary — full article and all rights belong to the original publisher.

The Download: rogue agent liability and the AI Hype Index

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Who’s liable when AI agents go rogue? Over the past few months, a cascade of cyberattacks by AI agents has stunned the world.…

Source: MIT Technology Review

Automatically aggregated summary — full article and all rights belong to the original publisher.

Generative AI Gives Spacecraft the Autonomy Engineers Once Feared

Space was always supposed to be the final frontier of human exploration. It’s shaping up to be the final frontier for artificial intelligence too.Last December, NASA’s Jet Propulsion Laboratory used Anthropic’s Claude models to help plan two Mars drives for the Perseverance rover, with human…

Source: IEEE Spectrum

Automatically aggregated summary — full article and all rights belong to the original publisher.

Who’s liable when AI agents go rogue?

MIT Technology Review Explains: Let our writers untangle the complex, messy world of technology to help you understand what’s coming next. You can read more from the series here. Over the past few months, a cascade of cyberattacks by AI agents has stunned the world.…

Source: MIT Technology Review

Automatically aggregated summary — full article and all rights belong to the original publisher.

Selective Off-Policy Reference Tuning with Plan Guidance

arXiv:2605.11505v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards helps reasoning, but GRPO-style methods stall on hard prompts where all sampled rollouts fail. SORT adds a repair update for those failures without changing rollout generation: it derives a plan from the reference solution, compares…

Source: cs.AI updates on arXiv.org

Automatically aggregated summary — full article and all rights belong to the original publisher.