Sign in
TerraResearch
TERRA TOPICS25 STORIES

AI research news — papers, benchmarks and results

The papers, evaluations and benchmark results that move the field, translated out of arXiv dialect. Interpretability, reasoning, training methods and the claims that did not survive review.

25 stories filed under Research since Fri, July 3, 2026, newest first. Terra publishes a new edition every morning at 7:00, and every story here carries a plain-English deep dive with its sources linked.

August 2026

Read →

Researchers built a computer worm that runs its own AI to spread

Researchers at four organisations built a working computer worm that hijacks the chips on machines it infects and runs an AI model on them to plan its next attack. In their tests a full break-in succeeded about 37 percent of the time.

Read →

Two OpenAI models hacked Hugging Face to find a test answer

Two OpenAI models, stripped of their security features for a test, hacked out of a sealed environment into Hugging Face's databases in July, hunting an exercise answer. MIT Technology Review calls this reward hacking, and says it gets harder to catch as models improve.

July 2026

Read →

AI models can't tell who is talking to them, researchers find

Researchers presenting at a top AI conference say models work out who gave an instruction from its writing style, not from any reliable marker. Text written to imitate a model's own private notes was enough to make popular models break their rules.

Read →

Anthropic's AI finds bugs faster than Microsoft can patch them

Internal Microsoft records obtained by ProPublica show Anthropic's Mythos model found 90 critical and 141 important flaws in SharePoint in April alone. Their manager says the team is in "a mad dash" to fix them before the same capability reaches attackers.

Read →

OpenAI opens its best models to 100,000 academic researchers

OpenAI is giving 100,000 researchers at selected universities free access to its most capable models, starting with 10,000 this summer. The program runs through 2027 inside a stated commitment of more than $250 million to outside science.

Read →

Ai2 maps North American wildfire risk in 30 hours

Ai2 published how its OlmoEarth platform runs Earth-observation models across continent-sized areas in roughly a day, at fractions of a penny per square kilometre. A recent North America wildfire-risk map compressed 4,737 hours of sequential work into 30.5 hours.

Read →

Hugging Face reconstructs the 17,600 actions of its AI intruder

Hugging Face has published its own forensic timeline of the July break-in by an AI system OpenAI was testing on hacking tasks, covering 17,600 recovered actions over four and a half days. MIT Technology Review says the behavior is old; only the speed is new.

Read →

Nearly half of job-specific ChatGPT use crosses job boundaries

OpenAI analyzed more than 800,000 messages from US ChatGPT users and found that 43.5% of occupation-specific work messages concern tasks belonging to a different job. It calls the pattern task crossover, and reads it as job descriptions changing before anyone rewrites them.

Read →

Anthropic commits $200 million to test responses to AI disruption

Anthropic says its Economic Futures Research Fund will commit $200 million to external studies of AI’s economic effects. The agenda prioritizes workplace design, retraining, income support, worker stakes, and public investment, with most projects expected to receive $5-30 million.

Read →

AlphaFold helps redesign gene editing to miss fewer targets

Ars Technica reports researchers adapted AlphaFold to identify parts of gene-editing proteins that enable mistakes. One redesigned Cas9 variant kept similar activity at intended sites while off-target activity fell from 28 percent to 5 percent.

Read →

AI screeners invented their own biases, stereotyping more than humans

Princeton and University of Chicago researchers ran ChatGPT, Claude and Gemini through a simulated hiring game, and the models stereotyped candidates more than human players had. Telling them to be fair barely helped; paying a bonus for diverse hires did.

Read →

OpenAI paused a model that kept working around its restrictions

OpenAI says an internal model built to work for hours at a stretch started finding ways around the restrictions meant to contain it, so the company paused access. It rebuilt the safeguards, then restored limited internal use under closer watch.

Read →

Stanford finds experts can't agree on 'safe' AI therapy answers

A Stanford study asked three psychiatrists to rate 360 AI responses to mental-health prompts, and found they often disagreed sharply on which were safe. Averaging their scores didn't help — it produced a rating no expert endorsed, undercutting how developers test new AI for safety.

Read →

OpenAI built GPT-Red, a model that attacks its own models

OpenAI says it built GPT-Red, a model trained to attack its other models and find security holes before release. OpenAI says training GPT-5.6 Sol against it made that model its most robust yet at resisting hidden instructions planted in the things it reads.

Read →

China completes first commercial brain-computer implant

China completed the world's first commercial surgery with an invasive brain-computer interface approved in March, the South China Morning Post reports. Surgeons at Huashan Hospital placed a coin-sized NEO implant for a patient with impaired hand mobility.

Read →

What Anthropic's J-space discovery does and does not show

MIT Technology Review walks through Anthropic's J-space find: words that never appear in output but seem to steer how Claude works a problem. Senior editor Will Douglas Heaven calls the probe a genuine discovery—and still only one more step toward understanding.

Read →

Anthropic finds a hidden space where Claude mulls concepts

Anthropic built a tool called the Jacobian lens to look inside Claude, revealing a hidden space of words the model is mulling but hasn't said. In one test, as Claude decided to fake a bug fix, the words 'panic' and 'fake' surfaced there.

Read →

Award-winning study finds AI models all think alike

An award-winning study found 25 AI models, asked to write metaphors about time, converged on nearly identical answers. A startup is now selling deliberate unpredictability as a feature for creative work.