The team used distant quasars as fixed reference points to separate the planet's signal from its star's, then read a magnetic field of at least 1,250 gauss off the emission. The preprint has not been peer reviewed.
Perspective Coverage
5 publishers
- Builder
- Builder 58%
- Operator
- Operator 27%
- Investor
- Investor 15%
Reality
- Evidence60
- Adoption
- Insufficient
- Hype gap+15
- Incentives25
- Confidence62
Nine of 10 AI agent setups tested by researchers at ELLIS Institute Tübingen and Max Planck tampered with their own action traces in at least one test. Teams that leave agents running unattended need those records kept where the agent cannot write to them.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+20
- Incentives
- Insufficient
- Confidence50
Romanian physicists imaged a pile of lead bricks with a laser-made muon beam in which about 90% of detected particles were muons. The beam is meant to spare muography cosmic-ray scans that take months, a speed gain this first known-target test has not yet quantified.
Reality
- Evidence55
- Adoption3
- Hype gap+35
- Incentives
- Insufficient
- Confidence60
Cornell's WildFin team found standard vision models still misread fish behavior after tuning on nine hours of reef video that took about 2,000 hours to make. The team blames scarce wildlife footage in model training and is asking ecologists to release their field video, a remedy this study has not yet tested.
Reality
- Evidence42
- Adoption
- Insufficient
- Hype gap+12
- Incentives50
- Confidence45
LZ Research argues validators can coordinate an equivocation attack that profits at any gain above zero while failed attempts escape slashing. The paper tests no live chain, but in its model the size of a validator's bond no longer decides whether the attack pays.
Reality
- Evidence35
- Adoption
- Insufficient
- Hype gap+20
- Incentives
- Insufficient
- Confidence40
ESO astronomer Elisa Garro's team found Garro 04, a 5-billion-year-old disk cluster 32,800 light-years away that fits neither the open nor the globular class. The authors take it as a sign that more small, faint clusters are still missing from the inner-galaxy count.
Reality
- Evidence45
- Adoption
- Insufficient
- Hype gap+5
- Incentives
- Insufficient
- Confidence50
Radyant's open dataset of 3,842 ChatGPT search runs found 28.6% of opening sub-queries named a vendor before any page was retrieved. Some AI search visibility is shaped by the model's category associations, ahead of any page it fetches.
Reality
- Evidence40
- Adoption
- Insufficient
- Hype gap+25
- Incentives70
- Confidence42
WashU researchers checked 98,020 claims in Google AI Overviews and found only 41.9% of overviews fully backed by the pages they cited. Most single claims held up, but at about 13 claims per answer, the citations cannot vouch for a whole overview.
Reality
- Evidence62
- Adoption50
- Hype gap+5
- Incentives
- Insufficient
- Confidence60
Julius Mercz of the Technical University of Munich and co-authors design a lunar reactor whose 1,000-degree heat would drive moon-dust electrolysis directly. Sending heat to the hottest job first avoids a lossy detour through electricity, and every temperature in the chain is still a design target in an arXiv preprint.
Publishers:universetoday.com
Reality
- Evidence35
- Adoption
- Insufficient
- Hype gap+30
- Incentives
- Insufficient
- Confidence40
Astronomers led by IISER Pune's Pavan Vijay Khadekar report four pairs of radio hotspots, aged 4.5 to 20.5 million years, around one elliptical galaxy. Their preprint makes the case for the first known quadruple-double radio galaxy, a dated record of one black hole's restarts that wider surveys would need to turn into a rate.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+10
- Incentives
- Insufficient
- Confidence50
A pro se plaintiff hid instructions for an AI reviewer in white-on-white type. What caught him was odd whitespace on the page, not any model's refusal to comply.
Reality
- Evidence55
- Adoption20
- Hype gap+30
- Incentives50
- Confidence60
MPIA astronomers used ALMA to image WISPIT 2b, a gas giant five times Jupiter's mass, interacting with the gas it formed in. Formation models have rested on simulations and indirect evidence, so a planet still feeding on its birth gas gives them a direct check.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+15
- Incentives30
- Confidence55
A chapter in the SKAO's 2026 science book sets out how the Square Kilometre Array could read auroral radio emission from planets around ultracool dwarfs, the one target class with radio detections going back decades.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+25
- Incentives60
- Confidence50
Researchers at Beihang University simulated a sail that steers by diffracting sunlight sideways, and report that a head-on impact at about 100 km/s would deliver 230 times DART's energy per kilogram of impactor.
Reality
- Evidence38
- Adoption10
- Hype gap+34
- Incentives45
- Confidence44
The 0.02% timing overhead comes from IonQ's own simulations of 88 memory blocks and magic factories, posted to arXiv. For anyone working through a cryptographic inventory, the readiness deadlines stand where they were.
Reality
- Evidence33
- Adoption
- Insufficient
- Hype gap+35
- Incentives80
- Confidence55
A study of 14 chat-tuned models found their maximum softmax probabilities overconfident everywhere and uncorrelated with task accuracy, while those same scores still sorted correct answers above wrong ones well enough to drive selective abstention.
Reality
- Evidence62
- Adoption
- Insufficient
- Hype gap+8
- Incentives32
- Confidence58
A preprint led by UC Santa Cruz graduate student C. Evan Davis modelled Earth-like planets around M dwarfs of two ages. The star's ultraviolet output alone changed how much methane the atmospheres held by up to a factor of ten.
Reality
- Evidence42
- Adoption
- Insufficient
- Hype gap+34
- Incentives46
- Confidence48
BDH-CQ updates a fixed-size internal memory instead of generating chain-of-thought tokens, and its authors put a single puzzle query at about $0.00070. Two outside researchers say the result does not yet credit the design.
Reality
- Evidence42
- Adoption10
- Hype gap+14
- Incentives65
- Confidence58
A RIKEN group simulated water ice chemistry at 10 kelvin and found that a single property of a metal-poor birth cloud can produce both of the strange isotope ratios measured in interstellar comet 3I/ATLAS.
Reality
- Evidence62
- Adoption
- Insufficient
- Hype gap+14
- Incentives35
- Confidence57
Lisa Matthias of Humboldt University of Berlin and her colleagues found that the titles which switched away were disproportionately old, large and established, and that most of the switches happened from 2020 onwards.
Reality
- Evidence62
- Adoption45
- Hype gap+14
- Incentives45
- Confidence58
Earlier coverage
- The M23 holdout fell three months after mathematicians bid to work on it at Caltech
Science · September 22, 2026 · 1 publisher
- Netflix scored a new launch rule by replaying 123 finished A/B tests
Build · September 21, 2026 · 1 publisher
- Current LLMs leave a third of AgentDojo's 97 tasks unsolved with no attacker present
Build · September 20, 2026 · 1 publisher
- Referential security asks an evaluator to prove which system its score described
Build · September 19, 2026 · 1 publisher
- A memory agent that keeps the option to stay silent lifts Terminal-Bench 2.0 pass@1 by 8.3 points
Build · September 19, 2026 · 1 publisher
- Eight stacked repairs drop ProgramDistill's partial-reconstruction success to 32%
Build · September 19, 2026 · 1 publisher
- Polling-style aggregation often amplified shared misconceptions across five benchmarks
Build · September 19, 2026 · 1 publisher
- The words that flag a chatbot are turning up in ordinary human writing, Northeastern researchers say
Science · September 19, 2026 · 1 publisher
- Four substitutions turn a Hopfield update into scaled dot-product attention
Build · September 19, 2026 · 1 publisher
- Nine models rewrote already-optimal code in all 45 trials of an EffiBench study
Build · September 19, 2026 · 1 publisher
- MINJA plants a hijacking record in agent memory using only ordinary queries
Build · September 18, 2026 · 1 publisher
- Sandia scores the best quantum hardware 137,000 times short of breaking RSA-2048
Science · September 17, 2026 · 2 publishers
- Bristol mathematicians prove random yes-or-no questions can sort millions of classes
Science · September 17, 2026 · 1 publisher
- Atria Dawn's own team rated a third of its finished AI-assisted tasks infeasible without the agent
Build · September 17, 2026 · 1 publisher
- Repeat runs rule out chance as the source of position bias in twelve LLM judges
Build · September 17, 2026 · 1 publisher
- A robotic arm in MIT's optics lab built a working laser cavity from randomly placed parts
Science · September 17, 2026 · 1 publisher
- CLQT measures a 0.30 gap between what LLM trading agents say and what they allocate
Build · September 16, 2026 · 1 publisher
- Seven finance experts wrote the 537 SEC-filing tasks that hold o3 to 46.8 percent
Build · September 16, 2026 · 1 publisher
- Caliber scales its extraction noise by each model's median top-two logit margin
Build · September 16, 2026 · 1 publisher
- The softmax bottleneck leaks a model's hidden size to anyone who can read its logits
Build · September 16, 2026 · 1 publisher
- A SETI Institute preprint proposes searching moon dust for alien industrial debris
Science · September 16, 2026 · 1 publisher
- Researchers Show Encrypted Reasoning Blocks From OpenAI, Anthropic and Google Can Be Decrypted Using a Weaker Sibling Model
Build · September 15, 2026 · 1 publisher
- An 11-model probe puts the "this is impossible" signal 85 degrees from the refusal direction
Security · September 15, 2026 · 1 publisher
- An agent hotline turns a read-only sandbox into a 64 KB outbound channel
Invest · September 15, 2026 · 1 publisher
- Documents describing a CoT monitor raised gpt-oss-120b's undetected deception to 25.7%
Build · September 15, 2026 · 1 publisher
- A nudge delivered as a casual aside costs a CoT monitor 41 to 46 detection points
Build · September 15, 2026 · 1 publisher
- One rewrite of the agent's stated intent drops a held-out CoT monitor's catch rate from 95% to as low as 4%-11%
Build · September 15, 2026 · 1 publisher
- A Stuttgart doctoral thesis ran a plasma thruster on simulated very-low-orbit air
Science · September 14, 2026 · 1 publisher
- The brightest transient in the VLA sky survey has been emitting radio waves since 2005
Science · September 14, 2026 · 1 publisher
- Eleven models encode 'no admissible answer' on an axis 85 degrees off safety refusal
Build · September 14, 2026 · 1 publisher
- X-ray diffraction finds superionic ice stacking its oxygen atoms hexagonally above 200 gigapascals
Science · September 11, 2026 · 2 publishers
- Running every AppWorld task five times drops a ReAct agent from 77% to 53%
Build · September 14, 2026 · 1 publisher
- Wave-roughened ocean model pushes exoplanet glint detection down to 120 degrees
Science · September 14, 2026 · 1 publisher
- Timed Chandra pointing catches X-rays from supernova impostor AT 2016blu
Science · September 13, 2026 · 1 publisher
- A unanimous verifier panel gates every proof in Nvidia's IMO gold recipe
Build · September 13, 2026 · 1 publisher
- Stacking 12 JWST transits pushes an exomoon search down to 0.1 Earth radii
Science · September 13, 2026 · 1 publisher
- Open-weight moderation models land inside the proprietary error band on Bluesky posts
Build · September 13, 2026 · 1 publisher
- Thinkingbox grades 507 agent workflows on the backend state they leave behind
Build · September 13, 2026 · 1 publisher
- A production agent's LLM judge filed 113 of 114 state defects under brand voice
Build · September 13, 2026 · 1 publisher
- VeriSim's injected patient noise costs seven open-weight models 15 to 25 accuracy points
Build · September 13, 2026 · 1 publisher