OBSERVATORY DEGRADED

Semantic review is behind discovery. Evidence and Pulse may be incomplete until the backlog is cleared.

2 material first-party/frontier/open-problem candidate(s) have waited more than 6 hours for semantic review.

0 critical overdue · 2 material overdue
Last semantic import 19 Sept, 04:12

EVIDENCE REGISTER

The evidence archive

A chronological register of published signals. Source class, verification state, evidence quality and relevance remain visible before interpretation.

PUBLISHED RECORDS

96 matching evidence records

76–96 / 96
19 Aug 2026AI for Science

Erdős Problem #501 resolved as independent of ZFC with Lean-verified human–AI work

A Lean 4 development resolves the first question of Erdős Problem #501 as independent of ZFC: CH gives a negative model, while adding sufficiently many random reals gives a positive model. The missing large-cardinal-free transfer is credited to Elliot Glazer working with Sol and Claude; prior components were human results.

PRIMARY CONFIRMEDFORMAL / REPOSITORYPrimary source ↗
Relevance7.3/10evidence 92/100
18 Aug 2026AI for Science

Stein dimension-free weak-(1,1) Riesz transform problem resolved

Ouyang, Spector and Stockdale prove a dimension-free weak-type (1,1) bound with constant 2 for the vector Riesz transform, settling Stein’s 1986 ICM problem. The paper states that the proof strategy was developed through LLM dialogues involving GPT-5.6 Sol, OpenAI reasoning agents and Claude Opus 5.0; the authors then independently checked, rewrote and validated the argument.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance8.4/10evidence 82/100
18 Aug 2026AI for Science

GPT-5.6 Sol-assisted work disproves Sato's weak F-equivalence conjecture

Avik Chakravarty, Daebeom Choi and Shengjing Xu construct counterexamples disproving Sato's weak F-equivalence conjecture for nonsingular projective toric weak Fano varieties in every dimension at least three, then formulate a Gorenstein refinement. The paper states that the results were developed with assistance from GPT-5.6 Sol.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance8.7/10evidence 88/100
18 Aug 2026AI for Science

Human–AI collaboration produces new representation-theory results with full Lean formalization

Haruhisa Enomoto reports a research collaboration with Fable 5 and GPT-5.6 Sol that proposes a quotient–submodule equidistribution conjecture and proves multiple nontrivial families, with a full Lean 4 formalization. This is evidence of frontier models participating in novel mathematical research, but it does not close a pre-existing named open problem.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance7.8/10evidence 91/100
18 Aug 2026AI for Science

Claude autonomously designs de novo protein binders validated in wet lab

Anthropic reports that Claude autonomously orchestrated open-source protein-design and folding tools from a human-written research prompt and produced binders against 14 of 15 targets. Adaptyv Bio and Twist Bioscience independently synthesized and tested the designs; reported hit rates were 22–35%, above the 10–15% field baseline cited by Anthropic.

INDEPENDENTLY CONFIRMEDFIRST-PARTY SOCIALPrimary source ↗
Relevance8.8/10evidence 88/100
17 Aug 2026AI for Science

FAR pipeline scales AI-assisted mathematical discovery across thousands of open problems

The Find-Attempt-Recommend pipeline scans thousands of combinatorics papers, identifies thousands of apparently open conjectures, surfaces hundreds of potential resolutions, and narrows them to a small expert-review set. This is a material AI-for-Science signal because it automates problem selection and triage, two human bottlenecks in research-level mathematics.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance8.6/10evidence 82/100
16 Aug 2026AI for Science

Talagrand convolution conjecture solved with AI-discovered proof

A new preprint claims a full proof of Talagrand's 1989 convolution conjecture on the Boolean hypercube. The authors state that the proof was discovered by the Odin Automatic AI Research Agent; they reorganized the final proofs. Independent mathematical review is still pending.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance8.1/10evidence 82/100
13 Aug 2026AI for ScienceBreakthrough

ChatGPT 5.6 Sol-assisted construction resolves elliptic regularity question negatively

Nam Q. Le, Qi Sun and Hung V. Tran construct uniformly elliptic nondivergence-form equations in dimension three whose bounded solutions have unbounded W1,1 variation, showing that no interior W1,p estimate depending only on ellipticity exists for p at least one and resolving negatively an open question raised by Nadirashvili, Tkachev and Vlăduţ. The authors state that the main results were obtained through chats with ChatGPT 5.6 Sol and that key strategies came from ChatGPT; they then reworked, rewrote and checked all arguments. The result remains without independent external proof reproduction.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance9.3/10evidence 89/100
13 Aug 2026AI for Science

Inherent Faraday outperforms frontier agents on a paper-replication benchmark

Inherent reports that Faraday, a 27B AI Scientist agent trained with long-horizon RL and using coding agents as tools, outperforms Claude Opus 4.8 and GPT-5.5 on held-out scientific paper-replication tasks in the new 310-task Replica benchmark. The result is first-party and benchmark/judge design is controlled by the same team, so independent replication remains pending.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance7.6/10evidence 84/100
11 Aug 2026AI for Science

Long-horizon AI research system helps set new bounds on the Grothendieck constant

A human-AI research collaboration tightened both lower and upper bounds on the Grothendieck constant. The companion methodology paper states that the AI system discovered and first proved the new lower bound K_G >= 6π/11, later independently checked and revised by the human authors; the exact value of K_G remains open.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance8.4/10evidence 92/100
11 Aug 2026AI for ScienceBreakthrough

Claude-assisted Riemann-related result improves a critical-line bound to 67.2%

A Claude-assisted research workflow reportedly improved the proportion of non-trivial simple zeros shown to lie on the critical line from about 41.6% to 67.2%. This was a substantial partial result, not a proof of the Riemann Hypothesis. The work used roughly 31M output tokens, about 60 sub-agents, around 650 discarded strategies, thousands of computational checks and Lean formalization.

PRIMARY CONFIRMEDANTHROPIC RESEARCH / VIBEMATHED
Relevance9.0/10
08 Aug 2026AI for Science

Dittert-5: near end-to-end AI mathematics workflow

GPT-5.6 Sol Ultra reportedly carried a Dittert-5 candidate through problem selection, discovery, certificate generation, Lean formalization, verification and manuscript preparation. The workflow was unusually autonomous, but independent human audit and review were still missing.

PRELIMINARYVIBEMATHEDPrimary source ↗
Relevance7.0/10
07 Aug 2026AI for Science

WeatherNext Cyclones signals faster AI-enabled forecasting progress

The 7 August Pulse recorded WeatherNext Cyclones as roughly one day of additional warning and about a decade of meteorological progress, without a documented ASI forecast revision.

PRIMARY CONFIRMEDORIGINAL SINGULARITY PULSE
Relevance8.0/10evidence 78/100
06 Aug 2026AI for Science

AI-assisted construction gives negative solution to the inverse generator problem on Hilbert spaces

Lorist, Meyries and Veraar construct a bounded strongly stable C0-semigroup generator on a Hilbert space whose inverse does not generate a C0-semigroup, resolving the inverse generator problem negatively. The current arXiv v2 AI disclosure states that ChatGPT 5.6 Pro explored Schauder-basis counterexamples, assisted the explicit basis construction in Proposition 2.1, and helped optimize constants.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance7.1/10evidence 84/100
06 Aug 2026AI for Science

K-U-M rank-3 mathematics result recorded as Lean-verified

The 6 August Pulse included a rank-3 K-U-M result recorded as Lean-verified. It was notable but did not justify an ASI forecast change.

PRIMARY CONFIRMEDORIGINAL SINGULARITY PULSE
Relevance7.0/10evidence 80/100
03 Aug 2026AI for ScienceBreakthrough

GPT-5.6 Sol-assisted construction gives counterexample to Connes' rigidity conjecture

Shuoxing Zhou constructs non-isomorphic ICC property-(T) groups with isomorphic group von Neumann algebras, giving a counterexample to Connes' rigidity conjecture for this class. The paper explicitly says the result was obtained with GPT-5.6 Sol assistance, independently of and concurrently with OpenAI work. The theorem and provenance are primary-source confirmed; independent specialist verification remains pending.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance9.4/10evidence 91/100
01 Aug 2026AI for ScienceBreakthrough

Astra reports 10 Lean-formalized mathematics results

Astra/OpenAI was recorded in the original Pulse as producing 10 mathematical results formalized in Lean, triggering the largest early ASI forecast revision in the archive.

PRIMARY CONFIRMEDORIGINAL SINGULARITY PULSE
Relevance9.0/10evidence 85/100
20 Jul 2026AI for Science

GigaPath-Flash cuts whole-slide pathology compute while preserving most predictive performance

Microsoft Research, Providence and University of Washington report two open-weight pathology foundation models for population-scale research. GigaPath-Flash retains 97% of GigaPath's average slide-level performance with roughly 50 times less compute. GigaTIME-Flash reports sixfold higher speed and eightfold lower memory while matching or improving spatial-protein prediction across tested cohorts.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance8.1/10evidence 76/100
20 Jul 2026AI for Science

ChatGPT 5.6 authored a preprint proving full replica symmetry breaking at zero temperature

A preprint attributed by its submitter to ChatGPT 5.6 proves full replica symmetry breaking for the zero-field Sherrington–Kirkpatrick model at zero temperature. A catalog audit found that the accompanying Lean development is conditional on seven identified analytic inputs and is not an assumption-free verification.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance8.7/10evidence 64/100
20 Jul 2026AI for ScienceBreakthrough

Claude-assisted counterexample refutes the Jacobian Conjecture in dimensions n≥3

Levent Alpöge announced an explicit polynomial self-map of C^3 with constant nonzero Jacobian determinant that is not injective, refuting the Jacobian Conjecture for dimensions n≥3. Alpöge explicitly credited Claude Fable 5 with work leading to the counterexample. The result was rapidly independently checked in exact arithmetic and formally verified in Isabelle/HOL and Lean-derived work. The two-dimensional case remains open.

INDEPENDENTLY CONFIRMEDFORMAL / REPOSITORYPrimary source ↗
Relevance9.2/10evidence 98/100
08 May 2026AI for ScienceBreakthrough

ChatGPT 5.4 Pro autonomously proves and disproves two Bruhat-order conjectures

Colin Defant's revised paper on the MacNeille completion of Bruhat order proves a conjecture of Escobar, Klein and Weigandt and gives a counterexample to a conjecture of Hamaker and Reiner. The abstract states that those two results were obtained autonomously by ChatGPT 5.4 Pro and presents the paper as a case study in LLM-automated mathematical research. The underlying paper was first submitted in May and revised September 6; the recovered revision is primary-source confirmed but not independently externally reproduced.

PRIMARY CONFIRMEDPREPRINTPrimary source ↗
Relevance9.4/10evidence 89/100
Page 4 / 4
← PreviousNext →