14 related articles

OpenAI launches GPT-5.6 Cyber hacker model with 95% response rate; Claude advances Riemann Hypothesis record from 41.6% to 67.2%; Meta open-sources 30B local agent model; Tencent generates 3D worlds from text.

A Reddit user used a GPT model to improve Anthropic's numerical bound on the Riemann Hypothesis zero ratio from 67.25% to 67.28%. Analyzing AI's discovery of Gram matrix spectral information loss and LLM capabilities vs. hallucination risks in frontier math.

A non-mathematician used ChatGPT to find a normalization error in two published Riemann Hypothesis papers, confirmed by the author. An analysis of AI-assisted academic auditing.

HyperSAE uses Poincaré ball hyperbolic geometry to replace Euclidean space in sparse autoencoders, reducing dead latents from 3.8% to 0.2% with zero inference cost.

Exploring how AI is successively solving Erdős math problems, analyzing the key factors of LLM reasoning breakthroughs and formal verification, plus the profound impact and debates AI brings to mathematical research.

OpenAI's internal model codenamed Astra reportedly solved 10 major open math problems. We examine the claim's credibility, AI math reasoning capabilities, and a rational evaluation framework.

The Theo Conjecture, unsolved for 35 years, has been cracked with an unexpected new term discovered. Exploring AI's evolving role in pure math research.

Fields Medalist Terence Tao on AI and mathematics: within a decade, AI will handle much of mathematicians' routine work — but that work was never the core of the discipline.

GPT-5.6 Soul Ultra proves the 50-year-old Cycle Double Cover Conjecture in under an hour. Plus: BCI clinical breakthrough, Apple vs. OpenAI, xAI privacy concerns, and EU dark pattern rules.
Can AI Prove Mathematical Conjectures?…
A PDF claiming GPT-5.6 Sol Ultra proved the Cycle Double Cover Conjecture sparked debate on Hacker News. We unpack the truth and the limits of LLMs in math proofs.

Block-sparse featurizers remap dense vision model activations into block-sparse representations, making the internal feature spaces of ViT, CNN, and other models readable and interpretable. This article explores their core principles, links to mechanistic interpretability, and applications.

A four-layer breakdown of why Chain-of-Thought (CoT) boosts LLM reasoning: compute allocation, external working memory, pretraining pattern activation, and DeepSeek R1 RL evidence.
OpenAI Model Disproves 80-Year-Old Erd…
OpenAI's AI model found a counterexample disproving an 80-year-old Erdős conjecture. Learn about the human-AI collaboration process, mathematical significance, and AI's breakthrough in pure math.
Tech FrontiersOpenAI CEO Sam Altman announces a general-purpose AI model has solved a major open math problem. We analyze this milestone, the leap from specialized to general AI, and its implications for science.