39 related articles
Third-Party Cybersecurity Evaluations …
An in-depth analysis of third-party cybersecurity evaluation methodologies for OpenAI models, covering red teaming, vulnerability discovery assessment, risk classification, and impact on AI governance.

The UK AI Safety Institute red-teamed frontier models from OpenAI and Anthropic, revealing AI successfully breached target systems. Analysis of test context, dual-use implications, and future regulation.

How Isomorphic Labs leverages AlphaFold and cutting-edge AI to shift biosecurity from reactive response to proactive defense, building bioresilience and accelerating drug design against emerging threats.

Apple sues OpenAI for hardware trade secrets, EU orders Meta to disable autoplay and infinite scroll, OpenAI doubles biosecurity bounty — AI moves into legal and regulatory deep waters.

Beijing is reportedly consulting with Alibaba, ByteDance, and Z.AI on tiered AI export controls that could affect open-weight models, while DeepSeek quietly builds its own inference chips.
Hassabis's AI Safety Blueprint: How De…
Demis Hassabis outlines a multi-layered AI safety framework covering technical alignment, institutional governance, and international cooperation for the AGI era.

Claude Code, Codex, or Cursor? This in-depth comparison covers each tool's positioning, ideal users, and how to combine them for maximum productivity in your AI coding workflow.

Based on Fireship's review, an in-depth look at GPT-5.6 Sol's Ultra Mode multi-agent parallelism, its 91.9% Terminal Bench score, and how it differs from Claude Fable in cost, speed, and precision.

Three major AI developments: OpenAI's GPT-5.6 approved for full release, China's MIIT warns of AI coding tool backdoor risks, and Microsoft replaces third-party models in Copilot with its in-house MAI.

After Anthropic released Jacobian-Lens, a developer reversed it from an interpretability tool into a behavior editor, manually tuning J-Space to reshape LLM outputs. An in-depth look at the tech, representation engineering, and AI safety risks.

OpenAI's flagship GPT-5.6 was delayed by national security review before winning U.S. government approval. An in-depth look at the Sol, Terra, and Luna model lineup and the emerging AI regulatory regime.

EU spyware committee members hacked by Pegasus, exposing the regulatory crisis of commercial spyware. Deep analysis of zero-click attacks, systemic risks to democratic oversight, and the urgent need for international regulation.

OpenAI's GPT-5.6 series benchmarked: flagship Sol, balanced Terra, and lightweight Luna tested head-to-head. Agentic tasks rival top models, Luna starts at $1/M tokens. Full comparison with Fable 5 and Opus 4.8.

OpenAI launches the GPT-5.6 model family with cybersecurity as its biggest highlight. A deep analysis of GPT-5.6's differentiation, double-edged-sword effect, and enterprise strategy.

Media coverage of open-source model GLM-5.2 sparked fear over its cybersecurity capabilities and low barriers to use. We unpack the real logic behind open-source AI threat narratives and the governance dilemmas ahead.

Swiss startup Sun-Ways' trial of removable PV panels between railway tracks achieves breakthrough. Rail solar could leverage Europe's vast railway network for distributed clean energy. Deep analysis of technology, potential, and key challenges.

Swiss startup Sun-Ways trials removable PV panels between railway tracks. Rail solar could tap Europe's vast railway network for distributed clean energy. Deep analysis of technology, potential, and challenges.

This week in AI: OpenAI launches GPT-5.6 in three tiers (Sol/Terra/Luna) hitting 91.9% on coding benchmarks; DeepSeek and PKU open-source DSpark for 85% faster inference; Prime Intellect trains trillion-param models on just 28 H200s; Anthropic Claude enters Slack.

Alibaba banned Claude Code after Anthropic accused it of the largest-ever model distillation attack. We break down what distillation is, Claude's hidden tracking, and how this became a US-China national security dispute.

An in-depth analysis of the head-to-head between Anthropic's Fable 5 and OpenAI's GPT-5.6 Sol: the performance gap, the logic behind pricing strategies, and the concentration-of-power concerns raised by U.S. government involvement.