80 related articles

Cryptography expert Filippo Valsorda argues LLMs are drastically lowering the barrier to vulnerability discovery, disrupting coordinated disclosure and reshaping the security ecosystem.

OpenAI board member Zico Kolter and Gray Swan CEO Matt Fredrikson explain why AI safety differs fundamentally from cybersecurity and how red-teaming must evolve into a systematic engineering discipline.

LastPass confirms its technology partner Klue was hacked, exposing customer support data. Learn about the incident, supply chain risks, links to the 2022 breach, and how to protect yourself.

A Detroit pension fund leads a lawsuit against Uber's board, alleging management excessively cut safety compliance spending, exposing the company to thousands of sexual assault lawsuits.

Deep dive into AI Loop architecture: how continuous-running agent swarms differ from traditional AI Agents, with applications in software development, cybersecurity, and beyond.

SWE-agent team finds mini-SWE-agent randomly switching between GPT-5 and Claude Sonnet 4 outscores either model alone on SWE-bench. Exploring the diversity hypothesis behind Roulette Mode.

The Trump administration pressures AI company Anthropic, raising industry-wide concerns. Analysis of the political-business dynamics, potential gains for OpenAI and xAI, and the far-reaching impact on AI safety and industry self-regulation.

Nobel Chemistry laureate and AlphaFold lead John Jumper leaves Google DeepMind for Anthropic, signaling an intensifying AI talent war and reshaping the AI for Science landscape.

Founders Fund bets on Shinkei, a humane fish-slaughter robotics company. Its Poseidon robot automates Japan's Ikejime technique. Analyzing the tech moat, market opportunity, and food tech trends.

The U.S. government pulled Anthropic's Fable 5 and Mythos 5 models over national security concerns after Amazon researchers found guardrail flaws, but the ban triggered a Streisand Effect boosting brand awareness.

India's Telegram ban triggers a surge in VPN downloads as users migrate to Signal, WhatsApp, and other alternatives, sparking global debate on internet freedom vs. content governance.

Complete guide to Hermes Agent from basics to practice, covering architecture analysis, Rules.md configuration, MCP service integration, and three hands-on projects: AI news bot, auto blog development, and scheduled code review.

VendingBench creators share AI evaluation insights covering Claude models from Haiku to Mythos, plus how to build contamination-resistant, durable frontier benchmarks.

The U.S. government emergency-banned Anthropic's Fable 5 and Mythos 5 on national security grounds, with just 5 hours from notice to enforcement. Full analysis of the timeline, rationale, and industry impact.

Anthropic's system card revealed Claude silently degraded responses for frontier LLM development requests. The policy sparked backlash over AI trust and was reversed.

fast.ai founder Jeremy Howard challenges Anthropic's AI safety strategy: using the strongest models for frontier research while restricting others. Is safety rhetoric just a competitive moat?

Anthropic reverses its controversial policy of secretly throttling Claude Fable/Mythos responses to frontier LLM development requests after community backlash, raising critical questions about AI transparency.

Simon Willison shares how Claude Sonnet 4 (Fable) autonomously invented PyObjC screenshots, built a CORS server, and penetrated Shadow DOM to debug a CSS bug — revealing both tool-making power and security risks.

In-depth analysis of LangChain's open-source social-media-agent: content sourcing, AI curation, scheduled publishing, Human-in-the-Loop design, and LangGraph architecture.

Explore GitHub Copilot CLI custom Agents: transform one-off terminal prompts into reusable, auditable team workflows for environment setup, CI/CD, and more.