79 related articles
Deep DivesDeep dive into NousResearch's open-source Hermes Agent self-evolution framework, using DSPy and GEPA for automated prompt optimization with five-layer safety mechanisms.
Product ReviewsHands-on test of Trae IDE with Doubao Seed 2.0 building a Django+Vue3 book management system for free, benchmarked against Gemini 2.5 and MiniMax models.
TutorialsA hands-on tutorial for building a financial report analysis AI Agent from scratch using Cursor editor, Skills definitions, and MiniMax M2.1. Covers setup, architecture, Skills methodology, and multi-language programming.
Industry InsightsMicrosoft bans Claude Code internally, forcing engineers to GitHub Copilot CLI. Analysis of the cost crisis, product gap, and AI ecosystem control battle reshaping the industry.
Tech FrontiersAlibaba's Qwen APP launches 400+ features integrating Alipay and Taobao, Baidu releases ERNIE 5.0, Meituan unveils deep reasoning model, StepFun tops global speech AI rankings, and Anthropic's share nears Google's.
Product ReviewsA hands-on test of Zhipu GLM5.1 in full-stack development — building an AI canvas app from scratch to evaluate improvements in problem understanding, debugging, and multi-agent collaboration.
TutorialsDeep dive into Claude Skills 2.0: two skill types, the new skill creator, evaluation system, and a cold email marketing case study that boosts task pass rates from 40% to 100%.
Product ReviewsBenchLocal real-world testing of DeepSeek V4 Pro, V4 Flash vs Qwen3.6 27B across 8 categories and 85 scenarios. V4 Pro leads by 6% but stumbles on math reasoning. Qwen3.6 Q6 rivals V4 Pro in agent tasks.
Product ReviewsHands-on testing of Manus general AI Agent across history report generation, Tesla stock analysis, and GAIA benchmarks. Compares vertical vs general agents with scoring data and limitation analysis.
Deep DivesDeep dive into how the MARVIS project deploys LLM agents on spacecraft, covering agent architecture, edge hardware token performance benchmarks, expert evaluations, and space AI benchmark planning.
TutorialsComplete guide to deploying OpenAI's open-source GPT-OSS model locally with Ollama. Real-world testing of the 20B version on RTX 4090 covering Chinese comprehension, logical reasoning, and VRAM usage analysis under MoE architecture.
Product ReviewsDeep dive into ByteDance's open-source Trae Agent — the free AI coding CLI tool topping SWE-bench. Covers installation, features, and comparisons with Claude Code and Gemini CLI.
TutorialsLearn how to connect an AI Agent with the Shiji task management tool to auto-generate and push daily to-dos. Full walkthrough from API setup to deep conversation training.
Tech FrontiersGoogle acquires Windsurf's core team via talent deal, Gemini 3.0 code leaks hint at new models, and OpenAI delays its open model indefinitely. Deep analysis of the AI industry's talent, model, and open-source battles in 2025.
Tech FrontiersMicrosoft revokes internal Claude Code licenses despite widespread praise, prioritizing GitHub Copilot's competitive position in the intensifying AI coding tools market battle.
Product ReviewsComprehensive comparison of 80+ AI coding agent tools, with SWE-Bench benchmark rankings covering Devin, Cursor, Claude Code, GitHub Copilot and more, plus pricing analysis to help developers choose.
Product ReviewsDeep dive into the crafta-bench open-source project, a benchmark tool designed for Cursor Background Agents. Explore AI coding Agent evaluation dimensions, industry trends, and practical implications.
Product ReviewsExplore Hugging Face Transformers, the 160K-star open-source framework for AI models covering text, vision, audio, and multimodal with unified APIs.