Research library
Artificial intelligence
Evaluate where AI is useful, understand the systems behind it, and make informed choices about models, data, and deployment.
158 articles · Page 12 of 14
Claude Sonnet 4.6: Near-Opus at 1/5 the Cost
Claude Sonnet 4.6 matches Opus 4.6 on key benchmarks while costing 60% less, reshaping the AI price-performance curve.
Alibaba's Qwen 3.5: Multilingual and Open
Qwen 3.5 supports 201 languages, operates autonomously across devices, and ships open-weight under Apache 2.0.
MiniMax M2.5: Frontier Coding for Less
Chinese startup MiniMax matches Claude Opus 4.6 on SWE-bench while charging roughly one-tenth the price per token.
Seedance 2.0: AI Video Meets Copyright War
Seedance 2.0 debuts at China's Spring Festival Gala, generates Hollywood-quality video from text prompts, and draws a Disney cease-and-desist.
The AI Software Selloff: A $1T Wake-Up Call
AI product launches wiped $1 trillion from software valuations in one week. Here's what the selloff signals for your business.
Claude Opus 4.6: Adaptive Thinking, 1M Context
Claude Opus 4.6 introduces adaptive thinking, a 1M-token context window, and leads agentic benchmarks with 80.8% on SWE-bench.
GPT-5.3-Codex: OpenAI's Agentic Code Model
OpenAI releases GPT-5.3-Codex with 77.3% on Terminal-Bench 2.0, an agentic coding model partially used in its own creation.
Emergent AI: When Models Surprise Creators
Why large AI models develop surprising capabilities like arithmetic and reasoning that smaller models lack. Emergent behaviors explained.
Foundations of Transformer Reasoning
A technical deep-dive into transformer architectures, attention mechanisms, scaling laws, and emerging techniques for reliable AI reasoning.
Prompt Engineering: Better Results From AI
Practical techniques for writing effective prompts that produce reliable AI outputs. Works across ChatGPT, Claude, Gemini, and other LLMs.
Taxonomy of AI: From ML to World Models
A map of AI systems — machine learning, deep learning, LLMs, multimodal models, and world models — with clear definitions and comparisons.
Why Bigger AI Models Work Better
The science behind AI scaling laws and chain-of-thought reasoning, without the PhD. Why larger models are smarter and how to use them.
