Research library

Artificial intelligence

Evaluate where AI is useful, understand the systems behind it, and make informed choices about models, data, and deployment.

158 articles · Page 12 of 14

Insight4 min read

Claude Sonnet 4.6: Near-Opus at 1/5 the Cost

Claude Sonnet 4.6 matches Opus 4.6 on key benchmarks while costing 60% less, reshaping the AI price-performance curve.

Insight5 min read

Alibaba's Qwen 3.5: Multilingual and Open

Qwen 3.5 supports 201 languages, operates autonomously across devices, and ships open-weight under Apache 2.0.

Insight5 min read

MiniMax M2.5: Frontier Coding for Less

Chinese startup MiniMax matches Claude Opus 4.6 on SWE-bench while charging roughly one-tenth the price per token.

Insight6 min read

Seedance 2.0: AI Video Meets Copyright War

Seedance 2.0 debuts at China's Spring Festival Gala, generates Hollywood-quality video from text prompts, and draws a Disney cease-and-desist.

Insight9 min read

The AI Software Selloff: A $1T Wake-Up Call

AI product launches wiped $1 trillion from software valuations in one week. Here's what the selloff signals for your business.

Insight5 min read

Claude Opus 4.6: Adaptive Thinking, 1M Context

Claude Opus 4.6 introduces adaptive thinking, a 1M-token context window, and leads agentic benchmarks with 80.8% on SWE-bench.

Insight5 min read

GPT-5.3-Codex: OpenAI's Agentic Code Model

OpenAI releases GPT-5.3-Codex with 77.3% on Terminal-Bench 2.0, an agentic coding model partially used in its own creation.

Insight10 min read

Emergent AI: When Models Surprise Creators

Why large AI models develop surprising capabilities like arithmetic and reasoning that smaller models lack. Emergent behaviors explained.

Technical guide22 min read

Foundations of Transformer Reasoning

A technical deep-dive into transformer architectures, attention mechanisms, scaling laws, and emerging techniques for reliable AI reasoning.

Insight11 min read

Prompt Engineering: Better Results From AI

Practical techniques for writing effective prompts that produce reliable AI outputs. Works across ChatGPT, Claude, Gemini, and other LLMs.

Technical guide14 min read

Taxonomy of AI: From ML to World Models

A map of AI systems — machine learning, deep learning, LLMs, multimodal models, and world models — with clear definitions and comparisons.

Insight10 min read

Why Bigger AI Models Work Better

The science behind AI scaling laws and chain-of-thought reasoning, without the PhD. Why larger models are smarter and how to use them.