Research library

Insights

Analysis of AI, automation, and software: what is changing, what holds up, and what it means for practical decisions.

155 articles · Page 10 of 13

Insight5 min read

NVIDIA Nemotron 3 Super: The Open Model Built for Multi-Agent Systems

Nemotron 3 Super: 120B open model blending Mamba and Transformer, with 4x memory efficiency, 1M token context, and 7.5x throughput advantage.

Insight6 min read

Anthropic's Claude Code Review: AI-Generated Code Reviewed by AI

Multi-agent code review for Claude Code: filters false positives, ranks bugs by severity, automates GitHub PR feedback with structured inline comments.

Insight4 min read

Google Open-Sources Always On Memory Agent: Ditching Vector Databases for LLM-Driven Persistence

Always On Memory Agent: open-source reference implementation for persistent agent memory without vector databases. LLM-driven memory consolidation.

Insight4 min read

OpenAI GPT-5.4 Brings Native Computer Use and Spreadsheet Intelligence

The new model can operate your computer like a human and build financial models in Excel. Here is what matters for developers and enterprises.

Insight5 min read

Microsoft Phi-4-Reasoning-Vision: A Model That Knows When Not to Think

Phi-4 15B model skips structured reasoning for simple perception tasks, uses it only for math and science. Selective reasoning improves latency and cost.

Insight8 min read

GPT-5.4: OpenAI's Workplace AI With Native Computer Use

OpenAI launches GPT-5.4 with native computer use, 1M-token context, extreme thinking mode, and Excel/Sheets integration targeting enterprise workflows.

Insight3 min read

Alibaba's Qwen3.5-9B: A 9-Billion Parameter Model Beating 120-Billion

Alibaba's new small model series proves that running frontier-class AI on a laptop isn't a distant dream — it's happening now.

Insight4 min read

Block Just Cut 4,000 Jobs Because of AI. The Numbers Are Staggering.

Jack Dorsey says AI efficiency is driving Block to slash 40% of its workforce. The real story is what comes next.

Insight4 min read

How AT&T Cut AI Costs by 90% With Small Language Models

AT&T processes 8B tokens daily on specialized small language models, cutting AI costs ~90% without sacrificing capability. How the architecture works.

Insight3 min read

Perplexity Computer: A $20B Bet on Model Specialization

Perplexity Computer orchestrates 19 models in one agent. Models are specializing, not consolidating — multi-model orchestration, not commodities.

Insight4 min read

GGML and llama.cpp Join Hugging Face: What It Means for Local AI

GGML creator Georgi Gerganov joins Hugging Face to secure the future of local AI inference. A landmark open-source move.

Insight6 min read

Nano Banana 2: What Flash-Speed Image Generation with Web Grounding Actually Means

Google's Nano Banana 2 combines Pro-quality image generation with Flash speed and search grounding. What the architecture implies.