Research library
Insights
Analysis of AI, automation, and software: what is changing, what holds up, and what it means for practical decisions.
155 articles · Page 10 of 13
NVIDIA Nemotron 3 Super: The Open Model Built for Multi-Agent Systems
Nemotron 3 Super: 120B open model blending Mamba and Transformer, with 4x memory efficiency, 1M token context, and 7.5x throughput advantage.
Anthropic's Claude Code Review: AI-Generated Code Reviewed by AI
Multi-agent code review for Claude Code: filters false positives, ranks bugs by severity, automates GitHub PR feedback with structured inline comments.
Google Open-Sources Always On Memory Agent: Ditching Vector Databases for LLM-Driven Persistence
Always On Memory Agent: open-source reference implementation for persistent agent memory without vector databases. LLM-driven memory consolidation.
OpenAI GPT-5.4 Brings Native Computer Use and Spreadsheet Intelligence
The new model can operate your computer like a human and build financial models in Excel. Here is what matters for developers and enterprises.
Microsoft Phi-4-Reasoning-Vision: A Model That Knows When Not to Think
Phi-4 15B model skips structured reasoning for simple perception tasks, uses it only for math and science. Selective reasoning improves latency and cost.
GPT-5.4: OpenAI's Workplace AI With Native Computer Use
OpenAI launches GPT-5.4 with native computer use, 1M-token context, extreme thinking mode, and Excel/Sheets integration targeting enterprise workflows.
Alibaba's Qwen3.5-9B: A 9-Billion Parameter Model Beating 120-Billion
Alibaba's new small model series proves that running frontier-class AI on a laptop isn't a distant dream — it's happening now.
Block Just Cut 4,000 Jobs Because of AI. The Numbers Are Staggering.
Jack Dorsey says AI efficiency is driving Block to slash 40% of its workforce. The real story is what comes next.
How AT&T Cut AI Costs by 90% With Small Language Models
AT&T processes 8B tokens daily on specialized small language models, cutting AI costs ~90% without sacrificing capability. How the architecture works.
Perplexity Computer: A $20B Bet on Model Specialization
Perplexity Computer orchestrates 19 models in one agent. Models are specializing, not consolidating — multi-model orchestration, not commodities.
GGML and llama.cpp Join Hugging Face: What It Means for Local AI
GGML creator Georgi Gerganov joins Hugging Face to secure the future of local AI inference. A landmark open-source move.
Nano Banana 2: What Flash-Speed Image Generation with Web Grounding Actually Means
Google's Nano Banana 2 combines Pro-quality image generation with Flash speed and search grounding. What the architecture implies.
