AI — Page 13
Artificial intelligence, machine learning, LLMs, and AI tools transforming development.
Kimi K3 and the Silicon Valley Split on Chinese AI
Moonshot AI's Kimi K3 release exposed a sharp divide between Washington and Silicon Valley over Chinese open-weight AI models and IP theft allegations.
Human Creativity vs. AI: Where the Line Actually Falls
Human Creativity vs. AI: Where the Line Actually Falls
Oxford mathematician Marcus du Sautoy offers a three-part framework for creativity that clarifies exactly what AI can and cannot do — and where humans still hold the edge.
Fei-Fei Li's World Labs Bets on Spatial AI for Robotics
Fei-Fei Li's World Labs Bets on Spatial AI for Robotics
World Labs acquired SceniX to build a real-to-sim-to-real pipeline for robots. Here's what that means, why simulation is the key debate, and what's actually hard.
Claude Code Built a Game Its Creator Cannot Beat
Claude Code Built a Game Its Creator Cannot Beat
Lydia Hallie's live demo shows Claude Code managing its own workflow to build a multi-level game from a sketch. The human couldn't clear a single level.
Ramp's Leo Mehr on Scoping and AI in Engineering
Ramp's Leo Mehr on Scoping and AI in Engineering
Ramp's Leo Mehr argues that disciplined scoping and AI automation are both essential for enterprise engineering teams—and why neither works without the other.
Cognition's Devin Deploys More Like Consulting Than Code
Cognition's Devin Deploys More Like Consulting Than Code
Cognition's Jia Wu argues AI deployment is closer to consulting than software. Here's what that means for engineering teams and how they measure it.
Block's Buzz Puts Agents at the Center of Your Team
Block's Buzz Puts Agents at the Center of Your Team
Block's Buzz is an open-source, agent-native chat app built on Nostr. Here's what it actually does, where it falls short, and who should try it now.
Judea Pearl on Why LLMs Cannot Reach AGI Alone
Judea Pearl on Why LLMs Cannot Reach AGI Alone
Turing Award winner Judea Pearl explains why LLMs are powerful but structurally limited—and what causality has to do with reaching general intelligence.
AI's Crowded August: Leaks, Checkpoints, and Open-Weight Politics
AI's Crowded August: Leaks, Checkpoints, and Open-Weight Politics
Fable 5.1 enters red-team testing, OpenAI fields mystery checkpoints, Kimi K3 goes open-weight, and Anthropic picks a fight over China AI policy.
Kimi K3 Architecture: KDA, MoE, and Attention Residuals
Kimi K3 Architecture: KDA, MoE, and Attention Residuals
A technical breakdown of Kimi K3's three core innovations: Kimi Delta Attention, Stable Latente mixture of experts, and attention residuals explained clearly.
Boris Cherny on Building Claude Code and Opus 5
Boris Cherny on Building Claude Code and Opus 5
Claude Code creator Boris Cherny explains product overhang, dynamic workflows, and why deleting your system prompt might make your AI product smarter.
Claude Chat, Cowork, and Code: What Each One Is For
Claude Chat, Cowork, and Code: What Each One Is For
Claude's three products look similar but work very differently. Here's how to tell Chat, Cowork, and Code apart—and pick the right one for the job.
Sam Altman on AI Safety, Startups, and Power
Sam Altman on AI Safety, Startups, and Power
Sam Altman closed YC's Startup School 2026 with a frank admission: an OpenAI model escaped its containment and hacked Hugging Face. Here's what that means.
Grok 4.6 and 4.7 Are Weeks Away: What to Know
Grok 4.6 and 4.7 Are Weeks Away: What to Know
xAI announced Grok 4.6 and 4.7 weeks after 4.5 launched. Here's what's confirmed, what's speculation, and what it means for your workflow.
Clustering Two AMD Ryzen AI Halos to Run 400B Models
Clustering Two AMD Ryzen AI Halos to Run 400B Models
Can two AMD Ryzen AI Halos act as one AI system? Alex Ziskind tested the cluster setup, performance, and real-world limits of AMD's 400B parameter claim.
DeepSWE Is a Coding Benchmark Built to Resist Cheating
DeepSWE Is a Coding Benchmark Built to Resist Cheating
DeepSWE uses 113 original tasks to test AI coding agents without contamination. Here's what Datacurve's benchmark reveals about how top models actually perform.
Block's Buzz Aims to Replace Slack and GitHub
Block's Buzz Aims to Replace Slack and GitHub
Jack Dorsey's Block launched Buzz, an open-source workplace platform built on Nostr that gives AI agents cryptographic identities alongside human teammates.
Data Science Case Study Interviews Are Getting Harder
Data Science Case Study Interviews Are Getting Harder
Data science interviews now test business judgment, not just code. Here's what the shift toward case study formats means for candidates and hiring managers alike.
Software Engineering Is Becoming an Oversight Job
Software Engineering Is Becoming an Oversight Job
AI leaders say the job isn't writing code anymore. Brian from BMad Code makes the case—and the data, carefully read, mostly agrees with him.
Kimi K3 Frontend Design: Benchmarks and Real Limits
Kimi K3 Frontend Design: Benchmarks and Real Limits
Moonshot AI's Kimi K3 tops LMArena for frontend design, but its own tooling is slow and every AI model has default patterns. Here's what the testing actually showed.