agentic AI
42 stories tagged agentic AI.
OpenAI's Persistent AI Agent: What It Is and What It Costs
OpenAI is testing a persistent, always-on mode for its Codex AI agent. Here's what that actually means—and why the privacy history matters.
GLM 5.3 Flash Benchmarks and the Ox Alpha Reveal
GLM 5.3 Flash Benchmarks and the Ox Alpha Reveal
Theo's Ox Alpha turned out to be GLM 5.3 Flash — a tiny, cheap model punching well above its weight in agentic coding tasks.
IBM Granite 4.2: Open Reasoning Models With an Agent Brain
IBM Granite 4.2: Open Reasoning Models With an Agent Brain
IBM's Granite 4.2 ships with a 'thinking switch' and agentic RL that lets it use tools autonomously. Here's what that actually means—and why it matters.
Digital Librarian AI Agents Bridge SQL and Vector Data
Digital Librarian AI Agents Bridge SQL and Vector Data
IBM's Shad Griffin explains how Digital Librarian AI agents connect SQL and vector databases to answer questions that neither system can handle alone.
NVIDIA Nemotron 3.5 Lightning Targets AI Agent Work
NVIDIA Nemotron 3.5 Lightning Targets AI Agent Work
NVIDIA's Nemotron 3.5 Lightning is a 30B MoE model built to handle the repetitive, high-volume work inside AI agents—faster and cheaper than frontier reasoning models.
Meta Muse Glimmer 30B Tested: Agent Strength, Coding Limits
Meta Muse Glimmer 30B Tested: Agent Strength, Coding Limits
Meta's Muse Glimmer 30B is built for agentic workflows, not coding. Here's what it actually does well—and where the 82% hallucination rate should give you pause.
Hermes Agent v0.20 Brings Live Web Browsing to Desktop
Hermes Agent v0.20 Brings Live Web Browsing to Desktop
Nous Research's Hermes Agent v0.20 adds live in-app web browsing, real-time voice, grounded citations, and agent-to-agent communication to its desktop app.
Stateless MCP Makes the Protocol Worth Using Again
Stateless MCP Makes the Protocol Worth Using Again
Anthropic's latest MCP spec goes stateless, dropping the persistent connection requirement. Here's what changed, what it costs to upgrade, and why skeptics are reversing course.
Qwen 3.8 Max Tests Open-Source Against Big AI
Qwen 3.8 Max Tests Open-Source Against Big AI
Alibaba's Qwen 3.8 Max challenges OpenAI and Anthropic with multimodal capability, a 1M token context window, and open weights coming soon.
MiniMax Agent: Real Utility or Overhyped AI Tool?
MiniMax Agent: Real Utility or Overhyped AI Tool?
MiniMax Agent promises to replace prompting with delegation. But its own engineering docs reveal a catch. Here's what the hands-on testing actually shows.
AI Agent Hallucination: Causes, Risks, and Fixes
AI Agent Hallucination: Causes, Risks, and Fixes
AI agents hallucinate differently than chatbots—and the stakes are higher. Here's what's driving confident AI errors and how system design can reduce them.
Designing GUIs for AI Agents: An Unsolved Problem
Designing GUIs for AI Agents: An Unsolved Problem
AI agents need interfaces humans can actually understand and control. The design choices made now will shape whether AI goes mainstream or stays a developer toy.
Alexandr Wang on AI, Vision, and Building Frontier Labs
Alexandr Wang on AI, Vision, and Building Frontier Labs
Scale AI founder Alexandr Wang argues AI's bottleneck is adoption, not capability. Here's what his argument gets right — and what it leaves unexamined.
AI Observability Cuts Telecom Faults by 50 Percent
AI Observability Cuts Telecom Faults by 50 Percent
A New Zealand telco reports cutting IT incidents by 50% using AI-driven observability. Here's what that claim actually means—and what it doesn't.
Booking Holdings CEO on AI, Scale, and Survival
Booking Holdings CEO on AI, Scale, and Survival
Booking Holdings CEO Glenn Fogel survived the dot-com crash and now faces the AI wave. His take on moats, agentic travel, and job displacement is worth your time.
Small Language Models Are Reshaping Agentic AI
Small Language Models Are Reshaping Agentic AI
Small language models are outperforming larger rivals on key AI agent benchmarks. Here's what the efficiency shift means for how AI gets built and deployed.
AI Agents in Production: What Actually Works
AI Agents in Production: What Actually Works
IBM's Shailaja Patel-Pranav breaks down why AI agents fail in production—and the coordination patterns that make them actually reliable in enterprise workflows.
How MCP and AI Agents Are Reshaping Software Design
How MCP and AI Agents Are Reshaping Software Design
IBM's Will Scott explains how design systems, context engineering, and MCP are combining to let AI agents build software that actually follows the rules.
Loop Engineering: Moving Beyond One-Shot AI Prompting
Loop Engineering: Moving Beyond One-Shot AI Prompting
From cron-job automations to multi-day autonomous goals, loop engineering is changing how developers interact with AI. Here's what that actually means.
Agentic AI Is Reshaping Sports Business Economics
Agentic AI Is Reshaping Sports Business Economics
How autonomous AI agents are moving from experimentation to operational deployment across NFL teams, sponsorship sales, and broadcast production.