AI agents: what can they actually do?
From a good answer to a finished task. A practical look at the capabilities, the limits, and the human decisions in between.
Not just
about technology.
About people, too.
Independent
AI magazine
Since 2026
From a good answer to a finished task. A practical look at the capabilities, the limits, and the human decisions in between.
Anthropic emphasizes lower task costs, clearer writing, and practical coding.
ElevenLabs introduces a general model and a low-latency Turbo variant.
Google highlights experiments ranging from simulations to interactive engineering models.
GitHub’s walkthrough shows an agent building a board around the task.
Google combines real-time dialogue with a generated avatar for enterprise use.
Anthropic reports progress in complex tasks and behavioral evaluations.
OpenAI positions two models around different balances of capability and cost.
NVIDIA updates its GPU-accelerated packages built on ROS.
The speech-to-text update supports recorded and streaming workflows.
The proposed company would retain major operations in Canada and Germany.
Background notes preserve conventions, decisions, and durable facts.
An OpenAI case study looks at testing, software changes, and production monitoring.
Choose the task, the hardware, and the operating boundaries before choosing the model.
Cohere releases open weights with a non-commercial license.
The managed service brings the Codex harness to longer-running sessions.
The announcement spans coding, computer use, cybersecurity, and science.
The release pairs a general model with a controlled-access cybersecurity variant.
Google introduces a dynamic alternative to fixed-rate video processing.
A browser-focused library includes versioned operations and a hardware test suite.