Python and TypeScript teams can move a code-first agent into a hosted LangSmith runtime that supplies durable execution, sandboxes, memory, identity and evaluation plumbing.
Archive
Latest AI intelligence — Page 3
AI reporting, analysis, and product updates across every desk, ordered by publication time.
387 sourced postsOne package can now carry skills and MCP server configuration across VS Code, Copilot CLI, the SDK and Copilot app, with enterprise controls applying across supported clients.
All ChatGPT plans can surface available tables and refine choices in conversation, with partner coverage varying by country and Work excluded.
The 11,016-instance benchmark separates static recognition from interactive planning, exposing a large gap between seeing topology and preserving it across actions.
The research system combines generated chest X-ray reports with confidence-scored outputs and deterministic measurements, but its retrospective results do not establish clinical safety or approval.
OpenAI has announced a higher-capacity seat with monthly and annual pricing, but general availability remains pending and workspace owners must still join a waitlist.
Approved security teams can now use Daybreak Red and Blue inside AWS environments through Bedrock, adding a governed cloud route for defensive and authorized testing workflows.
Two new studies suggest agent use is spreading beyond engineering, while a widening usage gap separates OpenAI's most active business customers from typical adopters.
Copilot customers can now trace input, output and cached tokens behind AI-credit consumption instead of seeing only the final credit charge.
Microsoft's smaller coding model is rolling out across Copilot with a lower stated list price, while teams have until September 10 to migrate from its predecessor.
The IDE plugin can carry selected context across agent chats, connect to Ollama as a bring-your-own-key provider, expose Codex sessions in debug logs and accept tighter enterprise controls.
The advertising pilot is now live in the United Kingdom, Mexico, Brazil, Japan and South Korea, extending a model that keeps paid plans ad-free and gives entry-tier users personalization controls.
Developers can route supported open-weight language models to Baseten from Hugging Face model pages and client libraries, using either provider credentials or Hugging Face billing.
The replay-based evaluation tests whether language models know when to scaffold and when to push students to reason, while its authors caution that simulated sessions do not measure learning.
The unified endpoint now leads Google's model-and-agent developer stack, adding managed sandboxes, background jobs and optional server-side state while the older API remains supported.
The August 7 change opens more everyday health, educational and clinical queries while continuing to reroute dual-use research requests to Opus 5.
The cloud-hosted browser runs in V8 isolates and is designed around AI-agent automation, offering developers a lighter alternative for some browser-based tasks.
The open-weight speech model adds Arabic, Korean and Brazilian Portuguese voices, plus a production NIM for teams deploying voice agents on their own infrastructure.
The Apache-2.0 model targets always-on agents on consumer hardware, combining tool use, long-horizon reasoning and interleaved text-and-image input.
The 30-billion-parameter open model targets specialised agent tasks, while an open routing library directs requests across mixed-model systems.
The voice stack separates real-time audio from tool calls, model delegation and context maintenance to avoid audible stalls.
Repository teams can launch documentation, error-investigation and follow-up workflows from issue or pull-request comments.
Paid-plan users can choose how much supported models reason for each delegated task, trading potential quality gains against token and credit use.
Existing users have a short export window, while deployed apps should continue running and AI-powered projects need a replacement inference provider.