# Agent Infrastructure

Published articles for Agent Infrastructure.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## AI agent credentials: Four architectures compared

DevFeed: [AI agent credentials: Four architectures compared](<https://devfeed.tech/articles/ai-agent-credentials-four-architectures-compared-57501.md>)

Original publisher: [Read original article](<https://workos.com/blog/agent-credential-architectures>)

Author: WorkOS

Published: 2026-09-21T00:00:00Z

Content type: comparison

Language: en

Sources: [WorkOS](<https://devfeed.tech/sources/workos-blog.md>)

Topics: [AI Agent](<https://devfeed.tech/topics/ai-agent.md>), [Claude](<https://devfeed.tech/topics/claude.md>), [Google](<https://devfeed.tech/topics/google.md>), [Nvidia](<https://devfeed.tech/topics/nvidia.md>), [OpenShell](<https://devfeed.tech/topics/openshell.md>), [proxy](<https://devfeed.tech/topics/proxy.md>), [Language models](<https://devfeed.tech/topics/language-models.md>), [Network](<https://devfeed.tech/topics/network.md>), [OAuth 2.0](<https://devfeed.tech/topics/oauth2.md>), [Model Context Protocol](<https://devfeed.tech/topics/model-context-protocol.md>)

Tags: [agent](<https://devfeed.tech/tags/agent.md>), [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-agent](<https://devfeed.tech/tags/ai-agent.md>), [architectures](<https://devfeed.tech/tags/architectures.md>), [credentials](<https://devfeed.tech/tags/credentials.md>), [google](<https://devfeed.tech/tags/google.md>), [mcp](<https://devfeed.tech/tags/mcp.md>), [network](<https://devfeed.tech/tags/network.md>), [nvidia](<https://devfeed.tech/tags/nvidia.md>), [oauth2](<https://devfeed.tech/tags/oauth2.md>), [production](<https://devfeed.tech/tags/production.md>), [proxy](<https://devfeed.tech/tags/proxy.md>), [secrets](<https://devfeed.tech/tags/secrets.md>)

### AI overview

A comparison of four approaches to securing credentials for production AI agents. Google and NVIDIA keep secrets outside the sandbox, Rubrik issues tokens per tool call, Microsoft inspects MCP traffic, and Opal delegates request enforcement. The article focuses on which failures each architecture prevents and the gaps that remain.

### Source excerpt

Google, NVIDIA, Rubrik, Microsoft and Opal each solve agent credentials differently. What each one protects against, and the gap they share.

## OpenAI's Model-Misalignment Reporting Framework Offers an Incident Template for Agent Teams

DevFeed: [OpenAI's Model-Misalignment Reporting Framework Offers an Incident Template for Agent Teams](<https://devfeed.tech/articles/openai-s-misalignment-reports-are-an-agent-operations-signal-56256.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/openai-model-misalignment-reporting-framework>)

Author: Developers Digest

Published: 2026-09-19T00:00:00Z

Content type: opinion

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [OpenAI](<https://devfeed.tech/topics/openai.md>), [incident](<https://devfeed.tech/topics/incident.md>), [Template](<https://devfeed.tech/topics/template.md>), [Tool](<https://devfeed.tech/topics/tool.md>)

Tags: [agent](<https://devfeed.tech/tags/agent.md>), [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [agents](<https://devfeed.tech/tags/agents.md>), [ai-agents](<https://devfeed.tech/tags/ai-agents.md>), [ai-safety](<https://devfeed.tech/tags/ai-safety.md>), [developer-tools](<https://devfeed.tech/tags/developer-tools.md>), [disclosure](<https://devfeed.tech/tags/disclosure.md>), [incident](<https://devfeed.tech/tags/incident.md>), [openai](<https://devfeed.tech/tags/openai.md>), [reporting](<https://devfeed.tech/tags/reporting.md>), [signal](<https://devfeed.tech/tags/signal.md>), [template](<https://devfeed.tech/tags/template.md>)

### AI overview

The document presents OpenAI's model-misalignment reporting framework as a practical template for teams operating tool-using agents, covering incident intake, severity labels, and evidence-led disclosure.

### Source excerpt

OpenAI's model-misalignment reporting framework is not just a safety-policy document. For teams shipping tool-using agents, it is a template for incident intake, severity labels, and evidence-led disclosure.

## Claude Code Plugin Evals Make Agent Extensions Testable

DevFeed: [Claude Code Plugin Evals Make Agent Extensions Testable](<https://devfeed.tech/articles/claude-code-plugin-evals-make-agent-extensions-testable-55978.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/claude-code-plugin-evals-ci>)

Author: Developers Digest

Published: 2026-09-13T00:00:00Z

Content type: release

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [Claude Code](<https://devfeed.tech/topics/claude-code.md>), [Code](<https://devfeed.tech/topics/code.md>)

Tags: [agent](<https://devfeed.tech/tags/agent.md>), [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [agent-skills](<https://devfeed.tech/tags/agent-skills.md>), [ai-coding](<https://devfeed.tech/tags/ai-coding.md>), [ci](<https://devfeed.tech/tags/ci.md>), [claude-code](<https://devfeed.tech/tags/claude-code.md>), [developer-tools](<https://devfeed.tech/tags/developer-tools.md>), [evals](<https://devfeed.tech/tags/evals.md>), [html](<https://devfeed.tech/tags/html.md>), [json](<https://devfeed.tech/tags/json.md>), [plugin](<https://devfeed.tech/tags/plugin.md>), [reports](<https://devfeed.tech/tags/reports.md>)

### AI overview

Claude Code 2.1.269 introduces plugin evaluations, baseline comparisons, JSON and HTML reporting, and CI gates, making agent extensions measurable.

### Source excerpt

Claude Code 2.1.269 adds plugin evals, baseline comparisons, JSON and HTML reports, and CI gates. That changes plugins from clever prompts into measurable agent infrastructure.

## Hermes Agent Adds Vercel AI Gateway and Sandbox Backends

DevFeed: [Hermes Agent Adds Vercel AI Gateway and Sandbox Backends](<https://devfeed.tech/articles/hermes-agent-gains-vercel-ai-gateway-and-sandbox-backends-the-agent-stack-goes-plug-and-play-56170.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/hermes-agent-vercel-ai-gateway-sandbox-2026>)

Author: Developers Digest

Published: 2026-08-08T00:00:00Z

Content type: news

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Vercel](<https://devfeed.tech/topics/vercel.md>), [Model Routing](<https://devfeed.tech/topics/model-routing.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [gateway](<https://devfeed.tech/topics/gateway.md>), [Terminal](<https://devfeed.tech/topics/terminal.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>), [control-plane](<https://devfeed.tech/topics/control-plane.md>), [Back end](<https://devfeed.tech/topics/backend.md>)

Tags: [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-gateway](<https://devfeed.tech/tags/ai-gateway.md>), [cloud](<https://devfeed.tech/tags/cloud.md>), [coding-agents](<https://devfeed.tech/tags/coding-agents.md>), [control-plane](<https://devfeed.tech/tags/control-plane.md>), [model-routing](<https://devfeed.tech/tags/model-routing.md>), [news](<https://devfeed.tech/tags/news.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [terminal](<https://devfeed.tech/tags/terminal.md>), [vercel](<https://devfeed.tech/tags/vercel.md>)

### AI overview

Vercel added Hermes Agent to its AI Gateway and made Vercel Sandbox a terminal backend for the open-source agent. The setup supports user-provided model routing across more than 200 models without markup and a cloud microVM for each agent command.

### Source excerpt

Vercel added Hermes Agent to AI Gateway and made Vercel Sandbox a terminal backend for the open-source agent. Hermes is now fully BYO: your own model routing through 200+ models at no markup, and your own cloud microVM for every agent command. Here is what that unlocks and why the agent control plane is consolidating.

## Cloudflare Folds Workers AI Into AI Gateway: One Control Plane for Every Model Provider

DevFeed: [Cloudflare Folds Workers AI Into AI Gateway: One Control Plane for Every Model Provider](<https://devfeed.tech/articles/cloudflare-folds-workers-ai-into-ai-gateway-one-control-plane-for-every-model-provider-56022.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/cloudflare-workers-ai-gateway-unified-control-plane-2026>)

Author: Developers Digest

Published: 2026-08-07T00:00:00Z

Content type: release

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [Cloudflare](<https://devfeed.tech/topics/cloudflare.md>), [control-plane](<https://devfeed.tech/topics/control-plane.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [API](<https://devfeed.tech/topics/api.md>), [REST API](<https://devfeed.tech/topics/rest-api.md>), [Routing (disambiguation)](<https://devfeed.tech/topics/routing.md>)

Tags: [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-gateway](<https://devfeed.tech/tags/ai-gateway.md>), [api](<https://devfeed.tech/tags/api.md>), [cloudflare](<https://devfeed.tech/tags/cloudflare.md>), [control-plane](<https://devfeed.tech/tags/control-plane.md>), [gateway](<https://devfeed.tech/tags/gateway.md>), [model-routing](<https://devfeed.tech/tags/model-routing.md>), [news](<https://devfeed.tech/tags/news.md>), [rest-api](<https://devfeed.tech/tags/rest-api.md>), [routing](<https://devfeed.tech/tags/routing.md>)

### AI overview

Cloudflare is merging Workers AI and AI Gateway into a unified control plane with a shared /ai/ REST API, automatically created default gateways, transferable AI Gateway credits for Workers AI, and model-first provider routing.

### Source excerpt

Cloudflare is merging Workers AI and AI Gateway into one control plane: unified /ai/ REST API, auto-created default gateways, AI Gateway credits spendable on Workers AI, and model-first routing that picks the provider for you. Here is what changes and what stays.

## Chat SDK Adds Durable Approvals: Agent Workflows That Wait For a Human

DevFeed: [Chat SDK Adds Durable Approvals: Agent Workflows That Wait For a Human](<https://devfeed.tech/articles/chat-sdk-adds-durable-approvals-agent-workflows-that-wait-for-a-human-56349.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/vercel-chat-sdk-durable-approvals-2026>)

Author: Developers Digest

Published: 2026-08-06T00:00:00Z

Content type: release

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [SDK](<https://devfeed.tech/topics/sdk.md>), [Vercel](<https://devfeed.tech/topics/vercel.md>)

Tags: [agent](<https://devfeed.tech/tags/agent.md>), [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai-agents](<https://devfeed.tech/tags/ai-agents.md>), [chat](<https://devfeed.tech/tags/chat.md>), [clicks](<https://devfeed.tech/tags/clicks.md>), [developer-workflow](<https://devfeed.tech/tags/developer-workflow.md>), [handler](<https://devfeed.tech/tags/handler.md>), [news](<https://devfeed.tech/tags/news.md>), [sdk](<https://devfeed.tech/tags/sdk.md>), [table](<https://devfeed.tech/tags/table.md>), [thread](<https://devfeed.tech/tags/thread.md>), [vercel](<https://devfeed.tech/tags/vercel.md>), [workflow](<https://devfeed.tech/tags/workflow.md>), [workflows](<https://devfeed.tech/tags/workflows.md>)

### AI overview

Vercel's Chat SDK adds durable approval controls for agent workflows. A workflow can pause until a person approves an action in a chat thread, with the wait surviving deployments and decisions scoped to authorized approvers.

### Source excerpt

Vercel's Chat SDK can now suspend a Workflow SDK run until someone clicks Approve in a chat thread. One requestApproval call replaces the approvals table, the onAction handler, and the polling loop - with verified decisions, scoped approvers, and a wait that survives deploys.

## DeepSeek V4 Flash Is 90% Off Through Novita on Vercel AI Gateway: The Cost Math

DevFeed: [DeepSeek V4 Flash Is 90% Off Through Novita on Vercel AI Gateway: The Cost Math](<https://devfeed.tech/articles/deepseek-v4-flash-is-90-off-through-novita-on-vercel-ai-gateway-the-cost-math-56065.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/deepseek-v4-flash-novita-90-off-vercel-ai-gateway>)

Author: Developers Digest

Published: 2026-08-05T00:00:00Z

Content type: news

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [deepseek](<https://devfeed.tech/topics/deepseek.md>), [gateway](<https://devfeed.tech/topics/gateway.md>), [Vercel](<https://devfeed.tech/topics/vercel.md>)

Tags: [agent](<https://devfeed.tech/tags/agent.md>), [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-gateway](<https://devfeed.tech/tags/ai-gateway.md>), [cost](<https://devfeed.tech/tags/cost.md>), [deepseek](<https://devfeed.tech/tags/deepseek.md>), [effective](<https://devfeed.tech/tags/effective.md>), [flash](<https://devfeed.tech/tags/flash.md>), [gateway](<https://devfeed.tech/tags/gateway.md>), [input](<https://devfeed.tech/tags/input.md>), [news](<https://devfeed.tech/tags/news.md>), [pricing](<https://devfeed.tech/tags/pricing.md>), [routing](<https://devfeed.tech/tags/routing.md>), [setup](<https://devfeed.tech/tags/setup.md>), [tokens](<https://devfeed.tech/tags/tokens.md>), [vercel](<https://devfeed.tech/tags/vercel.md>)

### AI overview

DeepSeek V4 Flash is temporarily available through Novita on Vercel AI Gateway at a 90% discount for Pro customers, reducing the effective input and output rates to $0.014 and $0.028 per million tokens. The article presents the before-and-after cost calculation, provider-pinning setup, and implications for routing decisions in inexpensive agent loops.

### Source excerpt

DeepSeek V4 Flash routed to Novita on Vercel AI Gateway is 90% off for Pro customers through August 11, dropping the effective rate to $0.014 input / $0.028 output per million tokens. Here is the verified before/after math, the provider-pinning setup, and what a 10x cheap agent loop means for routing decisions.

## The v0 API Is GA: Vercel Just Made Its App-Building Agent a Headless Service

DevFeed: [The v0 API Is GA: Vercel Just Made Its App-Building Agent a Headless Service](<https://devfeed.tech/articles/the-v0-api-is-ga-vercel-just-made-its-app-building-agent-a-headless-service-56356.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/vercel-v0-api-ga-2026>)

Author: Developers Digest

Published: 2026-08-05T00:00:00Z

Content type: release

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [API](<https://devfeed.tech/topics/api.md>), [App](<https://devfeed.tech/topics/app.md>), [Vercel](<https://devfeed.tech/topics/vercel.md>), [service](<https://devfeed.tech/topics/service.md>), [prompt](<https://devfeed.tech/topics/prompt.md>), [async](<https://devfeed.tech/topics/async.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>)

Tags: [agent](<https://devfeed.tech/tags/agent.md>), [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai-agents](<https://devfeed.tech/tags/ai-agents.md>), [ai-coding](<https://devfeed.tech/tags/ai-coding.md>), [api](<https://devfeed.tech/tags/api.md>), [app](<https://devfeed.tech/tags/app.md>), [async](<https://devfeed.tech/tags/async.md>), [building](<https://devfeed.tech/tags/building.md>), [deploy](<https://devfeed.tech/tags/deploy.md>), [headless](<https://devfeed.tech/tags/headless.md>), [live](<https://devfeed.tech/tags/live.md>), [news](<https://devfeed.tech/tags/news.md>), [preview](<https://devfeed.tech/tags/preview.md>), [prompt](<https://devfeed.tech/tags/prompt.md>), [running](<https://devfeed.tech/tags/running.md>), [service](<https://devfeed.tech/tags/service.md>), [streaming](<https://devfeed.tech/tags/streaming.md>), [vercel](<https://devfeed.tech/tags/vercel.md>)

### AI overview

The v0 API is generally available, providing programmatic, headless access to v0's app-building agent. Developers can send a prompt, receive a running app with an embeddable live preview URL, and deploy it to Vercel in one call. The release also describes synchronous, asynchronous, and streaming operation in agent workflows.

### Source excerpt

The v0 API is now generally available: programmatic, headless access to v0's app-building agent. Send a prompt, get a running app with a live preview URL you can embed, then deploy to Vercel in one call. Here is what changed, how the sync/async/streaming model works, and how it fits in an agent loop.

## Vercel AI Gateway Adds Team and Project Spend Budgets: The Cost-Cap Math for Agent Builders

DevFeed: [Vercel AI Gateway Adds Team and Project Spend Budgets: The Cost-Cap Math for Agent Builders](<https://devfeed.tech/articles/vercel-ai-gateway-adds-team-and-project-spend-budgets-the-cost-cap-math-for-agent-builders-56348.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/vercel-ai-gateway-spend-budgets-2026>)

Author: Developers Digest

Published: 2026-08-01T00:00:00Z

Content type: release

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [gateway](<https://devfeed.tech/topics/gateway.md>), [Vercel](<https://devfeed.tech/topics/vercel.md>), [Command-line interface](<https://devfeed.tech/topics/cli.md>)

Tags: [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-gateway](<https://devfeed.tech/tags/ai-gateway.md>), [cli](<https://devfeed.tech/tags/cli.md>), [cost](<https://devfeed.tech/tags/cost.md>), [cost-control](<https://devfeed.tech/tags/cost-control.md>), [gateway](<https://devfeed.tech/tags/gateway.md>), [limits](<https://devfeed.tech/tags/limits.md>), [news](<https://devfeed.tech/tags/news.md>), [projects](<https://devfeed.tech/tags/projects.md>), [team](<https://devfeed.tech/tags/team.md>), [vercel](<https://devfeed.tech/tags/vercel.md>)

### AI overview

Vercel AI Gateway introduces spend budgets scoped to teams and projects. Hard dollar limits can reject requests, while email alerts trigger at 50%, 75%, and 100%; defaults can be managed through the CLI.

### Source excerpt

AI Gateway spend budgets now scope to teams and projects, with hard dollar limits that reject requests, email alerts at 50/75/100%, and CLI-managed defaults. Here is how the three scopes compose and where it fits your cost stack.

## Tencent Hunyuan Paper Maps Coding-Agent Behavior Across Prompts, Tools, State, Permissions, and Runtime Policy

DevFeed: [Tencent Hunyuan Paper Maps Coding-Agent Behavior Across Prompts, Tools, State, Permissions, and Runtime Policy](<https://devfeed.tech/articles/harness-handbook-shows-the-missing-map-for-coding-agents-56164.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/harness-handbook-agent-behavior-map>)

Author: Developers Digest

Published: 2026-07-16T00:00:00Z

Content type: article

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [code search](<https://devfeed.tech/topics/code-search.md>), [Tool](<https://devfeed.tech/topics/tool.md>), [Authorization](<https://devfeed.tech/topics/authorization.md>), [Prompt Engineering](<https://devfeed.tech/topics/prompt-engineering.md>)

Tags: [2026](<https://devfeed.tech/tags/2026.md>), [agent](<https://devfeed.tech/tags/agent.md>), [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai-agents](<https://devfeed.tech/tags/ai-agents.md>), [code-search](<https://devfeed.tech/tags/code-search.md>), [codex](<https://devfeed.tech/tags/codex.md>), [developer-workflow](<https://devfeed.tech/tags/developer-workflow.md>), [evals](<https://devfeed.tech/tags/evals.md>), [harness](<https://devfeed.tech/tags/harness.md>), [permissions](<https://devfeed.tech/tags/permissions.md>), [runtime](<https://devfeed.tech/tags/runtime.md>), [search](<https://devfeed.tech/tags/search.md>), [tencent](<https://devfeed.tech/tags/tencent.md>)

### AI overview

A July 2026 Tencent Hunyuan paper maps coding-agent behavior across prompts, tools, state, permissions, and runtime policy, arguing that code search alone is insufficient.

### Source excerpt

A July 2026 paper from Tencent Hunyuan turns agent harnesses into behavior-level maps. The useful lesson for builders is simple: code search is not enough when one behavior spans prompts, tools, state, permissions, and runtime policy.

## Image Token Compression Is a Real Agent Cost Lever

DevFeed: [Image Token Compression Is a Real Agent Cost Lever](<https://devfeed.tech/articles/image-token-compression-is-a-real-agent-cost-lever-56177.md>)

Original publisher: [Read original article](<https://www.developersdigest.tech/blog/image-token-compression-agent-costs>)

Author: Developers Digest

Published: 2026-07-04T00:00:00Z

Content type: article

Language: en

Sources: [Developers Digest](<https://devfeed.tech/sources/developers-digest.md>)

Topics: [Compression](<https://devfeed.tech/topics/compression.md>), [context](<https://devfeed.tech/topics/context.md>)

Tags: [accounting](<https://devfeed.tech/tags/accounting.md>), [agent](<https://devfeed.tech/tags/agent.md>), [agent-infrastructure](<https://devfeed.tech/tags/agent-infrastructure.md>), [ai-coding](<https://devfeed.tech/tags/ai-coding.md>), [claude-code](<https://devfeed.tech/tags/claude-code.md>), [compression](<https://devfeed.tech/tags/compression.md>), [context](<https://devfeed.tech/tags/context.md>), [cost](<https://devfeed.tech/tags/cost.md>), [cost-optimization](<https://devfeed.tech/tags/cost-optimization.md>), [evals](<https://devfeed.tech/tags/evals.md>)

### AI overview

A Show HN project claims that rendering bulky agent context as images can reduce costs. The document emphasizes evaluating compression, enforcing byte-safety rules, and accounting for each request.

### Source excerpt

A Show HN project claims large agent-cost cuts by rendering bulky context as images. The useful lesson is not the trick itself. It is that compression needs evals, byte-safety rules, and per-request accounting.