<?xml version="1.0" encoding="utf-8"?>
<feed xmlns="http://www.w3.org/2005/Atom">
  <id>https://newruntime.com/feeds/coding-agents.xml</id>
  <title>New Runtime: Coding agents</title>
  <subtitle>A living map of coding agents as they move from autocomplete into supervised software work, review loops, environments, and operating cost. This bounded feed exists because the topic is a featured, evidence-dense New Runtime hub.</subtitle>
  <updated>2026-09-01T00:00:00.000Z</updated>
  <link rel="self" type="application/atom+xml" href="https://newruntime.com/feeds/coding-agents.xml"/>
  <link rel="hub" href="https://pubsubhubbub.appspot.com/"/>
  <link rel="alternate" type="text/html" href="https://newruntime.com/topics/coding-agents/"/>
  <author><name>New Runtime</name><uri>https://newruntime.com/</uri></author>
  <rights>New Runtime tracks the systems, signals, and shifts reshaping software as agents become first-class participants in work, products, and organizations.</rights>
  <entry>
    <id>https://newruntime.com/patterns/verification-bandwidth-is-the-scarce-resource/</id>
    <title>Verification bandwidth is the scarce engineering resource</title>
    <link rel="alternate" href="https://newruntime.com/patterns/verification-bandwidth-is-the-scarce-resource/"/>
    <published>2026-09-01T00:00:00.000Z</published>
    <updated>2026-09-01T00:00:00.000Z</updated>
    <summary>The primary bottleneck in agentic software delivery is moving from code production to the human and machine capacity required to verify it.</summary>
    <category term="coding-agents"/>
    <category term="engineering-management"/>
    <category term="evals"/>
    <link rel="related" href="https://alignment.anthropic.com/2026/automated-alignment-researchers/"/>
    <link rel="related" href="https://www.terminal-bench-science.ai/announcement"/>
    <link rel="related" href="https://thinkingmachines.ai/news/putting-task-expertise-into-rl/"/>
    <link rel="related" href="https://docs.langchain.com/langsmith/insights"/>
    <link rel="related" href="https://addyo.substack.com/p/own-the-outer-loop"/>
    <link rel="related" href="https://databricks.com/blog/benchmarking-coding-agents-databricks-multi-million-line-codebase"/>
    <link rel="related" href="https://hamel.dev/blog/posts/eval-smell"/>
    <link rel="related" href="https://addyosmani.com/blog/new-sdlc-vibe-coding"/>
    <link rel="related" href="https://anthropic.com/research/AI-assistance-coding-skills"/>
  </entry>
  <entry>
    <id>https://newruntime.com/projects/coding-agent-cost-controls/</id>
    <title>Coding Agent Cost Controls</title>
    <link rel="alternate" href="https://newruntime.com/projects/coding-agent-cost-controls/"/>
    <published>2026-08-15T00:00:00.000Z</published>
    <updated>2026-08-15T00:00:00.000Z</updated>
    <summary>A forkable setup for reducing coding-agent cost by shaping environment output, repo maps, model routing, cache stability, and task-level budgets.</summary>
    <category term="coding-agents"/>
    <category term="cost-control"/>
    <link rel="related" href="https://tech.autoscout24.com/blog/posts/3-techniques-to-reduce-token-consumption-claude-code-codex/"/>
    <link rel="related" href="https://github.com/rtk-ai/rtk"/>
  </entry>
  <entry>
    <id>https://newruntime.com/projects/verifiable-ai-coding-workflow/</id>
    <title>Verifiable AI Coding Workflow</title>
    <link rel="alternate" href="https://newruntime.com/projects/verifiable-ai-coding-workflow/"/>
    <published>2026-08-06T00:00:00.000Z</published>
    <updated>2026-08-06T00:00:00.000Z</updated>
    <summary>A small agent-assisted coding workflow where the useful artifact is not generated code, but a checked path from idea to tests, diff, review, and handoff.</summary>
    <category term="coding-agents"/>
    <category term="developer-workflow"/>
    <link rel="related" href="https://github.com/mattpocock/skills"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/agentic-sdlc-software-factory-loop/</id>
    <title>A Software Factory Connects Agents Through Verified Outcomes</title>
    <link rel="alternate" href="https://newruntime.com/posts/agentic-sdlc-software-factory-loop/"/>
    <published>2026-08-03T00:00:00.000Z</published>
    <updated>2026-08-03T00:00:00.000Z</updated>
    <summary>Augment and Warp describe team-level agent loops that move work from trigger and specification through implementation, verification, release, and measured improvement.</summary>
    <category term="agent-harness"/>
    <category term="coding-agents"/>
    <category term="evals"/>
    <category term="software-factories"/>
    <link rel="related" href="https://www.augmentcode.com/blog/what-is-loop-engineering-and-how-are-leading-software-engineering-teams-using-it"/>
    <link rel="related" href="https://www.warp.dev/blog/software-factory-build-guide"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/claude-code-auto-mode-action-gate/</id>
    <title>Claude Code Auto Mode Gates Actions Instead Of Explanations</title>
    <link rel="alternate" href="https://newruntime.com/posts/claude-code-auto-mode-action-gate/"/>
    <published>2026-08-03T00:00:00.000Z</published>
    <updated>2026-08-03T00:00:00.000Z</updated>
    <summary>Claude Code Auto Mode combines an input injection probe with a two-stage action classifier, preserving autonomy while exposing an honest residual miss rate.</summary>
    <category term="agent-harness"/>
    <category term="agent-security"/>
    <category term="coding-agents"/>
    <category term="evals"/>
    <link rel="related" href="https://www.anthropic.com/engineering/claude-code-auto-mode"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/cline-hooks-agent-harness-guardrails/</id>
    <title>Cline Hooks Put Deterministic Rules Inside The Agent Loop</title>
    <link rel="alternate" href="https://newruntime.com/posts/cline-hooks-agent-harness-guardrails/"/>
    <published>2026-08-03T00:00:00.000Z</published>
    <updated>2026-08-03T00:00:00.000Z</updated>
    <summary>Cline&apos;s plugin hooks show how an agent harness can journal every run and block dangerous tool calls without waiting for the model to choose a guardrail.</summary>
    <category term="agent-harness"/>
    <category term="coding-agents"/>
    <category term="mcp"/>
    <category term="observability"/>
    <link rel="related" href="https://cline.bot/blog/extend-cline-with-plugins-and-hooks"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/evocode-bench-multi-turn-regressions/</id>
    <title>EvoCode-Bench Exposes Multi-Turn Regression Risk</title>
    <link rel="alternate" href="https://newruntime.com/posts/evocode-bench-multi-turn-regressions/"/>
    <published>2026-08-03T00:00:00.000Z</published>
    <updated>2026-08-03T00:00:00.000Z</updated>
    <summary>EvoCode-Bench tests coding agents across persistent workspaces and evolving requirements, where regressions become the dominant failure mode.</summary>
    <category term="benchmarks"/>
    <category term="coding-agents"/>
    <category term="evals"/>
    <category term="regressions"/>
    <link rel="related" href="https://www.philschmid.de/evocode-bench"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/google-adk-loop-engineering-state-pruning/</id>
    <title>Loop Engineering Needs State Pruning, Not Infinite Chat</title>
    <link rel="alternate" href="https://newruntime.com/posts/google-adk-loop-engineering-state-pruning/"/>
    <published>2026-08-03T00:00:00.000Z</published>
    <updated>2026-08-03T00:00:00.000Z</updated>
    <summary>The Google ADK loop-engineering article frames self-correcting agents as desired-state systems with pruning, validation, and circuit breakers.</summary>
    <category term="agent-runtime"/>
    <category term="coding-agents"/>
    <category term="google-adk"/>
    <category term="loop-engineering"/>
    <link rel="related" href="https://medium.com/google-cloud/loop-engineering-in-self-correcting-code-migration-using-google-adk-2-0-61c30c9e36ca"/>
  </entry>
  <entry>
    <id>https://newruntime.com/shifts/coding-agent-general-workbench/</id>
    <title>Coding tool -&gt; general-purpose workbench</title>
    <link rel="alternate" href="https://newruntime.com/shifts/coding-agent-general-workbench/"/>
    <published>2026-08-01T00:00:00.000Z</published>
    <updated>2026-08-01T00:00:00.000Z</updated>
    <summary>Coding agents are expanding beyond software implementation into context-aware workbenches that assemble prototypes, interfaces, documents, workflows, and operational artifacts.</summary>
    <category term="agent-workbench"/>
    <category term="coding-agents"/>
    <category term="generative-ui"/>
    <category term="knowledge-work"/>
    <category term="skills"/>
    <link rel="related" href="https://www.anthropic.com/product/claude-code"/>
    <link rel="related" href="https://developers.openai.com/codex/use-cases"/>
    <link rel="related" href="https://hermes-agent.nousresearch.com/docs/user-guide/features/overview/"/>
    <link rel="related" href="https://docs.openclaw.ai/agent-workspace"/>
    <link rel="related" href="https://www.chatprd.ai/how-i-ai/stripe-owen-williams-on-buildling-internal-prototyping-studio"/>
    <link rel="related" href="https://github.com/1weiho/open-slide"/>
    <link rel="related" href="https://blog.google/innovation-and-ai/models-and-research/google-labs/stitch-design-md/"/>
    <link rel="related" href="https://modelcontextprotocol.io/extensions/apps/overview"/>
    <link rel="related" href="https://github.com/agentskills/agentskills"/>
    <link rel="related" href="https://openai.com/index/chatgpt-for-excel/"/>
    <link rel="related" href="https://www.microsoft.com/en-us/microsoft-365/blog/2026/04/22/copilots-agentic-capabilities-in-word-excel-and-powerpoint-are-generally-available/"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/openai-gpt-5-6-price-performance-frontier/</id>
    <title>GPT-5.6 Turns Efficiency Work Into API Economics</title>
    <link rel="alternate" href="https://newruntime.com/posts/openai-gpt-5-6-price-performance-frontier/"/>
    <published>2026-08-01T00:00:00.000Z</published>
    <updated>2026-08-01T00:00:00.000Z</updated>
    <summary>OpenAI turned GPT-5.6 serving and kernel efficiency gains into lower Luna and Terra prices, plus a faster Sol mode for latency-sensitive API workloads.</summary>
    <category term="coding-agents"/>
    <category term="enterprise-ai"/>
    <category term="inference"/>
    <category term="model-economics"/>
    <link rel="related" href="https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/langchain-reviewbench-review-agent-evals/</id>
    <title>ReviewBench Turns Code Review Into An Agent Eval</title>
    <link rel="alternate" href="https://newruntime.com/posts/langchain-reviewbench-review-agent-evals/"/>
    <published>2026-08-01T00:00:00.000Z</published>
    <updated>2026-08-01T00:00:00.000Z</updated>
    <summary>LangChain&apos;s ReviewBench uses real PR review history to test whether code-review agents can recover substantive reviewer findings without flooding humans with noise.</summary>
    <category term="coding-agents"/>
    <category term="developer-tools"/>
    <category term="evals"/>
    <category term="verification"/>
    <link rel="related" href="https://x.com/LangChain/status/2083236117839499511"/>
    <link rel="related" href="https://www.langchain.com/blog/towards-automating-eval-engineering"/>
    <link rel="related" href="https://www.langchain.com/blog/unified-stack-for-evaluating-agents"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/github-stacked-prs-agent-review/</id>
    <title>GitHub Stacked PRs Turn Large Agent Changes Into Reviewable Chains</title>
    <link rel="alternate" href="https://newruntime.com/posts/github-stacked-prs-agent-review/"/>
    <published>2026-07-31T00:00:00.000Z</published>
    <updated>2026-07-31T00:00:00.000Z</updated>
    <summary>GitHub&apos;s stacked pull request preview gives large dependent code changes a native review path, which matters as agents produce broader diffs.</summary>
    <category term="coding-agents"/>
    <category term="developer-tools"/>
    <category term="verification"/>
    <category term="workflows"/>
    <link rel="related" href="https://x.com/github/status/2082894271653306445"/>
    <link rel="related" href="https://docs.github.com/en/pull-requests/how-tos/stacked-pull-requests"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/cline-recursive-self-improvement-coding-agent/</id>
    <title>Cline Turns Recursive Self-Improvement Into Harness Work</title>
    <link rel="alternate" href="https://newruntime.com/posts/cline-recursive-self-improvement-coding-agent/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>Cline&apos;s Terminal-Bench run is not a singularity story; it is a concrete loop where an agent reads traces, patches the harness, reruns evals, and hands a PR to humans.</summary>
    <category term="agent-economics"/>
    <category term="agent-harness"/>
    <category term="coding-agents"/>
    <category term="evals"/>
    <category term="open-source"/>
    <link rel="related" href="https://x.com/cline/status/2082544251519611187"/>
    <link rel="related" href="https://cline.bot/blog/recursive-self-improvement-for-coding-agents"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/cursor-cloud-agent-environment-product/</id>
    <title>Cursor Treats the Cloud Agent Environment as the Product</title>
    <link rel="alternate" href="https://newruntime.com/posts/cursor-cloud-agent-environment-product/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>Cursor&apos;s cloud-agent environment write-up shows why agent performance depends on dependencies, commands, security boundaries, end-to-end tests, and self-healing diagnostics.</summary>
    <category term="agent-environments"/>
    <category term="coding-agents"/>
    <category term="developer-tools"/>
    <category term="verification"/>
    <link rel="related" href="https://x.com/cursor_ai/status/2082841399838327289"/>
    <link rel="related" href="https://cursor.com/blog/cloud-agent-environment"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/cursor-agent-swarms-sqlite/</id>
    <title>Cursor&apos;s SQLite Swarm Makes Coordination the Expensive Part</title>
    <link rel="alternate" href="https://newruntime.com/posts/cursor-agent-swarms-sqlite/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>Cursor&apos;s SQLite experiment shows why agent-swarm economics depend on task trees, shared memory, conflict handling, review lenses, and selective use of expensive planners.</summary>
    <category term="agents"/>
    <category term="coding-agents"/>
    <category term="developer-tools"/>
    <link rel="related" href="https://cursor.com/blog/agent-swarm-model-economics"/>
    <link rel="related" href="https://github.com/cursor/minisqlite"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/factory-comarch-night-shift-agents/</id>
    <title>Factory and Comarch Show the Night-Shift Shape of Agent Work</title>
    <link rel="alternate" href="https://newruntime.com/posts/factory-comarch-night-shift-agents/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>Factory&apos;s Comarch case study is less about one coding assistant and more about governed agent missions that continue execution while humans set direction and review.</summary>
    <category term="agent-orchestration"/>
    <category term="coding-agents"/>
    <category term="enterprise-ai"/>
    <category term="software-delivery"/>
    <link rel="related" href="https://x.com/FactoryAI/status/2082834849925046589"/>
    <link rel="related" href="https://factory.ai/case-studies/comarch"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/factory-at-comarch-hundreds-of-engineers-now-use-the-factory-platform-to-run-age/</id>
    <title>Factory: At Comarch, hundreds of engineers now use the Factory platform to run agents end-to-end across...</title>
    <link rel="alternate" href="https://newruntime.com/signals/factory-at-comarch-hundreds-of-engineers-now-use-the-factory-platform-to-run-age/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>A public X post from Factory with a linked primary source flags At Comarch, hundreds of engineers now use the Factory platform to run agents end-to-end across the software development lifecycle. • 40% greater engineering efficiency • &gt;30% high...</summary>
    <category term="agents"/>
    <category term="coding-agents"/>
    <category term="interfaces"/>
    <link rel="related" href="https://x.com/FactoryAI/status/2082834849925046589"/>
    <link rel="related" href="https://factory.ai/case-studies/comarch"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/github-stacked-prs-now-on-github/</id>
    <title>GitHub: Stacked PRs now on GitHub 🥞</title>
    <link rel="alternate" href="https://newruntime.com/signals/github-stacked-prs-now-on-github/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>A public X post from GitHub with a linked primary source flags Stacked PRs now on GitHub 🥞</summary>
    <category term="coding-agents"/>
    <category term="github"/>
    <link rel="related" href="https://x.com/github/status/2082894271653306445"/>
    <link rel="related" href="https://docs.github.com/en/pull-requests/how-tos/stacked-pull-requests?utm_source=X-post-2&amp;utm_medium=social&amp;utm_campaign=stacked-prs-gtm-public-preview-2026"/>
    <link rel="related" href="https://docs.github.com/en/pull-requests/how-tos/stacked-pull-requests?utm_source=X-post-3&amp;utm_medium=social&amp;utm_campaign=stacked-prs-gtm-public-preview-2026"/>
    <link rel="related" href="https://docs.github.com/en/pull-requests/how-tos/stacked-pull-requests?utm_source=X-post-4&amp;utm_medium=social&amp;utm_campaign=stacked-prs-gtm-public-preview-2026"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/google-gemini-create-edit-and-summarize-with-gemini-on-macos-using-your-voice/</id>
    <title>Google Gemini: Create, edit, and summarize with Gemini on macOS using your voice.</title>
    <link rel="alternate" href="https://newruntime.com/signals/google-gemini-create-edit-and-summarize-with-gemini-on-macos-using-your-voice/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>A public X post from Google Gemini with a linked primary source flags Create, edit, and summarize with Gemini on macOS using your voice. Stay in your flow by speaking into any active window: Dictate clean text or ask Gemini to transform highlighted...</summary>
    <category term="ai"/>
    <category term="coding-agents"/>
    <category term="context-engineering"/>
    <category term="generative-ui"/>
    <link rel="related" href="https://x.com/GeminiApp/status/2082858829541458008"/>
    <link rel="related" href="https://blog.google/innovation-and-ai/products/gemini-app/speak-naturally-gemini-app-mac-os/"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/mem0-experiment-2-does-a-preference-survive-a-hard-context-wipe/</id>
    <title>mem0: Experiment 2: Does a preference survive a hard context wipe?</title>
    <link rel="alternate" href="https://newruntime.com/signals/mem0-experiment-2-does-a-preference-survive-a-hard-context-wipe/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>A public X post from mem0 with a linked primary source flags Experiment 2: Does a preference survive a hard context wipe? We told Claude(mid-conversation) to write functions with standard for loops instead of list comprehensions. Then ran /...</summary>
    <category term="agent-memory"/>
    <category term="coding-agents"/>
    <category term="context-engineering"/>
    <category term="interfaces"/>
    <category term="models"/>
    <link rel="related" href="https://x.com/mem0ai/status/2082850644927611257"/>
    <link rel="related" href="https://mem0.ai/blog/how-mem0-cut-claude-code-s-memory-footprint-by-97"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/minisqlite-makes-the-swarm-claim-inspectable/</id>
    <title>miniSQLite Makes the Coding-Swarm Claim Inspectable</title>
    <link rel="alternate" href="https://newruntime.com/posts/minisqlite-makes-the-swarm-claim-inspectable/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>The released Rust database turns Cursor&apos;s swarm experiment into an artifact that can be read, built, tested, and challenged instead of accepted as a benchmark chart.</summary>
    <category term="agents"/>
    <category term="coding-agents"/>
    <category term="developer-tools"/>
    <link rel="related" href="https://github.com/cursor/minisqlite"/>
    <link rel="related" href="https://cursor.com/blog/agent-swarm-model-economics"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/openai-developers-imagegen-in-codex-just-got-a-new-lightbox-and-canvas/</id>
    <title>OpenAI Developers: ImageGen in Codex just got a new lightbox and canvas.</title>
    <link rel="alternate" href="https://newruntime.com/signals/openai-developers-imagegen-in-codex-just-got-a-new-lightbox-and-canvas/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>A public X post from OpenAI Developers with a linked primary source flags ImageGen in Codex just got a new lightbox and canvas. Now, it’s even easier to explore and refine visuals in your workflow.</summary>
    <category term="coding-agents"/>
    <category term="models"/>
    <category term="workflows"/>
    <link rel="related" href="https://x.com/OpenAIDevs/status/2082944138635595782"/>
    <link rel="related" href="https://learn.chatgpt.com/docs/changelog"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/openai-we-re-also-upgrading-auto-review-in-the-chatgpt-app-and-codex-cli-from-gp/</id>
    <title>OpenAI: We’re also upgrading Auto-review in the ChatGPT app and Codex CLI from GPT-5.4 to GPT-5.6 Luna.</title>
    <link rel="alternate" href="https://newruntime.com/signals/openai-we-re-also-upgrading-auto-review-in-the-chatgpt-app-and-codex-cli-from-gp/"/>
    <published>2026-07-30T00:00:00.000Z</published>
    <updated>2026-07-30T00:00:00.000Z</updated>
    <summary>A public X post from OpenAI with a linked primary source flags We’re also upgrading Auto-review in the ChatGPT app and Codex CLI from GPT-5.4 to GPT-5.6 Luna. Combined with Luna’s new price, we expect Auto-review to cost about 10x less, makin...</summary>
    <category term="agents"/>
    <category term="api-design"/>
    <category term="coding-agents"/>
    <category term="context-engineering"/>
    <category term="models"/>
    <category term="workflows"/>
    <link rel="related" href="https://x.com/OpenAI/status/2082878180478910571"/>
    <link rel="related" href="https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/codex-multi-agent-orchestration-skill/</id>
    <title>A Codex Skill Turns Multi-Agent Work Into a Reusable Control Surface</title>
    <link rel="alternate" href="https://newruntime.com/posts/codex-multi-agent-orchestration-skill/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A shared Codex orchestration skill points to a future where agent workflows are packaged as portable operational routines.</summary>
    <category term="ai-infrastructure"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://x.com/pvncher/status/2080707291603407077"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/codex-multi-agent-orchestration-skill/</id>
    <title>A Codex Skill Turns Multi-Agent Work Into a Reusable Control Surface</title>
    <link rel="alternate" href="https://newruntime.com/signals/codex-multi-agent-orchestration-skill/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A shared Codex orchestration skill points to a future where agent workflows are packaged as portable operational routines.</summary>
    <category term="ai-infrastructure"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://x.com/pvncher/status/2080707291603407077"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/agentic-engineering-full-stack-workflow/</id>
    <title>Agentic Engineering Looks Like Workflow Design, Not Hands-Free Coding</title>
    <link rel="alternate" href="https://newruntime.com/posts/agentic-engineering-full-stack-workflow/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A full-stack agentic engineering walkthrough shows the operating pattern around planning, validation, browser checks, and human review.</summary>
    <category term="coding-agents"/>
    <category term="human-computer-interaction"/>
    <link rel="related" href="https://www.youtube.com/watch?v=kPN564Kol14"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/agentic-engineering-full-stack-workflow/</id>
    <title>Agentic Engineering Looks Like Workflow Design, Not Hands-Free Coding</title>
    <link rel="alternate" href="https://newruntime.com/signals/agentic-engineering-full-stack-workflow/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A full-stack agentic engineering walkthrough shows the operating pattern around planning, validation, browser checks, and human review.</summary>
    <category term="coding-agents"/>
    <category term="human-computer-interaction"/>
    <link rel="related" href="https://www.youtube.com/watch?v=kPN564Kol14"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/augment-code-the-best-value-for-every-token-gpt-5-6-sol-is-now-our-default-in-co/</id>
    <title>Augment Code: The best value for every token: GPT-5.6 Sol is now our default in Cosmos Eight models in eight...</title>
    <link rel="alternate" href="https://newruntime.com/signals/augment-code-the-best-value-for-every-token-gpt-5-6-sol-is-now-our-default-in-co/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A public X post from Augment Code as a public source in its own right flags The best value for every token: GPT-5.6 Sol is now our default in Cosmos Eight models in eight weeks! After testing them all on real, long-horizon software development tasks, we’r...</summary>
    <category term="coding-agents"/>
    <category term="models"/>
    <link rel="related" href="https://x.com/augmentcode/status/2082594768987574480"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/claude-code-certification-credential-layer/</id>
    <title>Claude Code Certifications Turn Agent Use Into a Credential Layer</title>
    <link rel="alternate" href="https://newruntime.com/posts/claude-code-certification-credential-layer/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>Anthropic certification activity around Claude Code is a signal that agentic development is becoming a managed enterprise capability.</summary>
    <category term="ai-business-models"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://www.linkedin.com/posts/introducing-three-new-certifications-for-share-7486412944876998657-hIBP"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/claude-code-certification-credential-layer/</id>
    <title>Claude Code Certifications Turn Agent Use Into a Credential Layer</title>
    <link rel="alternate" href="https://newruntime.com/signals/claude-code-certification-credential-layer/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>Anthropic certification activity around Claude Code is a signal that agentic development is becoming a managed enterprise capability.</summary>
    <category term="ai-business-models"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://www.linkedin.com/posts/introducing-three-new-certifications-for-share-7486412944876998657-hIBP"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/claude-md-prompt-cruft-shrinkage/</id>
    <title>Claude.md Shrinkage Says the Harness Is Learning What Not to Say</title>
    <link rel="alternate" href="https://newruntime.com/posts/claude-md-prompt-cruft-shrinkage/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A reported 80 percent reduction in Claude Code prompt material points to a maturing harness discipline around concise system context.</summary>
    <category term="ai-infrastructure"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://x.com/trq212/status/2080710971228918066"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/claude-md-prompt-cruft-shrinkage/</id>
    <title>Claude.md Shrinkage Says the Harness Is Learning What Not to Say</title>
    <link rel="alternate" href="https://newruntime.com/signals/claude-md-prompt-cruft-shrinkage/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A reported 80 percent reduction in Claude Code prompt material points to a maturing harness discipline around concise system context.</summary>
    <category term="ai-infrastructure"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://x.com/trq212/status/2080710971228918066"/>
  </entry>
  <entry>
    <id>https://newruntime.com/shifts/verification-ownership/</id>
    <title>Code production -&gt; verification ownership</title>
    <link rel="alternate" href="https://newruntime.com/shifts/verification-ownership/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>As agents produce more implementation, engineering responsibility is moving toward specifications, evidence, acceptance decisions, and release accountability.</summary>
    <category term="code-review"/>
    <category term="coding-agents"/>
    <category term="engineering-management"/>
    <category term="verification"/>
    <link rel="related" href="https://ir.gitlab.com/news/news-details/2026/GitLab-Research-Reveals-Organizations-Are-Generating-AI-Code-Faster-Than-They-Can-Control-It/default.aspx"/>
    <link rel="related" href="https://www.databricks.com/blog/benchmarking-coding-agents-databricks-multi-million-line-codebase"/>
    <link rel="related" href="https://www.anthropic.com/research/AI-assistance-coding-skills"/>
    <link rel="related" href="https://learn.chatgpt.com/docs/hooks"/>
    <link rel="related" href="https://www.imperialviolet.org/2026/07/26/zstd-lean.html"/>
    <link rel="related" href="https://coles.codes/posts/reviewing-code-you-didnt-write"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/code-review-becomes-agent-bottleneck/</id>
    <title>Code Review Becomes the Agent Bottleneck</title>
    <link rel="alternate" href="https://newruntime.com/posts/code-review-becomes-agent-bottleneck/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>Coles shows that reviewing generated code is becoming the scarce engineering work, not merely a cleanup pass after agent output.</summary>
    <category term="coding-agents"/>
    <category term="human-computer-interaction"/>
    <link rel="related" href="https://coles.codes/posts/reviewing-code-you-didnt-write"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/code-review-becomes-agent-bottleneck/</id>
    <title>Code Review Becomes the Agent Bottleneck</title>
    <link rel="alternate" href="https://newruntime.com/signals/code-review-becomes-agent-bottleneck/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>Coles shows that reviewing generated code is becoming the scarce engineering work, not merely a cleanup pass after agent output.</summary>
    <category term="coding-agents"/>
    <category term="human-computer-interaction"/>
    <link rel="related" href="https://coles.codes/posts/reviewing-code-you-didnt-write"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/codex-hooks-close-the-type-error-loop/</id>
    <title>Codex Hooks Close the Type Error Loop</title>
    <link rel="alternate" href="https://newruntime.com/posts/codex-hooks-close-the-type-error-loop/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>Codex hooks make validation output an active part of the agent loop, so type errors and checks can be returned before the task drifts.</summary>
    <category term="ai-infrastructure"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://learn.chatgpt.com/docs/hooks"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/codex-hooks-close-the-type-error-loop/</id>
    <title>Codex Hooks Close the Type Error Loop</title>
    <link rel="alternate" href="https://newruntime.com/signals/codex-hooks-close-the-type-error-loop/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>Codex hooks make validation output an active part of the agent loop, so type errors and checks can be returned before the task drifts.</summary>
    <category term="ai-infrastructure"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://learn.chatgpt.com/docs/hooks"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/codex-security-cli-scan-workbench/</id>
    <title>Codex Security CLI Turns Security Review Into a Scannable Workbench</title>
    <link rel="alternate" href="https://newruntime.com/posts/codex-security-cli-scan-workbench/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>OpenAI&apos;s Codex Security CLI packages repository, diff, working-tree, export, validation, and patch flows into a security-review workbench rather than a single scanner command.</summary>
    <category term="codex"/>
    <category term="coding-agents"/>
    <category term="developer-tools"/>
    <category term="security"/>
    <link rel="related" href="https://x.com/OpenAI/status/2082263717916586117"/>
    <link rel="related" href="https://www.npmjs.com/package/@openai/codex-security"/>
    <link rel="related" href="https://github.com/openai/codex-security"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/cursor-cursor-is-now-on-ipad/</id>
    <title>Cursor: Cursor is now on iPad.</title>
    <link rel="alternate" href="https://newruntime.com/signals/cursor-cursor-is-now-on-ipad/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A public X post from Cursor with a linked primary source flags Cursor is now on iPad. All the power of Cursor on iPhone, with more room to work with agents.</summary>
    <category term="agents"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://x.com/cursor_ai/status/2082532273421955513"/>
    <link rel="related" href="https://apps.apple.com/us/app/cursor/id6767085653"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/openai-developers-we-used-gpt-5-6-sol-in-codex-to-optimize-its-own-infrastructur/</id>
    <title>OpenAI Developers: We used GPT-5.6 Sol in Codex to optimize its own infrastructure and performance.</title>
    <link rel="alternate" href="https://newruntime.com/signals/openai-developers-we-used-gpt-5-6-sol-in-codex-to-optimize-its-own-infrastructur/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A public X post from OpenAI Developers as a public source in its own right flags We used GPT-5.6 Sol in Codex to optimize its own infrastructure and performance. These improvements compound across inference and the agent loop, producing more useful work from...</summary>
    <category term="agents"/>
    <category term="coding-agents"/>
    <category term="models"/>
    <link rel="related" href="https://x.com/OpenAIDevs/status/2082580211552457102"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/openai-we-quietly-released-the-open-source-codex-security-cli-but-hacker-news-fo/</id>
    <title>OpenAI: We quietly released the open-source Codex Security CLI, but Hacker News found it before we had...</title>
    <link rel="alternate" href="https://newruntime.com/signals/openai-we-quietly-released-the-open-source-codex-security-cli-but-hacker-news-fo/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>A public X post from OpenAI with a linked primary source flags We quietly released the open-source Codex Security CLI, but Hacker News found it before we had a chance to share it here... You can now use it to scan repositories, track findings...</summary>
    <category term="coding-agents"/>
    <category term="interfaces"/>
    <category term="models"/>
    <category term="security"/>
    <link rel="related" href="https://x.com/OpenAI/status/2082263717916586117"/>
    <link rel="related" href="https://www.npmjs.com/package/@openai/codex-security"/>
    <link rel="related" href="https://github.com/openai/codex-security"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/prompt-caching-turns-context-into-infrastructure/</id>
    <title>Prompt Caching Turns Agent Context Into Infrastructure</title>
    <link rel="alternate" href="https://newruntime.com/posts/prompt-caching-turns-context-into-infrastructure/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>Earendil frames prompt caching as an agent systems primitive, where stable context becomes a cost, latency, and architecture concern.</summary>
    <category term="ai-infrastructure"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://earendil.com/posts/prompt-caching/"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/prompt-caching-turns-context-into-infrastructure/</id>
    <title>Prompt Caching Turns Agent Context Into Infrastructure</title>
    <link rel="alternate" href="https://newruntime.com/signals/prompt-caching-turns-context-into-infrastructure/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>Earendil frames prompt caching as an agent systems primitive, where stable context becomes a cost, latency, and architecture concern.</summary>
    <category term="ai-infrastructure"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://earendil.com/posts/prompt-caching/"/>
  </entry>
  <entry>
    <id>https://newruntime.com/posts/raptor-loop-hunt-security-agent-loop/</id>
    <title>RAPTOR Loop Hunt Packages Security Hunting as an Agent Skill</title>
    <link rel="alternate" href="https://newruntime.com/posts/raptor-loop-hunt-security-agent-loop/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>RAPTOR Loop Hunt shows security research moving toward looped agent skills with altitude changes, evidence collection, and review checkpoints.</summary>
    <category term="ai-security"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://github.com/dinosn/raptor-loop-hunt"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/raptor-loop-hunt-security-agent-loop/</id>
    <title>RAPTOR Loop Hunt Packages Security Hunting as an Agent Skill</title>
    <link rel="alternate" href="https://newruntime.com/signals/raptor-loop-hunt-security-agent-loop/"/>
    <published>2026-07-29T00:00:00.000Z</published>
    <updated>2026-07-29T00:00:00.000Z</updated>
    <summary>RAPTOR Loop Hunt shows security research moving toward looped agent skills with altitude changes, evidence collection, and review checkpoints.</summary>
    <category term="ai-security"/>
    <category term="coding-agents"/>
    <link rel="related" href="https://github.com/dinosn/raptor-loop-hunt"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/augment-code-our-users-have-been-really-excited-to-try-kimi-k3-and-we-re-excited/</id>
    <title>Augment Code: Our users have been really excited to try Kimi K3 and we&apos;re excited to bring it to Augment&apos;s p...</title>
    <link rel="alternate" href="https://newruntime.com/signals/augment-code-our-users-have-been-really-excited-to-try-kimi-k3-and-we-re-excited/"/>
    <published>2026-07-28T00:00:00.000Z</published>
    <updated>2026-07-28T00:00:00.000Z</updated>
    <summary>A public X post from Augment Code as a public source in its own right flags Our users have been really excited to try Kimi K3 and we&apos;re excited to bring it to Augment&apos;s product family. It is the most capable open-source model that we have tested to date....</summary>
    <category term="agents"/>
    <category term="coding-agents"/>
    <category term="models"/>
    <link rel="related" href="https://x.com/augmentcode/status/2081926971249054019"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/cursor-today-we-re-launching-cursor-start-a-new-649-month-plan-for-developers-in/</id>
    <title>Cursor: Today we&apos;re launching Cursor Start, a new ₹649/month plan for developers in India.</title>
    <link rel="alternate" href="https://newruntime.com/signals/cursor-today-we-re-launching-cursor-start-a-new-649-month-plan-for-developers-in/"/>
    <published>2026-07-28T00:00:00.000Z</published>
    <updated>2026-07-28T00:00:00.000Z</updated>
    <summary>A public X post from Cursor with a linked primary source flags Today we&apos;re launching Cursor Start, a new ₹649/month plan for developers in India. Start includes generous access to Grok 4.5 and Composer, so you can plan, build, test, and ship...</summary>
    <category term="agents"/>
    <category term="coding-agents"/>
    <category term="interfaces"/>
    <category term="mcp"/>
    <category term="workflows"/>
    <link rel="related" href="https://x.com/cursor_ai/status/2081978255004053560"/>
    <link rel="related" href="https://cursor.com/blog/cursor-start-india"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/factory-we-re-firm-believers-in-the-open-security-ecosystem-contributing-across/</id>
    <title>Factory: We&apos;re firm believers in the open security ecosystem, contributing across our platform: • Publi...</title>
    <link rel="alternate" href="https://newruntime.com/signals/factory-we-re-firm-believers-in-the-open-security-ecosystem-contributing-across/"/>
    <published>2026-07-28T00:00:00.000Z</published>
    <updated>2026-07-28T00:00:00.000Z</updated>
    <summary>A public X post from Factory with a linked primary source flags We&apos;re firm believers in the open security ecosystem, contributing across our platform: • Publishing two open-weight models behind Droid Shield 2.0 • Using Autonomous Security Revi...</summary>
    <category term="coding-agents"/>
    <category term="interfaces"/>
    <category term="models"/>
    <category term="security"/>
    <link rel="related" href="https://x.com/FactoryAI/status/2082138137065669106"/>
    <link rel="related" href="https://blogs.nvidia.com/blog/open-secure-ai-alliance/"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/github-5/</id>
    <title>GitHub: 5.</title>
    <link rel="alternate" href="https://newruntime.com/signals/github-5/"/>
    <published>2026-07-28T00:00:00.000Z</published>
    <updated>2026-07-28T00:00:00.000Z</updated>
    <summary>A public X post from GitHub with a linked primary source flags 5. Turn on code scanning Code scanning uses CodeQL to identify patterns that lead to vulnerabilities, including injection flaws, unsafe deserialization, and insecure GitHub Action...</summary>
    <category term="coding-agents"/>
    <category term="github"/>
    <category term="interfaces"/>
    <category term="security"/>
    <category term="workflows"/>
    <link rel="related" href="https://x.com/github/status/2082156050279739854"/>
    <link rel="related" href="https://github.blog/security/6-security-settings-every-github-maintainer-should-enable-this-week/"/>
  </entry>
  <entry>
    <id>https://newruntime.com/signals/github-you-re-not-behind/</id>
    <title>GitHub: You&apos;re not behind.</title>
    <link rel="alternate" href="https://newruntime.com/signals/github-you-re-not-behind/"/>
    <published>2026-07-28T00:00:00.000Z</published>
    <updated>2026-07-28T00:00:00.000Z</updated>
    <summary>A public X post from GitHub as a public source in its own right flags You&apos;re not behind. There&apos;s no secret everyone else has. There&apos;s just the harness, and it&apos;s mostly all you need. @burkeholland gives you a simple, repeatable workflow for GitHub Co...</summary>
    <category term="coding-agents"/>
    <category term="github"/>
    <category term="workflows"/>
    <link rel="related" href="https://x.com/github/status/2082201573976056245"/>
  </entry>
</feed>
