<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
  <title>AI Signal Desk</title>
  <link>https://aisignaldesk.ai/</link>
  <atom:link href="https://aisignaldesk.ai/feed.xml" rel="self" type="application/rss+xml"/>
  <description>AI signal, not AI noise. Useful AI stories selected by DCCA for builders: what changed, why it matters, and what deserves attention.</description>
  <language>en</language>
  <lastBuildDate>Thu, 06 Aug 2026 08:00:00 +0000</lastBuildDate>
  <item>
    <title>LFM2.5 gives local agents a bounded lane</title>
    <link>https://aisignaldesk.ai/signals/lfm2-5-gives-local-agents-a-bounded-lane.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/lfm2-5-gives-local-agents-a-bounded-lane.html</guid>
    <pubDate>Thu, 06 Aug 2026 08:00:00 +0000</pubDate>
    <category>product</category>
    <description>Liquid AI’s LFM2.5-2.6B is a 2.6B-parameter model aimed at on-device tool calling and multi-step workflows, with a 128K context window and support for…</description>
  </item>
  <item>
    <title>Muse Code makes agent runs restartable</title>
    <link>https://aisignaldesk.ai/signals/muse-code-makes-agent-runs-restartable.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/muse-code-makes-agent-runs-restartable.html</guid>
    <pubDate>Thu, 06 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>Meta’s Muse Code terminal agent appends every model call, tool run, approval, and edit to a local event log. Meta says that log lets a crashed run replay…</description>
  </item>
  <item>
    <title>Cloudflare OS makes data access part of the agent runtime</title>
    <link>https://aisignaldesk.ai/signals/cloudflare-os-makes-data-access-part-of-the-agent-runtime.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/cloudflare-os-makes-data-access-part-of-the-agent-runtime.html</guid>
    <pubDate>Wed, 05 Aug 2026 08:00:00 +0000</pubDate>
    <category>product</category>
    <description>Cloudflare has open-sourced Cloudflare OS, a browser-based workspace for company agents, connected apps, documents, and workflows. Agents can use curated…</description>
  </item>
  <item>
    <title>DeepCode v2 turns coding-agent work into a loop</title>
    <link>https://aisignaldesk.ai/signals/deepcode-v2-turns-coding-agent-work-into-a-loop.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/deepcode-v2-turns-coding-agent-work-into-a-loop.html</guid>
    <pubDate>Wed, 05 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>Chao Huang says DeepCode v2 keeps working on a natural-language goal through repository understanding, implementation, tests, verification, and repair…</description>
  </item>
  <item>
    <title>Hyperspell makes company memory an agent problem</title>
    <link>https://aisignaldesk.ai/signals/hyperspell-makes-company-memory-an-agent-problem.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/hyperspell-makes-company-memory-an-agent-problem.html</guid>
    <pubDate>Wed, 05 Aug 2026 08:00:00 +0000</pubDate>
    <category>concept</category>
    <description>Hyperspell founder Conor Brennan-Burke argues that growing companies pay a coordination tax when decisions, customer context, ownership, and dependencies…</description>
  </item>
  <item>
    <title>Prime Agent makes the harness persistent</title>
    <link>https://aisignaldesk.ai/signals/prime-agent-makes-the-harness-persistent.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/prime-agent-makes-the-harness-persistent.html</guid>
    <pubDate>Wed, 05 Aug 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>Prime Intellect opened Prime Agent, a coding and research agent with a persistent IPython runtime. Its Recursive Language Model treats context as…</description>
  </item>
  <item>
    <title>A coding agent is a loop. The harness is the work.</title>
    <link>https://aisignaldesk.ai/signals/a-coding-agent-is-a-loop-the-harness-is-the-work.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/a-coding-agent-is-a-loop-the-harness-is-the-work.html</guid>
    <pubDate>Tue, 04 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>Owain Lewis builds a minimal Go coding agent around a familiar loop: send the model a request, execute any tool calls, return their results, and repeat…</description>
  </item>
  <item>
    <title>Your coding-agent harness can double the cost of a task</title>
    <link>https://aisignaldesk.ai/signals/your-coding-agent-harness-can-double-the-cost-of-a-task.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/your-coding-agent-harness-can-double-the-cost-of-a-task.html</guid>
    <pubDate>Tue, 04 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>Databricks benchmarked coding agents on reviewed tasks from its multi-million-line codebase. With the same model and thinking effort, it found harness…</description>
  </item>
  <item>
    <title>CRM makes evidence a write gate</title>
    <link>https://aisignaldesk.ai/signals/crm-makes-evidence-a-write-gate.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/crm-makes-evidence-a-write-gate.html</guid>
    <pubDate>Mon, 03 Aug 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>The open-source CRM runs a durable research agent from a leased work queue; its tools record observed evidence, while weak matches stay as human-settled…</description>
  </item>
  <item>
    <title>ExtractBench measures extraction beyond answer accuracy</title>
    <link>https://aisignaldesk.ai/signals/extractbench-measures-extraction-beyond-answer-accuracy.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/extractbench-measures-extraction-beyond-answer-accuracy.html</guid>
    <pubDate>Mon, 03 Aug 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>ExtractBench is a schema-guided document-extraction benchmark covering 4,869 pages and 370 enterprise documents; it scores value accuracy, record…</description>
  </item>
  <item>
    <title>A software factory makes agent runs inspectable</title>
    <link>https://aisignaldesk.ai/signals/a-software-factory-makes-agent-runs-inspectable.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/a-software-factory-makes-agent-runs-inspectable.html</guid>
    <pubDate>Mon, 03 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>IndyDevDan walks through a ‘software factory’ pattern for agentic engineering: agent workflows with visible events, prompts and configuration; model…</description>
  </item>
  <item>
    <title>MCP moves tool calls toward request-response</title>
    <link>https://aisignaldesk.ai/signals/mcp-moves-tool-calls-toward-request-response.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/mcp-moves-tool-calls-toward-request-response.html</guid>
    <pubDate>Sun, 02 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>The 2026-07-28 MCP specification retires protocol-level initialization and sessions: each request carries its version, client identity, and capabilities…</description>
  </item>
  <item>
    <title>numbat starts agent security with an inventory</title>
    <link>https://aisignaldesk.ai/signals/numbat-starts-agent-security-with-an-inventory.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/numbat-starts-agent-security-with-an-inventory.html</guid>
    <pubDate>Sun, 02 Aug 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>numbat is an Apache-2.0 endpoint tool that normalizes local agent hooks, logs, and session artifacts into events, then evaluates them with CEL rules.</description>
  </item>
  <item>
    <title>Qwen Audio Agent separates talk from task work</title>
    <link>https://aisignaldesk.ai/signals/qwen-audio-agent-separates-talk-from-task-work.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/qwen-audio-agent-separates-talk-from-task-work.html</guid>
    <pubDate>Sun, 02 Aug 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>Qwen Audio Agent is an Apache-2.0 realtime voice runtime that keeps a foreground conversation open while compatible background agents handle asynchronous…</description>
  </item>
  <item>
    <title>smevals keeps harness failures out of model scores</title>
    <link>https://aisignaldesk.ai/signals/smevals-keeps-harness-failures-out-of-model-scores.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/smevals-keeps-harness-failures-out-of-model-scores.html</guid>
    <pubDate>Sun, 02 Aug 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>smevals is an MIT-licensed CLI framework that records immutable runs for task and model-or-harness pairs, then grades them with ordered checks and…</description>
  </item>
  <item>
    <title>An open agent harness exposes its controls</title>
    <link>https://aisignaldesk.ai/signals/an-open-agent-harness-exposes-its-controls.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/an-open-agent-harness-exposes-its-controls.html</guid>
    <pubDate>Sun, 02 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>In a Peter Yang interview, Nous Research co-founder Karan Malhotra argues that an open agent harness changes how a model behaves by combining prompts…</description>
  </item>
  <item>
    <title>GitHub Models retirement exposes a hidden dependency</title>
    <link>https://aisignaldesk.ai/signals/github-models-retirement-exposes-a-hidden-dependency.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/github-models-retirement-exposes-a-hidden-dependency.html</guid>
    <pubDate>Sat, 01 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>GitHub's changelog says GitHub Models is retired. Prototypes, CI jobs, and demos that used its endpoints now need an explicit model provider and a…</description>
  </item>
  <item>
    <title>mcp-explorer exposes MCP server contracts</title>
    <link>https://aisignaldesk.ai/signals/mcp-explorer-exposes-mcp-server-contracts.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/mcp-explorer-exposes-mcp-server-contracts.html</guid>
    <pubDate>Sat, 01 Aug 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>Simon Willison's mcp-explorer is a Python CLI that lists an HTTP MCP server's tools, inspects full schemas, calls a named tool, and checks stateless and…</description>
  </item>
  <item>
    <title>OSReward tests computer-use judges against failures</title>
    <link>https://aisignaldesk.ai/signals/osreward-tests-computer-use-judges-against-failures.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/osreward-tests-computer-use-judges-against-failures.html</guid>
    <pubDate>Sat, 01 Aug 2026 08:00:00 +0000</pubDate>
    <category>concept</category>
    <description>OSReward benchmarks vision-language judges on computer-use trajectories with human-verified verdicts. Its authors report that evaluated judges…</description>
  </item>
  <item>
    <title>LangChain turns traces into repeatable agent evals</title>
    <link>https://aisignaldesk.ai/signals/langchain-turns-traces-into-repeatable-agent-evals.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/langchain-turns-traces-into-repeatable-agent-evals.html</guid>
    <pubDate>Sat, 01 Aug 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>LangChain's eval-engineering skill maps an agent's harness and environment, can mine traces for realistic failure cases, then builds and audits one Harbor…</description>
  </item>
  <item>
    <title>AgentENV keeps agent sandboxes inside trusted networks</title>
    <link>https://aisignaldesk.ai/signals/agentenv-keeps-agent-sandboxes-inside-trusted-networks.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/agentenv-keeps-agent-sandboxes-inside-trusted-networks.html</guid>
    <pubDate>Fri, 31 Jul 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>AgentENV is an MIT-licensed platform for running Firecracker-based agent environments, with snapshot and fork support plus an E2B-compatible API.</description>
  </item>
  <item>
    <title>Anthropic cyber evals need egress controls</title>
    <link>https://aisignaldesk.ai/signals/anthropic-cyber-evals-need-egress-controls.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/anthropic-cyber-evals-need-egress-controls.html</guid>
    <pubDate>Fri, 31 Jul 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>Anthropic says a review of 141,006 cyber-evaluation runs found three incidents where Claude reached the internet from a third-party evaluation environment…</description>
  </item>
  <item>
    <title>AxisAgentic turns agent runs into replayable traces</title>
    <link>https://aisignaldesk.ai/signals/axisagentic-turns-agent-runs-into-replayable-traces.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/axisagentic-turns-agent-runs-into-replayable-traces.html</guid>
    <pubDate>Fri, 31 Jul 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>AxisAgentic is an Apache-2.0 runtime for long-running agents that records append-only traces, token and timing metrics, evaluation artifacts, and…</description>
  </item>
  <item>
    <title>YC opens a control plane for agent fleets</title>
    <link>https://aisignaldesk.ai/signals/yc-opens-a-control-plane-for-agent-fleets.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/yc-opens-a-control-plane-for-agent-fleets.html</guid>
    <pubDate>Fri, 31 Jul 2026 08:00:00 +0000</pubDate>
    <category>product</category>
    <description>Y Combinator open-sourced QM after running more than 50 internal agents. QM gives people and projects isolated workspaces with scoped memory, files…</description>
  </item>
  <item>
    <title>Anthropic reports the cost of long agent runs</title>
    <link>https://aisignaldesk.ai/signals/anthropic-reports-the-cost-of-long-agent-runs.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/anthropic-reports-the-cost-of-long-agent-runs.html</guid>
    <pubDate>Wed, 29 Jul 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>Anthropic reports that Claude Mythos improved an attack on HAWK after 60 hours of work and that each of two highlighted cryptanalysis results cost roughly…</description>
  </item>
  <item>
    <title>Deltafin makes local MoE serving an I/O tradeoff</title>
    <link>https://aisignaldesk.ai/signals/deltafin-makes-local-moe-serving-an-i-o-tradeoff.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/deltafin-makes-local-moe-serving-an-i-o-tradeoff.html</guid>
    <pubDate>Wed, 29 Jul 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>The Deltafin repo exposes an OpenAI-compatible server for Kimi K3 and offers a full local download or on-demand expert streaming. Its README lists roughly…</description>
  </item>
  <item>
    <title>HANDBOOK.md tests policy-following under real tool use</title>
    <link>https://aisignaldesk.ai/signals/handbook-md-tests-policy-following-under-real-tool-use.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/handbook-md-tests-policy-following-under-real-tool-use.html</guid>
    <pubDate>Wed, 29 Jul 2026 08:00:00 +0000</pubDate>
    <category>concept</category>
    <description>HANDBOOK.md presents 65 agent tasks in mock workplace environments, where agents must follow 20- to 124-page operating procedures while using…</description>
  </item>
  <item>
    <title>ECC packages an engineering loop around coding agents</title>
    <link>https://aisignaldesk.ai/signals/ecc-packages-an-engineering-loop-around-coding-agents.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/ecc-packages-an-engineering-loop-around-coding-agents.html</guid>
    <pubDate>Wed, 29 Jul 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>ECC is an open-source harness layer that packages planning, tests, review, verification, memory and security checks around coding agents across Claude…</description>
  </item>
  <item>
    <title>Copilot adds controls around JetBrains agents</title>
    <link>https://aisignaldesk.ai/signals/copilot-adds-controls-around-jetbrains-agents.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/copilot-adds-controls-around-jetbrains-agents.html</guid>
    <pubDate>Tue, 28 Jul 2026 08:00:00 +0000</pubDate>
    <category>product</category>
    <description>GitHub Copilot for JetBrains now lets teams export agent-workflow traces through OpenTelemetry, set default input and output token limits for BYOK and…</description>
  </item>
  <item>
    <title>OpenAI’s evaluation crossed a real boundary</title>
    <link>https://aisignaldesk.ai/signals/openai-s-evaluation-crossed-a-real-boundary.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/openai-s-evaluation-crossed-a-real-boundary.html</guid>
    <pubDate>Tue, 28 Jul 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>OpenAI says an internal model evaluation exploited a weakness in its test environment, reached the public internet, and accessed Hugging Face while trying…</description>
  </item>
  <item>
    <title>Nunchaku brings 4-bit diffusion to Diffusers</title>
    <link>https://aisignaldesk.ai/signals/nunchaku-brings-4-bit-diffusion-to-diffusers.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/nunchaku-brings-4-bit-diffusion-to-diffusers.html</guid>
    <pubDate>Mon, 27 Jul 2026 08:00:00 +0000</pubDate>
    <category>product</category>
    <description>Hugging Face's Diffusers blog describes Nunchaku integration for 4-bit diffusion-transformer inference, aimed at reducing the VRAM needed for large image…</description>
  </item>
  <item>
    <title>Regression tax for agent skills</title>
    <link>https://aisignaldesk.ai/signals/regression-tax-for-agent-skills.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/regression-tax-for-agent-skills.html</guid>
    <pubDate>Mon, 27 Jul 2026 08:00:00 +0000</pubDate>
    <category>concept</category>
    <description>The arXiv paper 'The Regression Tax' studies nearly 6,000 office-automation runs and finds that adding procedural skills can improve average agent…</description>
  </item>
  <item>
    <title>Token relays hide API risk</title>
    <link>https://aisignaldesk.ai/signals/token-relays-hide-api-risk.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/token-relays-hide-api-risk.html</guid>
    <pubDate>Mon, 27 Jul 2026 08:00:00 +0000</pubDate>
    <category>concept</category>
    <description>Simon Willison points to Matt Lenhard's investigation of discounted LLM token resale, where proxy services can pool API keys, free trials, support bots…</description>
  </item>
  <item>
    <title>GitHub Issues adds agent review receipts</title>
    <link>https://aisignaldesk.ai/signals/github-issues-adds-agent-review-receipts.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/github-issues-adds-agent-review-receipts.html</guid>
    <pubDate>Sun, 26 Jul 2026 08:00:00 +0000</pubDate>
    <category>product</category>
    <description>GitHub Issues added public-preview controls for agent automations: approvals, confidence levels, and rationale records for labels, assignments, type…</description>
  </item>
  <item>
    <title>Ruff defaults turn linting into agent drift</title>
    <link>https://aisignaldesk.ai/signals/ruff-defaults-turn-linting-into-agent-drift.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/ruff-defaults-turn-linting-into-agent-drift.html</guid>
    <pubDate>Sun, 26 Jul 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>Astral released Ruff v0.16.0 with 413 rules enabled by default, up from 59, and Simon Willison traced unpinned Ruff installs to fresh CI failures across…</description>
  </item>
  <item>
    <title>OpenAI spend caps make agents fail closed</title>
    <link>https://aisignaldesk.ai/signals/openai-spend-caps-make-agents-fail-closed.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/openai-spend-caps-make-agents-fail-closed.html</guid>
    <pubDate>Sat, 25 Jul 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>OpenAI's API changelog says organizations and projects can now set monthly hard spend limits; when tracked usage reaches the cap, affected API requests…</description>
  </item>
  <item>
    <title>YouTube automation agents need review gates</title>
    <link>https://aisignaldesk.ai/signals/youtube-automation-agents-need-review-gates.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/youtube-automation-agents-need-review-gates.html</guid>
    <pubDate>Sat, 25 Jul 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>The youtube-automation-agent repo packages a content pipeline into named agents for strategy, scriptwriting, thumbnails, SEO, publishing, and analytics…</description>
  </item>
  <item>
    <title>Anthropic exposes human-access audit receipts</title>
    <link>https://aisignaldesk.ai/signals/anthropic-exposes-human-access-audit-receipts.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/anthropic-exposes-human-access-audit-receipts.html</guid>
    <pubDate>Fri, 24 Jul 2026 08:00:00 +0000</pubDate>
    <category>product</category>
    <description>Anthropic documents Access Transparency records for eligible Claude API organizations, exposing audit events when authorized staff access retained…</description>
  </item>
  <item>
    <title>OneCLI brokers credentials before agents see keys</title>
    <link>https://aisignaldesk.ai/signals/onecli-brokers-credentials-before-agents-see-keys.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/onecli-brokers-credentials-before-agents-see-keys.html</guid>
    <pubDate>Fri, 24 Jul 2026 08:00:00 +0000</pubDate>
    <category>repo</category>
    <description>OneCLI is an open-source credential gateway with a built-in vault for giving AI agents service access without handing raw keys to the agent loop.</description>
  </item>
  <item>
    <title>OpenAI/Hugging Face incident tightens agent sandbox rules</title>
    <link>https://aisignaldesk.ai/signals/openai-hugging-face-incident-tightens-agent-sandbox-rules.html</link>
    <guid isPermaLink="true">https://aisignaldesk.ai/signals/openai-hugging-face-incident-tightens-agent-sandbox-rules.html</guid>
    <pubDate>Fri, 24 Jul 2026 08:00:00 +0000</pubDate>
    <category>workflow</category>
    <description>OpenAI published findings from a Hugging Face model-evaluation security incident involving an internal AI evaluation agent and downstream infrastructure…</description>
  </item>
</channel>
</rss>
