<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>Automation on Cognaptus</title>
    <link>https://cognaptus.com/tags/automation/</link>
    <description>Recent content in Automation on Cognaptus</description>
    <generator>Hugo -- 0.145.0</generator>
    <language>en-us</language>
    <lastBuildDate>Sun, 30 Aug 2026 00:00:00 +0000</lastBuildDate>
    <atom:link href="https://cognaptus.com/tags/automation/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Rules, RPA, ML, LLMs, and Agents: The Decision Ladder</title>
      <link>https://cognaptus.com/academy/foundations/rules-rpa-ml-llm-agent-decision-ladder/</link>
      <pubDate>Thu, 23 Apr 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/foundations/rules-rpa-ml-llm-agent-decision-ladder/</guid>
      <description>Learn how to pick the simplest automation approach that actually fits the task, instead of jumping too quickly to complex AI.</description>
    </item>
    <item>
      <title>AI-Powered Email Sorting</title>
      <link>https://cognaptus.com/academy/operations/ai-powered-email-sorting/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/operations/ai-powered-email-sorting/</guid>
      <description>A grounded lesson on AI-assisted email triage, including taxonomy design, confidence thresholds, escalation rules, multilingual handling, and operational safeguards.</description>
    </item>
    <item>
      <title>Build a Simple AI Classification Pipeline</title>
      <link>https://cognaptus.com/academy/tools/build-a-simple-ai-classification-pipeline/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/tools/build-a-simple-ai-classification-pipeline/</guid>
      <description>A practical guide to building AI classification workflows, including schema design, threshold logic, review routing, MVP architecture, logging, and maintenance over time.</description>
    </item>
    <item>
      <title>AI for Ticket Triage and Case Routing</title>
      <link>https://cognaptus.com/academy/operations/ai-for-ticket-triage-and-case-routing/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/operations/ai-for-ticket-triage-and-case-routing/</guid>
      <description>A practical guide to AI-assisted ticket triage and case routing, including taxonomy design, ownership rules, confidence thresholds, escalation paths, and SLA-aware workflows.</description>
    </item>
    <item>
      <title>Smart Invoicing with AI</title>
      <link>https://cognaptus.com/academy/finance/smart-invoicing-with-ai/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/finance/smart-invoicing-with-ai/</guid>
      <description>A finance-control-oriented guide to AI-assisted invoicing, including field extraction, duplicate detection, three-way-match boundaries, exception queues, and audit design.</description>
    </item>
    <item>
      <title>Automate Reports with AI</title>
      <link>https://cognaptus.com/academy/operations/automate-reports-with-ai/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/operations/automate-reports-with-ai/</guid>
      <description>Learn how to apply AI to weekly and monthly reporting workflows with source-to-draft design, review checkpoints, KPI preservation, and narrative consistency rules.</description>
    </item>
    <item>
      <title>Generate Marketing Content at Scale</title>
      <link>https://cognaptus.com/academy/marketing/generate-marketing-content-at-scale/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/marketing/generate-marketing-content-at-scale/</guid>
      <description>A practical guide to using AI for content systems, including audience signals, quality gates, duplication control, editorial review, brand-risk checks, distribution planning, and pipeline metrics.</description>
    </item>
    <item>
      <title>AI Agents vs Workflows</title>
      <link>https://cognaptus.com/academy/foundations/ai-agents-vs-workflows/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/foundations/ai-agents-vs-workflows/</guid>
      <description>A practical guide to the difference between AI agents and workflows, with business examples, design trade-offs, and governance advice.</description>
    </item>
    <item>
      <title>Progress Is Not Completion: What CAP Reveals About Browser-Agent Readiness</title>
      <link>https://cognaptus.com/blog/2026-08-30-progress-is-not-completion-what-cap-reveals-about-browseragent-readiness/</link>
      <pubDate>Sun, 30 Aug 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-08-30-progress-is-not-completion-what-cap-reveals-about-browseragent-readiness/</guid>
      <description>CAP shows why browser-agent readiness depends less on visible progress than on reliably completing every required action and perception step across real websites.</description>
    </item>
    <item>
      <title>The Robot Needs a Shift Supervisor</title>
      <link>https://cognaptus.com/blog/2026-07-03-the-robot-needs-a-shift-supervisor/</link>
      <pubDate>Fri, 03 Jul 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-07-03-the-robot-needs-a-shift-supervisor/</guid>
      <description>A systematic study of hierarchical VLA agents shows that robot reliability depends less on simply adding hierarchy than on how planning, control, memory, observation, and handoff are orchestrated.</description>
    </item>
    <item>
      <title>Borrowed Hands Still Need a Grip</title>
      <link>https://cognaptus.com/blog/2026-06-27-borrowed-hands-still-need-a-grip/</link>
      <pubDate>Sat, 27 Jun 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-06-27-borrowed-hands-still-need-a-grip/</guid>
      <description>GLAM shows how heterogeneous robot demonstrations become useful only when their effects are grounded into a target-executable latent action space.</description>
    </item>
    <item>
      <title>OCR and the City: Why Document AI Still Needs Eyes</title>
      <link>https://cognaptus.com/blog/2026-06-08-ocr-and-the-city-why-document-ai-still-needs-eyes/</link>
      <pubDate>Mon, 08 Jun 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-06-08-ocr-and-the-city-why-document-ai-still-needs-eyes/</guid>
      <description>A comparison-based reading of arXiv 2606.02162, showing when OCR text, document images, fine-tuned Transformers, and prompt-based LLMs actually help enterprise document classification.</description>
    </item>
    <item>
      <title>Provenance, Not Providence: Why AI Answers Need Receipts</title>
      <link>https://cognaptus.com/blog/2026-05-09-provenance-not-providence-why-ai-answers-need-receipts/</link>
      <pubDate>Sat, 09 May 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-05-09-provenance-not-providence-why-ai-answers-need-receipts/</guid>
      <description>A business-focused reading of DataDignity, a new benchmark and method suite for tracing LLM outputs back to likely supporting training documents.</description>
    </item>
    <item>
      <title>Credit Where It’s Due: The New Reasoning Stack for Agentic AI</title>
      <link>https://cognaptus.com/blog/2026-05-07-credit-where-its-due-the-new-reasoning-stack-for-agentic-ai/</link>
      <pubDate>Thu, 07 May 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-05-07-credit-where-its-due-the-new-reasoning-stack-for-agentic-ai/</guid>
      <description>A research-cluster analysis of why reliable AI agents need better task structure, process evaluation, and credit assignment—not just larger models or longer chains of thought.</description>
    </item>
    <item>
      <title>Prompt and Circumstance: Why One Accuracy Number Is Not a Reliability Audit</title>
      <link>https://cognaptus.com/blog/2026-05-07-prompt-and-circumstance-why-one-accuracy-number-is-not-a-reliability-audit/</link>
      <pubDate>Thu, 07 May 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-05-07-prompt-and-circumstance-why-one-accuracy-number-is-not-a-reliability-audit/</guid>
      <description>A practical reading of a new multi-variant audit showing why AI model reliability depends on prompts, evaluators, calibration definitions, and parseability—not just benchmark accuracy.</description>
    </item>
    <item>
      <title>Receipts, Please: RAG’s New Evidence Stack</title>
      <link>https://cognaptus.com/blog/2026-05-07-receipts-please-rags-new-evidence-stack/</link>
      <pubDate>Thu, 07 May 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-05-07-receipts-please-rags-new-evidence-stack/</guid>
      <description>A research-cluster reading of why practical RAG systems now need retrieval discipline, sufficiency control, faithfulness training, verification tooling, and privacy-aware governance.</description>
    </item>
    <item>
      <title>Edge Cases: Why Graph World Models May Make AI Agents Less Lost</title>
      <link>https://cognaptus.com/blog/2026-05-04-edge-cases-why-graph-world-models-may-make-ai-agents-less-lost/</link>
      <pubDate>Mon, 04 May 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-05-04-edge-cases-why-graph-world-models-may-make-ai-agents-less-lost/</guid>
      <description>A practical reading of graph world models: how structured relational memory could make AI agents more reliable, inspectable, and useful in complex business environments.</description>
    </item>
    <item>
      <title>Look Who’s Reasoning Now: UpstreamQA and the Fine Print of Video AI</title>
      <link>https://cognaptus.com/blog/2026-05-02-look-whos-reasoning-now-upstreamqa-and-the-fine-print-of-video-ai/</link>
      <pubDate>Sat, 02 May 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-05-02-look-whos-reasoning-now-upstreamqa-and-the-fine-print-of-video-ai/</guid>
      <description>A practical reading of UpstreamQA: why modular reasoning can make video AI more interpretable, more accurate in some cases, and worse in others.</description>
    </item>
    <item>
      <title>Org-Charted Territory: Why AI Agents Need Middle Management</title>
      <link>https://cognaptus.com/blog/2026-04-28-orgcharted-territory-why-ai-agents-need-middle-management/</link>
      <pubDate>Tue, 28 Apr 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-04-28-orgcharted-territory-why-ai-agents-need-middle-management/</guid>
      <description>A practical reading of OneManCompany and why enterprise AI agents need organisational design, not just sharper prompts and shinier tools.</description>
    </item>
    <item>
      <title>Search Me If You Can: Why AI Agent Discovery Needs Receipts</title>
      <link>https://cognaptus.com/blog/2026-04-28-search-me-if-you-can-why-ai-agent-discovery-needs-receipts/</link>
      <pubDate>Tue, 28 Apr 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-04-28-search-me-if-you-can-why-ai-agent-discovery-needs-receipts/</guid>
      <description>AgentSearchBench shows why finding the right AI agent requires execution evidence, not just pretty descriptions.</description>
    </item>
    <item>
      <title>Clawing Back the Benchmark: When AI Agents Start Testing Themselves</title>
      <link>https://cognaptus.com/blog/2026-04-23-clawing-back-the-benchmark-when-ai-agents-start-testing-themselves/</link>
      <pubDate>Thu, 23 Apr 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-04-23-clawing-back-the-benchmark-when-ai-agents-start-testing-themselves/</guid>
      <description>ClawEnvKit shows how agent evaluation may shift from fixed benchmark artifacts to generated, verified, continuously refreshed test environments.</description>
    </item>
    <item>
      <title>Mind the Cut: Where Your AI Strategy Quietly Breaks</title>
      <link>https://cognaptus.com/blog/2026-04-11-mind-the-cut-where-your-ai-strategy-quietly-breaks/</link>
      <pubDate>Sat, 11 Apr 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-04-11-mind-the-cut-where-your-ai-strategy-quietly-breaks/</guid>
      <description>A business-oriented reading of the Cartesian cut: why the boundary between model and runtime determines whether AI agents remain governable, brittle, or truly autonomous.</description>
    </item>
    <item>
      <title>From Memory to Machinery: Why AI Agents Are Learning to Write Themselves</title>
      <link>https://cognaptus.com/blog/2026-03-19-from-memory-to-machinery-why-ai-agents-are-learning-to-write-themselves/</link>
      <pubDate>Thu, 19 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-03-19-from-memory-to-machinery-why-ai-agents-are-learning-to-write-themselves/</guid>
      <description>AgentFactory shows why the next useful step in AI agents may be less about remembering better and more about preserving executable work as reusable, auditable capability.</description>
    </item>
    <item>
      <title>The Memory Gap Nobody Budgeted For: Why Your AI Agents Keep Forgetting Each Other</title>
      <link>https://cognaptus.com/blog/2026-03-19-the-memory-gap-nobody-budgeted-for-why-your-ai-agents-keep-forgetting-each-other/</link>
      <pubDate>Thu, 19 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-03-19-the-memory-gap-nobody-budgeted-for-why-your-ai-agents-keep-forgetting-each-other/</guid>
      <description>A business reading of Governed Memory, showing why multi-agent AI needs shared memory, policy routing, schema feedback, and entity isolation—not just another RAG store.</description>
    </item>
    <item>
      <title>From Retry to Recovery: Teaching AI Agents to Learn from Their Own Mistakes</title>
      <link>https://cognaptus.com/blog/2026-03-18-from-retry-to-recovery-teaching-ai-agents-to-learn-from-their-own-mistakes/</link>
      <pubDate>Wed, 18 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-03-18-from-retry-to-recovery-teaching-ai-agents-to-learn-from-their-own-mistakes/</guid>
      <description>A close reading of LEAFE, a reflective-experience training framework that shifts AI agents from blind retry loops toward internalized recovery behavior.</description>
    </item>
    <item>
      <title>AI Academy</title>
      <link>https://cognaptus.com/academy/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/</guid>
      <description>Role-based learning paths, workflow playbooks, build specifications, governance controls, prototype evaluation, and reusable practice cases.</description>
    </item>
    <item>
      <title>AI for Business Operations</title>
      <link>https://cognaptus.com/academy/operations/</link>
      <pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/academy/operations/</guid>
      <description>Map real operations, design AI-assisted workflows, preserve ownership and service levels, and manage adoption from discovery through pilot.</description>
    </item>
    <item>
      <title>Audit the Bots: When AI Judges the Work of Other AI</title>
      <link>https://cognaptus.com/blog/2026-03-13-audit-the-bots-when-ai-judges-the-work-of-other-ai/</link>
      <pubDate>Fri, 13 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-03-13-audit-the-bots-when-ai-judges-the-work-of-other-ai/</guid>
      <description>A practical reading of CUAAudit and what its evidence says about using vision-language models to audit autonomous computer-use agents.</description>
    </item>
    <item>
      <title>Teaching Reinforcement Learning to Think Before It Acts</title>
      <link>https://cognaptus.com/blog/2026-03-09-teaching-reinforcement-learning-to-think-before-it-acts/</link>
      <pubDate>Mon, 09 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-03-09-teaching-reinforcement-learning-to-think-before-it-acts/</guid>
      <description>A mechanism-first reading of H2RL, a neuro-symbolic reinforcement learning framework that uses logic as training scaffolding rather than inference-time baggage.</description>
    </item>
    <item>
      <title>From Chatbots to Co‑Workers: The Architecture of Agentic AI</title>
      <link>https://cognaptus.com/blog/2026-03-07-from-chatbots-to-coworkers-the-architecture-of-agentic-ai/</link>
      <pubDate>Sat, 07 Mar 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-03-07-from-chatbots-to-coworkers-the-architecture-of-agentic-ai/</guid>
      <description>A mechanism-first reading of agentic AI: how planning, tools, memory, and feedback loops turn language models into operational systems—and why that also makes them harder to trust.</description>
    </item>
    <item>
      <title>Lost in the Links: When World Knowledge Isn’t Enough</title>
      <link>https://cognaptus.com/blog/2026-02-21-lost-in-the-links-when-world-knowledge-isnt-enough/</link>
      <pubDate>Sat, 21 Feb 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-02-21-lost-in-the-links-when-world-knowledge-isnt-enough/</guid>
      <description>LLM-WikiRace shows why agent reliability depends less on stored knowledge and more on planning, recovery, and loop control.</description>
    </item>
    <item>
      <title>The Reliability Gap: Why Smarter AI Agents Still Fail When It Matters</title>
      <link>https://cognaptus.com/blog/2026-02-19-the-reliability-gap-why-smarter-ai-agents-still-fail-when-it-matters/</link>
      <pubDate>Thu, 19 Feb 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-02-19-the-reliability-gap-why-smarter-ai-agents-still-fail-when-it-matters/</guid>
      <description>A mechanism-first reading of why agent accuracy is not the same as production reliability, and how firms should evaluate consistency, robustness, predictability, and safety before deployment.</description>
    </item>
    <item>
      <title>Lost in Translation: When 14% WER Hides a 44% Failure Rate</title>
      <link>https://cognaptus.com/blog/2026-02-13-lost-in-translation-when-14-wer-hides-a-44-failure-rate/</link>
      <pubDate>Fri, 13 Feb 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-02-13-lost-in-translation-when-14-wer-hides-a-44-failure-rate/</guid>
      <description>Why speech models can look reliable on benchmark metrics while still failing on the named entities that drive real-world routing, cost, and fairness.</description>
    </item>
    <item>
      <title>From Features to Actions: Why Agentic AI Needs a New Explainability Playbook</title>
      <link>https://cognaptus.com/blog/2026-02-09-from-features-to-actions-why-agentic-ai-needs-a-new-explainability-playbook/</link>
      <pubDate>Mon, 09 Feb 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-02-09-from-features-to-actions-why-agentic-ai-needs-a-new-explainability-playbook/</guid>
      <description>A practical reading of why feature attribution explains static predictions, but trajectory-level diagnostics are needed to understand failures in agentic AI systems.</description>
    </item>
    <item>
      <title>When Images Pretend to Be Interfaces: Stress‑Testing Generative Models as GUI Environments</title>
      <link>https://cognaptus.com/blog/2026-02-09-when-images-pretend-to-be-interfaces-stresstesting-generative-models-as-gui-environments/</link>
      <pubDate>Mon, 09 Feb 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-02-09-when-images-pretend-to-be-interfaces-stresstesting-generative-models-as-gui-environments/</guid>
      <description>GEBench shows why beautiful generated interfaces are not yet reliable environments for training or testing GUI agents.</description>
    </item>
    <item>
      <title>Simulate This: When LLMs Stop Talking and Start Modeling</title>
      <link>https://cognaptus.com/blog/2026-02-06-simulate-this-when-llms-stop-talking-and-start-modeling/</link>
      <pubDate>Fri, 06 Feb 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-02-06-simulate-this-when-llms-stop-talking-and-start-modeling/</guid>
      <description>A practical decision map for using LLMs in modeling and simulation without mistaking prompts, RAG, or temperature settings for engineering discipline.</description>
    </item>
    <item>
      <title>When Models Guess the Verb by Looking at the Drawer</title>
      <link>https://cognaptus.com/blog/2026-01-24-when-models-guess-the-verb-by-looking-at-the-drawer/</link>
      <pubDate>Sat, 24 Jan 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-01-24-when-models-guess-the-verb-by-looking-at-the-drawer/</guid>
      <description>A case-first reading of RCORE shows why video models can still confuse actions when object priors overpower temporal evidence.</description>
    </item>
    <item>
      <title>Cosmos Policy: When Video Models Stop Watching and Start Acting</title>
      <link>https://cognaptus.com/blog/2026-01-23-cosmos-policy-when-video-models-stop-watching-and-start-acting/</link>
      <pubDate>Fri, 23 Jan 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-01-23-cosmos-policy-when-video-models-stop-watching-and-start-acting/</guid>
      <description>A mechanism-first reading of Cosmos Policy, showing how latent frame injection turns a video diffusion model into a robot policy, world model, and planner.</description>
    </item>
    <item>
      <title>Probe, Then Commit: Why Solver Tuning Finally Grew Up</title>
      <link>https://cognaptus.com/blog/2026-01-19-probe-then-commit-why-solver-tuning-finally-grew-up/</link>
      <pubDate>Mon, 19 Jan 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-01-19-probe-then-commit-why-solver-tuning-finally-grew-up/</guid>
      <description>A practical reading of the Probe and Solve Algorithm, a two-phase method for tuning constraint programming solvers under real time budgets.</description>
    </item>
    <item>
      <title>When Goals Collide: Synthesizing the Best Possible Outcome</title>
      <link>https://cognaptus.com/blog/2026-01-16-when-goals-collide-synthesizing-the-best-possible-outcome/</link>
      <pubDate>Fri, 16 Jan 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-01-16-when-goals-collide-synthesizing-the-best-possible-outcome/</guid>
      <description>How multi-property LTLf synthesis turns impossible all-or-nothing specifications into computable frontiers of guaranteed outcomes.</description>
    </item>
    <item>
      <title>NPCs With Short-Term Memory Loss: Benchmarking Agents That Actually Live in the World</title>
      <link>https://cognaptus.com/blog/2026-01-10-npcs-with-shortterm-memory-loss-benchmarking-agents-that-actually-live-in-the-world/</link>
      <pubDate>Sat, 10 Jan 2026 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2026-01-10-npcs-with-shortterm-memory-loss-benchmarking-agents-that-actually-live-in-the-world/</guid>
      <description>A mechanism-first reading of MineNPC-Task, a Minecraft benchmark that shows how memory-aware agents should be tested before anyone trusts them in real workflows.</description>
    </item>
    <item>
      <title>Echoes, Not Amnesia: Teaching GUI Agents to Remember What Worked</title>
      <link>https://cognaptus.com/blog/2025-12-23-echoes-not-amnesia-teaching-gui-agents-to-remember-what-worked/</link>
      <pubDate>Tue, 23 Dec 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-12-23-echoes-not-amnesia-teaching-gui-agents-to-remember-what-worked/</guid>
      <description>A mechanism-first look at EchoTrail-GUI, a framework that turns stateless GUI agents into memory-augmented systems by collecting, filtering, retrieving, and reusing successful operating traces.</description>
    </item>
    <item>
      <title>When Rewards Learn to See: Teaching Humanoids What the Ground Looks Like</title>
      <link>https://cognaptus.com/blog/2025-12-21-when-rewards-learn-to-see-teaching-humanoids-what-the-ground-looks-like/</link>
      <pubDate>Sun, 21 Dec 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-12-21-when-rewards-learn-to-see-teaching-humanoids-what-the-ground-looks-like/</guid>
      <description>A mechanism-first reading of E-SDS, a framework that makes automated reward generation environment-aware for humanoid locomotion.</description>
    </item>
    <item>
      <title>Prompt-to-Parts: When Language Learns to Build</title>
      <link>https://cognaptus.com/blog/2025-12-20-prompttoparts-when-language-learns-to-build/</link>
      <pubDate>Sat, 20 Dec 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-12-20-prompttoparts-when-language-learns-to-build/</guid>
      <description>A mechanism-first reading of Prompt-to-Parts, where language models become useful for physical design not by imagining perfect 3D objects, but by compiling intent into constrained, inspectable part assemblies.</description>
    </item>
    <item>
      <title>ImplicitRDP: When Robots Stop Guessing and Start Feeling</title>
      <link>https://cognaptus.com/blog/2025-12-13-implicitrdp-when-robots-stop-guessing-and-start-feeling/</link>
      <pubDate>Sat, 13 Dec 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-12-13-implicitrdp-when-robots-stop-guessing-and-start-feeling/</guid>
      <description>A mechanism-first reading of ImplicitRDP, showing why force-aware robot policies need causal structure, not just extra sensor channels.</description>
    </item>
    <item>
      <title>When AI Becomes the Reviewer: Pairwise Judgment at Scale</title>
      <link>https://cognaptus.com/blog/2025-12-12-when-ai-becomes-the-reviewer-pairwise-judgment-at-scale/</link>
      <pubDate>Fri, 12 Dec 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-12-12-when-ai-becomes-the-reviewer-pairwise-judgment-at-scale/</guid>
      <description>A mechanism-first look at how LLMs can turn expensive proposal review into pairwise ranking, audit signals, and similarity checks without pretending the committee has disappeared.</description>
    </item>
    <item>
      <title>Replan, Rethink, Repeat: Why Vision-Language Models Make Better Closed‑Loop Planners</title>
      <link>https://cognaptus.com/blog/2025-11-16-replan-rethink-repeat-why-visionlanguage-models-make-better-closedloop-planners/</link>
      <pubDate>Sun, 16 Nov 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-11-16-replan-rethink-repeat-why-visionlanguage-models-make-better-closedloop-planners/</guid>
      <description>A comparison-led reading of why VLM robot planners need feedback, memory, and carefully tuned replanning rather than blind faith in more model calls.</description>
    </item>
    <item>
      <title>From Prototype to Profit: How IBM&#39;s CUGA Redefines Enterprise Agents</title>
      <link>https://cognaptus.com/blog/2025-11-02-from-prototype-to-profit-how-ibms-cuga-redefines-enterprise-agents/</link>
      <pubDate>Sun, 02 Nov 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-11-02-from-prototype-to-profit-how-ibms-cuga-redefines-enterprise-agents/</guid>
      <description>IBM’s CUGA pilot shows that enterprise agent value depends less on leaderboard glory than on governed tool use, provenance, regression testing, and measurable workflow compression.</description>
    </item>
    <item>
      <title>Fast but Flawed: What Happens When AI Agents Try to Work Like Humans</title>
      <link>https://cognaptus.com/blog/2025-11-01-fast-but-flawed-what-happens-when-ai-agents-try-to-work-like-humans/</link>
      <pubDate>Sat, 01 Nov 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-11-01-fast-but-flawed-what-happens-when-ai-agents-try-to-work-like-humans/</guid>
      <description>A workflow-level reading of human and AI work shows why agents are cheap, quick, programmatic, and still risky in the places businesses most want to automate.</description>
    </item>
    <item>
      <title>The Mr. Magoo Problem: When AI Agents &#39;Just Do It&#39;</title>
      <link>https://cognaptus.com/blog/2025-10-09-the-mr-magoo-problem-when-ai-agents-just-do-it/</link>
      <pubDate>Thu, 09 Oct 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-10-09-the-mr-magoo-problem-when-ai-agents-just-do-it/</guid>
      <description>A business-focused reading of Blind Goal-Directedness: why computer-use agents need trajectory-level judgement, not just better task completion.</description>
    </item>
    <item>
      <title>When More Becomes Smarter: The Unreasonable Effectiveness of Scaling Agents</title>
      <link>https://cognaptus.com/blog/2025-10-09-when-more-becomes-smarter-the-unreasonable-effectiveness-of-scaling-agents/</link>
      <pubDate>Thu, 09 Oct 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-10-09-when-more-becomes-smarter-the-unreasonable-effectiveness-of-scaling-agents/</guid>
      <description>A mechanism-first look at why wide scaling, behavior narratives, and comparative judging may matter more for computer-use agents than another heroic single rollout.</description>
    </item>
    <item>
      <title>Failures, Taxonomized: How Multi‑Level Reflection Turns Agents Into Self‑Learners</title>
      <link>https://cognaptus.com/blog/2025-10-02-failures-taxonomized-how-multilevel-reflection-turns-agents-into-selflearners/</link>
      <pubDate>Thu, 02 Oct 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-10-02-failures-taxonomized-how-multilevel-reflection-turns-agents-into-selflearners/</guid>
      <description>How SaMuLe turns failed agent traces into a reusable diagnostic layer—and what that means for enterprise automation.</description>
    </item>
    <item>
      <title>Paths &gt; Outcomes: Measuring Agent Quality Beyond the Final State</title>
      <link>https://cognaptus.com/blog/2025-10-02-paths-outcomes-measuring-agent-quality-beyond-the-final-state/</link>
      <pubDate>Thu, 02 Oct 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-10-02-paths-outcomes-measuring-agent-quality-beyond-the-final-state/</guid>
      <description>A practical reading of CORE, a path-based evaluation framework showing why tool-using AI agents must be judged by the sequence of actions they take, not only the state they leave behind.</description>
    </item>
    <item>
      <title>Repo, Meet Your Agent: Turning GitHub into a Workforce with EnvX</title>
      <link>https://cognaptus.com/blog/2025-09-14-repo-meet-your-agent-turning-github-into-a-workforce-with-envx/</link>
      <pubDate>Sun, 14 Sep 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-09-14-repo-meet-your-agent-turning-github-into-a-workforce-with-envx/</guid>
      <description>EnvX shows how repositories can become callable agents, but the real business value is disciplined software reuse—not fantasy staff replacement.</description>
    </item>
    <item>
      <title>Agents on the Clock: Turning a 3‑Layer Taxonomy into a Build‑Ready Playbook</title>
      <link>https://cognaptus.com/blog/2025-08-26-agents-on-the-clock-turning-a-3layer-taxonomy-into-a-buildready-playbook/</link>
      <pubDate>Tue, 26 Aug 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-08-26-agents-on-the-clock-turning-a-3layer-taxonomy-into-a-buildready-playbook/</guid>
      <description>A mechanism-first reading of agentic reasoning frameworks as operational control loops, not just model upgrades.</description>
    </item>
    <item>
      <title>From Text to Motion: How Manimator Turns Dense Papers into Dynamic Learning</title>
      <link>https://cognaptus.com/blog/2025-07-22-from-text-to-motion-how-manimator-turns-dense-papers-into-dynamic-learning/</link>
      <pubDate>Tue, 22 Jul 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-07-22-from-text-to-motion-how-manimator-turns-dense-papers-into-dynamic-learning/</guid>
      <description>Manimator shows how LLM pipelines can turn dense STEM material into first-draft explanatory animations, but its real value is production leverage rather than guaranteed pedagogy.</description>
    </item>
    <item>
      <title>The Butterfly Defect: Diagnosing LLM Failures in Tool-Agent Chains</title>
      <link>https://cognaptus.com/blog/2025-07-22-the-butterfly-defect-diagnosing-llm-failures-in-toolagent-chains/</link>
      <pubDate>Tue, 22 Jul 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-07-22-the-butterfly-defect-diagnosing-llm-failures-in-toolagent-chains/</guid>
      <description>A mechanism-first reading of how small parameter errors in LLM tool agents propagate into failed automation chains, and what operators should govern before they scale agents.</description>
    </item>
    <item>
      <title>Bridges and Biases: How LLMs Are Learning to Inspect Infrastructure</title>
      <link>https://cognaptus.com/blog/2025-07-21-bridges-and-biases-how-llms-are-learning-to-inspect-infrastructure/</link>
      <pubDate>Mon, 21 Jul 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-07-21-bridges-and-biases-how-llms-are-learning-to-inspect-infrastructure/</guid>
      <description>A mechanism-first reading of how multimodal LLMs can turn bridge NDE contour maps into inspection support, and why the real value is triage rather than autonomous certification.</description>
    </item>
    <item>
      <title>Mind the Gap: Fixing the Flaws in Agentic Benchmarking</title>
      <link>https://cognaptus.com/blog/2025-07-04-mind-the-gap-fixing-the-flaws-in-agentic-benchmarking/</link>
      <pubDate>Fri, 04 Jul 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-07-04-mind-the-gap-fixing-the-flaws-in-agentic-benchmarking/</guid>
      <description>Agentic benchmark scores can look precise while measuring broken graders, leaky environments, and trivial shortcuts rather than real agent capability.</description>
    </item>
    <item>
      <title>Half-Life Crisis: Why AI Agents Fade with Time (and What It Means for Automation)</title>
      <link>https://cognaptus.com/blog/2025-05-11-halflife-crisis-why-ai-agents-fade-with-time-and-what-it-means-for-automation/</link>
      <pubDate>Sun, 11 May 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-05-11-halflife-crisis-why-ai-agents-fade-with-time-and-what-it-means-for-automation/</guid>
      <description>A mechanism-first reading of why AI-agent reliability may decay exponentially with task length, and what that means for automation design.</description>
    </item>
    <item>
      <title>Body of Proof: Why Embodied AI Needs More Than One Mind</title>
      <link>https://cognaptus.com/blog/2025-05-09-body-of-proof-why-embodied-ai-needs-more-than-one-mind/</link>
      <pubDate>Fri, 09 May 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-05-09-body-of-proof-why-embodied-ai-needs-more-than-one-mind/</guid>
      <description>A category-based field map for understanding why multi-agent embodied AI is not just single-agent robotics with extra hardware.</description>
    </item>
    <item>
      <title>Evolving Beyond Bottlenecks: How Agentic Workflows Revolutionize Optimization</title>
      <link>https://cognaptus.com/blog/2025-05-08-evolving-beyond-bottlenecks-how-agentic-workflows-revolutionize-optimization/</link>
      <pubDate>Thu, 08 May 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-05-08-evolving-beyond-bottlenecks-how-agentic-workflows-revolutionize-optimization/</guid>
      <description>A mechanism-first reading of how foundation-model agents and evolutionary search could reduce the expert bottleneck in practical optimization—without pretending the experts can retire.</description>
    </item>
    <item>
      <title>How AI-Powered Automation SaaS Can Reshape Real Estate Brokerage in Southeast Asia</title>
      <link>https://cognaptus.com/blog/2025-03-23-how-aipowered-automation-saas-can-reshape-real-estate-brokerage-in-southeast-asia/</link>
      <pubDate>Sun, 23 Mar 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-03-23-how-aipowered-automation-saas-can-reshape-real-estate-brokerage-in-southeast-asia/</guid>
      <description>A practical look at how agent-centred AI SaaS can clean listings, qualify leads, and improve broker productivity in Southeast Asia without pretending to replace trust.</description>
    </item>
    <item>
      <title>Beyond Words: How Transformer Models Are Revolutionizing SaaS for Small Businesses</title>
      <link>https://cognaptus.com/blog/2025-03-21-beyond-words-how-transformer-models-are-revolutionizing-saas-for-small-businesses/</link>
      <pubDate>Fri, 21 Mar 2025 00:00:00 +0000</pubDate>
      <guid>https://cognaptus.com/blog/2025-03-21-beyond-words-how-transformer-models-are-revolutionizing-saas-for-small-businesses/</guid>
      <description>A practical explanation of how transformer models can move small-business SaaS from static record-keeping toward context-aware automation.</description>
    </item>
  </channel>
</rss>
