The Agent Is 20% of the Work. The Platform Is the Other 80%.
An enterprise AI agent deployed for payroll processing achieved 94% accuracy in testing but only 70% in production due to unanticipated real-world input variations. The team improved accuracy to 98% through shadow testing and infrastructure enhancements, not model changes. This highlights that successful AI deployment depends more on platform engineering than on the agent model itself.
- ▪The payroll AI agent processed over 3,000 emails daily, performing six-step data extraction and classification.
- ▪Test accuracy was 94%, but production accuracy dropped to 70% due to unanticipated inputs like typos, screenshots, and conflicting instructions.
- ▪Shadow testing over four weeks, where AI outputs were reviewed alongside human work, increased effective accuracy to 98% without changing the agent's model.
- ▪The agent engine represented only 20% of the work, while platform infrastructure like evaluation pipelines and input governance made up the remaining 80%.
- ▪Without proper evaluation infrastructure, teams risk outsourcing their AI learning to vendors who control their testing and feedback loops.
DEV.to (Top) files mainly under programming. We currently carry 4,924 of its stories.
Story provenance
Source · retrieval · rights · ranking — open for full record
inspect →
Story provenance
Attribution is not the same as permission. This drawer separates discovery metadata, excerpts, WeSearch-generated summaries, reuse status, and whether the publisher receives the visit. Nothing here claims a legal grant the publisher has not made.
Record
| Original publisher | DEV.to (Top) |
| Canonical URL | https://dev.to/todd_linnertz_871a076f68e/the-agent-is-20-of-the-work-the-platform-is-the-other-80-4cf8 |
| Publication time | Sun, 17 May 2026 04:56:38 +0000 |
| Retrieval time | 2026-05-17T05:33:58.675Z |
| Last seen | 2026-05-17T05:33:58.675Z |
| Headline source | Publisher (no WeSearch rewrite) |
| Excerpt source | publisher body |
| Excerpt method | First ~120 words (~800 chars) of extracted publisher body, fair-use limited. |
| Summary | WeSearch · cerebras-chat (WeSearch summarizer) |
| Summary source text | contentText |
| Citation coverage | Summary is a WeSearch-generated derivative; primary citation is the original publisher URL. |
| Cluster | 49ca9i0df3k8 |
| Cluster logic | Grouped by semantic title/content similarity across sources within a rolling window. Same-publisher template collisions are excluded from coverage comparison. |
| Ranking reason | Story pages are not engagement-ranked. Hub feeds use recency, with optional source-diversified chronological ordering (cap consecutive stories per source). No personalized ranking. |
| Publisher visit | Yes — open original |
| Substitutes article? | No — link-out required for full text |
Rights status (four layers)
WeSearch handling by dimension
| Indexing | May the item be indexed (stored, ranked, made findable)? | Allowed |
| Snippet | May a short excerpt of the publisher's text be shown? | Allowed |
| AI summary | May WeSearch generate its own short summary of the article? | Limited |
| Retrieval / RAG | May the content be exposed for third-party retrieval-augmented generation? | Not asserted |
| Model training | May the content be used to train AI models? | Not asserted |
| Commercial reuse | May the content be reused commercially? | Not permitted |
Basis: Derived from the published RSS/Atom feed. Contact: [email protected]. Reviewed: 2026-07-24.
Opening excerpt (first ~120 words) tap to expand
try { if(localStorage) { let currentUser = localStorage.getItem('current_user'); if (currentUser) { currentUser = JSON.parse(currentUser); if (currentUser.id === 3861685) { document.getElementById('article-show-container').classList.add('current-user-is-article-author'); } } } } catch (e) { console.error(e); } Todd Linnertz Posted on May 17 • Originally published at devopsdiary.blog The Agent Is 20% of the Work. The Platform Is the Other 80%. #ai #platformengineering #devops #mlops Governing AI in the Enterprise (2 Part Series) 1 AI Doesn't Fix Your Development Problems. It Accelerates Them. 2 The Agent Is 20% of the Work. The Platform Is the Other 80%. Originally published at devopsdiary.blog. Post F-AID1 in the "Governing AI in the Enterprise" series.
…
Excerpt limited to ~120 words for fair-use compliance. The full article is at DEV.to (Top).