Estate ratings and performance
One rating scale, four levels of zoom
Every agent rolls into a sub-function, every sub-function into a function, every function into the enterprise. The rating is not a satisfaction score. It is a judgment about whether this estate could be handed to an auditor tomorrow, and it is deliberately harsher than the autonomy number sitting next to it.
Health
How each function is performing
The measures every function scorecard draws on, pulled from the pages that own them.
Enterprise estate rating
Sound, with a named weakness under active work.
The seven dimensions behind every rating
Enterprise mean across all rated sub-functions, with the weight each dimension carries
Control Coverage
75 · w 18%Share of decision paths that pass through a named gate rather than around one.
Evidence Quality
75 · w 16%Whether the trail an agent leaves would survive an external inspection unaided.
Override Discipline
83 · w 14%How often humans reverse the agent, and whether the reversal is recorded with a reason.
Data Integrity
77 · w 14%Confidence in the upstream records the agent is reasoning over.
Recovery Readiness
76 · w 14%Time and cost to reverse a bad run, tested rather than assumed.
Cost Transparency
79 · w 10%Whether cost to serve, including inference, is measured per unit rather than estimated.
Judgment Readiness
59 · w 14%Capacity to handle the cases the rules do not cover. This score is deliberately lowest where autonomy is highest.
Rating distribution
How the 154 sub-function estates actually land
No estate is rated AAA. That is intentional. A perfect control rating on a two-year-old operating model would say more about the rater than the estate.
Function scorecards
Ranked by estate rating, not by autonomy. The two frequently disagree.
| Function | Rating | Score | Autonomy now / cap | Agents | Tasks 24h | Success | Override | Falling | Open esc. | Incidents | Open the function record |
|---|---|---|---|---|---|---|---|---|---|---|---|
| A | 77 | 78% / 88% | 90 | 108,899 | 93% | 7.5% | 0 | 3 | 1 | Line map | |
| A | 77 | 75% / 84% | 110 | 141,911 | 93% | 7.2% | 0 | 1 | 1 | Line map | |
| A | 76 | 66% / 77% | 90 | 104,682 | 93% | 7.6% | 0 | 7 ·1 | 1 | Line map | |
| A | 75 | 70% / 79% | 98 | 112,472 | 93% | 7.5% | 0 | 2 | 1 | Line map | |
| A | 75 | 74% / 83% | 88 | 117,271 | 93% | 7.1% | 0 | 2 | 1 | Line map | |
| A | 75 | 75% / 82% | 92 | 111,744 | 92% | 7% | 0 | 6 | 1 | Line map | |
| A | 75 | 74% / 81% | 91 | 128,484 | 93% | 7.1% | 0 | 2 ·1 | 1 | Line map | |
| A | 75 | 76% / 84% | 89 | 110,068 | 93% | 7.7% | 0 | 5 | 1 | Line map | |
| A | 75 | 66% / 76% | 111 | 143,904 | 93% | 7.1% | 0 | 4 ·1 | 1 | Line map | |
| A | 75 | 75% / 84% | 119 | 134,875 | 93% | 7.7% | 0 | 1 | 3 | Line map | |
| BBB | 74 | 72% / 78% | 99 | 138,436 | 92% | 7.1% | 0 | 2 | 1 | Line map | |
| BBB | 73 | 69% / 77% | 86 | 92,612 | 93% | 8.2% | 0 | 2 | 1 | Line map | |
| BBB | 73 | 61% / 69% | 106 | 129,901 | 93% | 6.8% | 0 | 2 | 1 | Line map | |
| BBB | 72 | 66% / 72% | 92 | 109,355 | 93% | 7.2% | 0 | 7 ·2 | 1 | Line map |
The "open esc." column shows escalations still waiting on a person, with tier-1 count after the dot. A function can hold a strong rating and still carry escalations. The rating measures whether the estate is controlled, not whether it is quiet.
Ten estates to fix first
Lowest rated sub-functions across all 14 functions
Employment Legal
Legal · 8 agents · 10,567 units/wk
51% / 56%
The lowest ceiling in the entire estate alongside Litigation. Employment law is jurisdiction-specific and personally consequential; agents prepare, qualified counsel always decides.
Network Design
Supply Chain · 8 agents · 7,324 units/wk
56% / 60%
Deliberately capped low. Network modeling is fully agentic, but opening or closing a facility is a capital decision with employment consequences that only a board can take.
Discovery Research
Research & Development · 7 agents · 2,471 units/wk
71% / 78%
Literature synthesis and hypothesis generation are exactly where agents outperform — they read everything. A scientist validates before anything becomes a program.
Innovation Portfolio
Research & Development · 8 agents · 11,992 units/wk
58% / 64%
Agents assemble evidence and even recommend killing a program. Stopping or funding research is a capital decision with careers attached, so a committee owns it.
Environment Health & Safety
Administration · 11 agents · 8,482 units/wk
50% / 54%
Deliberately the lowest ceiling in Administration. Where a decision can injure someone, a competent human signs. Agents accelerate investigation, never the verdict.
Brand & Positioning
Marketing · 9 agents · 7,593 units/wk
58% / 64%
Agents author the positioning with evidence and dissent attached. What the company says about itself is ratified by a named executive.
Territory & Quota Planning
Sales · 9 agents · 135 units/wk
46% / 66%
Segmentation and capacity math are fully agentic. Quota is a compensation commitment to named people, so the final allocation carries a human signature.
Account Planning
Sales · 9 agents · 4,088 units/wk
64% / 72%
Agents author the account strategy end to end. The rep owns the relationship and ratifies the plan they will actually run.
Commercial Advisory
Legal · 8 agents · 14,214 units/wk
63% / 64%
Agents draft the advice and the reasoning. Legal advice is privileged and relied upon, so a qualified human signs every answer that leaves the function.
Corporate & Governance
Legal · 8 agents · 7,927 units/wk
68% / 72%
Assembly, calendaring and filing mechanics are automated. Minutes and authority changes are approved by named officers because they are the legal record of what the company decided.
Ten estates ready for more
Highest rated, and the cap that is still holding them
Intake & Routing
Customer Service · 10 agents · 4,225 units/wk
84% / 94%
Management Reporting
Finance · 8 agents · 9,301 units/wk
81% / 90%
H2R Control Tower
AI GBS & Global Business Services · 12 agents · 25,526 units/wk
78% / 91%
Onboarding
Human Resources · 9 agents · 18,493 units/wk
80% / 88%
Warehouse & Fulfillment
Supply Chain · 8 agents · 18,305 units/wk
84% / 88%
CRM Data Stewardship
Revenue Operations · 8 agents · 12,025 units/wk
83% / 92%
Demand Qualification
Sales · 9 agents · 9,651 units/wk
83% / 90%
Knowledge Management
Customer Service · 10 agents · 8,585 units/wk
71% / 80%
Tier-1 Resolution
Customer Service · 10 agents · 3,499 units/wk
81% / 90%
Hospitality & Reception
Administration · 9 agents · 9,810 units/wk
83% / 90%
Agent watchlist
Trust falling, override rate high, or currently not running clean
Build-Buy-Partner Agent
Research & Development · Technology Scouting
86% ok · 17% ovr
47 tasks · A2 / A3
Brand Health Agent
Marketing · Brand & Positioning
87% ok · 17% ovr
66 tasks · A2 / A2
Opportunity Management Challenger Agent
Sales · Opportunity Management
87% ok · 17% ovr
189 tasks · A2 / A2
Investor Relations Support Challenger Agent
Finance · Investor Relations Support
87% ok · 17% ovr
97 tasks · A2 / A2
IP Strategy Agent
Legal · IP Portfolio
88% ok · 17% ovr
16 tasks · A2 / A3
Portfolio Bets Agent
Research & Development · Innovation Portfolio
89% ok · 17% ovr
171 tasks · A2 / A2
R&D Analytics Orchestrator
Research & Development · Portfolio Analytics
89% ok · 17% ovr
102 tasks · A3 / A4
Record-to-Report & Close Challenger Agent
Finance · Record-to-Report & Close
89% ok · 17% ovr
208 tasks · A2 / A4
Proposal Quality Agent
Sales · Solution & Proposal
90% ok · 17% ovr
111 tasks · A3 / A3
Culture Orchestrator
Human Resources · Culture & Engagement
90% ok · 17% ovr
122 tasks · A2 / A2
Conflict Detection Agent
Information Technology · Change & Release
90% ok · 17% ovr
63 tasks · A4 / A4
Retention Doctrine Agent
Sales · Expansion & Renewals
91% ok · 17% ovr
9 tasks · A2 / A3
Sourcing & RFx Challenger Agent
Procurement & Source-to-Pay · Sourcing & RFx
91% ok · 17% ovr
197 tasks · A2 / A3
Network Orchestrator
Information Technology · Network
91% ok · 17% ovr
190 tasks · A3 / A3
Agents earning their ceiling
Highest success rate at real volume, lowest override
Renewal Calendar Agent
Revenue Operations · Renewals & Churn Ops
99% ok · 1% ovr
3,898 tasks · A4 / A4
Goods Receipt Agent
AI GBS & Global Business Services · S2P Control Tower
99% ok · 1% ovr
3,610 tasks · A4 / A4
Phishing Response Agent
Information Technology · Cybersecurity Operations
99% ok · 1% ovr
3,160 tasks · A3 / A3
Reharvest Agent
Information Technology · Asset & License
99% ok · 1% ovr
1,163 tasks · A4 / A4
Reliability Trend Agent
Engineering · Engineering Analytics
99% ok · 2% ovr
3,104 tasks · A4 / A4
Handoff Agent
Customer Service · Tier-1 Resolution
99% ok · 2% ovr
2,481 tasks · A4 / A4
Release Coordination Agent
Information Technology · Application Management
99% ok · 2% ovr
490 tasks · A3 / A3
Model vs Commit Agent
Revenue Operations · Forecast Integrity
99% ok · 3% ovr
3,840 tasks · A4 / A4
Channel Allocation Agent
Marketing · Demand Generation
99% ok · 3% ovr
3,651 tasks · A4 / A4
Travel Coordination Agent
Administration · Executive Support
99% ok · 3% ovr
2,702 tasks · A3 / A3
Stakeholder Map Agent
Sales · Account Planning
99% ok · 3% ovr
2,136 tasks · A3 / A3
Feedback Ingestion Agent
Customer Service · Voice of Customer
99% ok · 3% ovr
1,523 tasks · A3 / A3
Order Validation Agent
Revenue Operations · Booking & Order Management
99% ok · 3% ovr
1,340 tasks · A4 / A4
Discovery Agent
Information Technology · Asset & License
99% ok · 3% ovr
1,204 tasks · A4 / A4
Actions
What is waiting on a person
A scorecard is only worth reading if somebody is expected to act on a red. These are this week’s reds.
Ratify or retire an unratified position
A strategy position the scorecard cannot yet score against.
Clear an unresolved escalation
Carried on the scorecard of the function that raised it.
Answer for a breached service level
Shows red on the owning function’s scorecard.
Answer for a breached budget cap
Cost line above the cap on the record.
Operations
What this desk is allowed to start
A surface that only reports is not operable. This is the work this page can set in motion, and the bound it runs into.
Trigger and bound
This desk can compose a scorecard: pull each measure from its owning page, apply the declared threshold and publish the result. It cannot change a threshold, override a rating, or exclude a measure — a scorecard here is a view over other pages, never a separate set of numbers.
Live observability
What the record shows right now
What is currently red across the measures every function scorecard draws on.
Current distribution
61 reds
Is policy and strategy coming to fruition
Whether the written intent is holding here
14 functions are scored against 61 ratified positions and 395 active policy items.
Nothing on the record settles this
The intent is a single scorecard per function that a leader can read in a minute and act on. Structurally that holds: every measure shown is drawn from the page that owns it, so there is no second version of a number anywhere on this platform. What it cannot yet do is weight: the scorecard treats a breached service level and an unratified strategy position as comparable signals, because no agreed weighting exists. Reading it as a ranking of functions would be wrong.
How to read this without fooling yourself
A high success rate on a high-volume transactional agent is close to meaningless on its own. The work is repeatable, so it should be near perfect. The number that matters there is override rate and the reason attached to each override.
Judgment Readiness is scored lowest exactly where autonomy is highest. That is the honest tension in this model: the more of the routine an estate has automated, the less practiced its people are on the cases that fall outside the rules.
Nothing on this page is a benchmark. These are this estate's own measurements against its own control set. Comparing a function to an external figure would require the same definitions on both sides, and those definitions rarely survive contact.