The dashboard every client gets.
Every build ships with reporting like this for your systems alone: runs, hours removed, success rate, approvals waiting, and incidents. This page is a working demo with representative figures, so you can see exactly what you would be looking at after launch.
Demo data · refreshed on load
Runs, last 30 days.
One run is one triggered execution, from trigger to final result. Weekends dip because most of what we run follows office hours.
Runs per day, all systems
Sep 1 to Sep 30 · 64,374 total- Weekday
- Weekend
- Today, partial
Hours removed per month, trailing 12.
We time the manual process at scoping, multiply by the volume each system handled that month, and subtract the time people still spend on reviews.
Manual hours removed, by month
Oct 2025 to Sep 2026of 64,374 runs finished without an engineer touching them
open at the last nightly summary, oldest waiting 14 minutes
on evaluation sets of past cases with known answers, across 23 agents, re-scored monthly
Systems in production.
The ten highest-volume systems this month, with clients named by industry only. The remaining 136 are mostly small automations that run a few times a day.
| Client | System | Type | Runs / mo | Status |
|---|---|---|---|---|
| Freight brokerage | Quote extraction | Workflow | 2,318 | Healthy |
| B2B SaaS | First-line support agent | Agent | 4,860 | Healthy |
| Dental group | Recall reminders | Automation | 6,120 | Healthy |
| Specialty retail | Order exception routing | Workflow | 4,380 | Healthy |
| Marine services | Invoice to accounting sync | Automation | 3,940 | Healthy |
| Property management | Maintenance ticket triage | Agent | 3,420 | Degraded The CRM vendor has limited request rates since Sep 26. Retries are extended and the backlog clears nightly. |
| Freight brokerage | Carrier onboarding checks | Automation | 1,210 | Healthy |
| Professional services | Proposal drafting assistant | Agent | 890 | Healthy |
| Home builder | Permit document intake | Workflow | 640 | Healthy |
| Insurance brokerage | Policy renewal outreach | Automation | 0 | Paused Seasonal. Resumes Nov 1 at the client’s request. |
Incidents, last 90 days.
Anything that stopped a production system for more than ten minutes, or caused a run to end in the wrong state. Three this quarter, all resolved, none with data loss.
- Sep 18, 2026 41 min
CRM vendor change broke field mapping
ResolvedCause. The CRM vendor renamed two properties in a scheduled release. The proposal drafting assistant began failing on contact lookups. Runs queued and none were lost.
Fix. We updated the mapping within the hour. A nightly test now runs against the CRM test environment and alerts us to any field change before it reaches live systems.
- Aug 27, 2026 2 h 10 min
AI provider outage triggered backup
ResolvedCause. The primary AI model provider returned errors for just over two hours. The backup provider took over automatically after four minutes of elevated failures.
Fix. No data was lost. Response time rose about 30% during the window, and sampled accuracy stayed within limits. We lowered the switch-over threshold from four minutes to ninety seconds.
- Jul 14, 2026 26 min
Expired credential on an accounting sync
ResolvedCause. A login token for the accounting system expired after an admin change on the client side. The first failed run alerted us at 06:12.
Fix. We renewed the credential and replayed the queued batch. Every credential we hold now has an expiry alert that warns 14 days ahead.
How these numbers are made.
What counts as a run
One run is one triggered execution of a system, from the trigger to its final result. A workflow started by a signal from another tool is one run, no matter how many steps follow. An agent handling a support ticket is one run per ticket, however many messages it takes. A scheduled update counts once each time it fires, even when it processes several hundred records.
How hours saved are measured
During scoping we time the manual process as it exists. We take the median of at least ten observed instances per step. We multiply that baseline by the volume the system handled in the month, then subtract the time people still spend on reviews and approvals in the new process. We re-measure at the 90-day review and keep whichever figure is lower.
What success means
A run succeeds when it reaches its intended final result without an engineer stepping in. Automatic retries that eventually complete count as success. Runs that end in a deliberate handoff to a person count as success when the handoff was the right call. We sample and check that weekly. Runs still waiting on an approval are excluded until they finish.
How agent accuracy is scored
Before launch, we score every agent against an evaluation set: at least 200 of the client's own past cases with known right answers. We re-score monthly as the model, the instructions, or the data change. The figure on this page is the median across all agents in production. Individual results are shared with each client, and one is public in the Halcyon case study.
Why some systems are excluded
Systems in their first 14 days are left out while their baseline settles. Systems paused at the client's request are listed but contribute no runs. Pilots that have not gone live are not counted, and neither are systems that clients have taken fully in-house, since we no longer see their data.
Rounding and timing
We round hours to the nearest hundred and percentages to one decimal place. Run counts are exact as of the nightly summary. The page is regenerated as demo data each time the page loads. We replace client names with industries and never publish a figure that could identify a single client's volume.
Want these numbers for your own operations?
Every build ships with a dashboard like this for your systems alone.