01Discovery
We map the workflow, the systems, the volumes and the exceptions.
You getA written opportunity assessment with the cost of the status quo.We design, build and operate autonomous agents that qualify leads, triage support, reconcile data and run multi-step workflows end to end, inside the systems your team already uses, with a person in the loop wherever being wrong would be expensive.
Most organisations have already tried a chatbot. It answered questions, deflected some tickets, and left the actual work exactly where it was. An agent is a different thing: it is given a goal, a set of tools and a boundary, and it operates inside your systems until the task is complete or it needs a human.
That shift (from answering to acting) is what makes agents worth the engineering. It is also what makes them harder to build well. An agent that acts on your live systems has to be constrained, observable and reversible, or it will create more work than it removes.
The failure mode is rarely the model. It is the plumbing around it: whether the agent can see the right data, whether it has real permission to act, what happens when it is unsure, and whether anyone can reconstruct afterwards what it did and why.
That plumbing is the majority of the build, and it is the part we have spent 23 years doing for enterprise systems that could not afford to break.
We would rather talk you out of one than sell you the wrong thing. This is the test we apply before proposing anything.
| If the process… | Use | Because |
|---|---|---|
| Has stable rules and structured input | Rules-based automation (RPA) | Cheaper to run, trivially auditable, and it cannot improvise. Adding a model here buys you nothing and costs you predictability. |
| Varies in wording, format or judgement | An AI agent | The variation is exactly what a rules engine cannot absorb. This is where an agent earns its cost. |
| Is high-volume but mostly routine, with a difficult tail | Agent plus escalation | The agent clears the routine majority; the tail routes to a person with the context already assembled. |
| Is rare, high-stakes and relationship-driven | Leave it with people | Low volume means no payback, and the cost of being wrong is high. Automate the preparation around it instead. |
These are the patterns we see clear their cost quickest, high volume, well-bounded, and currently absorbing hours of skilled attention.
Enrich, score and route inbound enquiries the moment they arrive, so nothing waits for the next working day.
CRM · web forms · enrichment APIsClassify, prioritise and draft the first reply, with the genuinely difficult tickets escalated, not guessed at.
Ticketing · knowledge base · order historyMatch invoices, payments and ledger entries; surface only the exceptions a person needs to judge.
ERP · banking feeds · ledgerRead contracts, claims, forms and statements into structured data, with confidence scores and a review queue.
Document store · OCR · line-of-business systemAssemble recurring reports from live systems and write the commentary, ready for a human to sign off.
Data warehouse · BI · emailAnswer staff questions from your own policies and documentation, and cite the source so the answer can be checked.
Intranet · HR system · policy libraryMost clients start with a single scoped agent and a clear success measure. Some come to us with a stalled internal build and need the integration and governance layer finished properly.
Nothing is replaced. The agent sits alongside your systems and uses them the way a member of staff would; through interfaces, with permissions, leaving a record.
We build the answers in from the first sprint, because retrofitting them after a successful pilot is where most agent programmes stall.
Least-privilege service accounts, scoped per agent and per system. It gets the access the task needs and nothing beyond it.
Reversible actions run autonomously. Anything irreversible (payment, deletion, external communication) waits at an approval gate.
Every decision, tool call and input is logged and reviewable, so an outcome can be reconstructed months later.
An evaluation suite runs against every change. Real failures become test cases, so the same mistake cannot return silently.
Chosen deliberately, not by default. Where residency or confidentiality demands it, models run inside your own environment.
Per-task cost instrumented from day one, with ceilings, alerting, and cheaper models handling the routine steps.
We map the workflow, the systems, the volumes and the exceptions.
You getA written opportunity assessment with the cost of the status quo.Use cases prioritised, the first agent scoped, success measure agreed.
You getA costed roadmap and a defined pilot.Agent developed against your data, with guardrails and evaluation from sprint one.
You getA working agent in staging, reviewable every two weeks.Connected to live systems and validated against real historical cases.
You getEvaluation results against your own data, not a benchmark.Launch, monitor, retrain and widen scope as confidence builds.
You getMonthly performance and cost reporting, and a documented handover.The gap between a convincing demo and an agent that survives a Tuesday is where these projects are won and lost, and it is almost entirely about your data, your edge cases and your people. That is not work anyone can do from a distance, which is why we staff it with forward deployed engineers.
The engineer sits inside your rituals and your repository, learning the process the agent is meant to run instead of a written description of it.
Real documents, real exceptions, the record that has been wrong since 2014. Nobody finds these from a specification, and they decide whether the agent is trusted.
Measured on whether the workflow runs in your hands, not on whether the sprint closed. It changes what somebody does when a requirement turns out to be wrong.
Evaluation sets, drift, the prompt that stops working after a model update. An agent with nobody watching it degrades quietly and nobody notices until a customer does.
Your engineers work in the same repository from week one. The point is that the capability stays with you when the engagement ends.
If you keep the engineer on, they already know the system. If you do not, the documentation was written as we went. Either way there is no re-comprehension bill.
| Shape | When it fits | Duration | Commercial model |
|---|---|---|---|
| Readiness assessment | You suspect agents apply but cannot yet name the workflow. | 2–3 weeks | Fixed price |
| Pilot agent | One workflow identified, and you want it proven before committing further. | 6–8 weeks | Fixed price |
| Dedicated team | The pilot worked and there is a backlog behind it. | Rolling monthly | Per-person monthly, your backlog |
The three below are the ones this page does not already answer. Anything more specific, put it to us directly.
Per-task cost is instrumented from the first week, with ceilings and alerting per workflow, so you can see what each completed task costs rather than receiving one opaque monthly number. Cheaper models handle routine steps and expensive ones are reserved for the steps that need judgement, that routing is usually where most of the saving sits. If a workflow stops paying for itself, you will see it in the numbers before we do.
That is a decision you make, not one we make for you. Where confidentiality or residency rules require it, we deploy open-weight models inside your own cloud tenancy or data centre, and nothing leaves your boundary. Where a commercial API is acceptable, we use providers with zero-retention terms and restrict what is sent to the minimum the task needs. Either way, the agent sees only the systems and fields you scope it to, and every retrieval is logged.
You own the orchestration code, the prompts, the evaluation sets and the integration layer, they are in your repositories from the first commit, not ours. Because the build is model-agnostic, you are not tied to a provider either. We hand over with documentation and a working session for your engineers; a number of clients run their agents themselves now and call us only when they want another one built.