Workflow & Process Engine
Everyone's busy. Nothing's moving.
i. what is ops?
Ops follows work from beginning to end and explains how it actually gets done.
Describe a process, an organization, or something that keeps going wrong, and Ops traces it: what starts it, what happens at each stage, who does the work, where it waits, where responsibility changes hands, where it breaks, and what comes out.
The work can be almost anything. A hospital discharge, a permit application, a restaurant on a Friday night, a supply chain, a volunteer rota, a morning routine with three kids and eight minutes. If something moves through stages and comes out changed, Ops can follow it.
ii. who is ops for?
something is failing and you can't see where
You don't want to engineer a workflow. You want a painful headache to go away. Something in the setup is broken, everyone is busy, and nobody can point at the place it stops.
you just want to know how it actually works
You are curious about the machinery behind ordinary things. How the package gets there. How the ward gets emptied. How the colony reorganizes. Not because you run one — because you want to see it.
you're new and everyone assumes you already know
New job, new industry, or a promotion into something you are expected to manage but nobody explained. You would rather find out than ask and look like you should have known.
you're writing it and it has to be right
A novel, an article, a script, a piece of research. You need the real mechanics — how long it takes, who signs it, what happens when it goes wrong — not a plausible guess.
you run the thing and it has to hold
A kitchen, a clinic, a shop, a crew, a rota. It works now. You want to know what breaks first when it gets busier, and what to watch so you see it coming.
iii. what can i ask?
Ordinary language. No process terminology required.
understand it
details
what actually happens
"What actually happens after I hit submit on a tax return?" · "What happens to my suitcase after I check it?" · "What happens after I file an insurance claim?" · "How does a package cross the country?" · "How does a building permit go from submission to approval?" · "How does recycling actually get processed?" · "What happens to my blood after I donate it?" · "How does a court case get scheduled?"
how do they handle it
"How do they handle surges of people at a concert gate?" · "How does a hotel turn over three hundred rooms in four hours?" · "How does ER triage decide who goes first?" · "How does a blood drive keep track of what it has?" · "How are organ transplants matched in real time?" · "How does a warehouse avoid running out of stock?" · "How do food delivery apps coordinate drivers?" · "How does rideshare surge pricing work behind the scenes?"
what's the step by step
"What's the step-by-step for getting a liquor license?" · "What's the step-by-step for registering a small business?" · "What's the step-by-step for getting a car through an inspection?" · "What's the step-by-step for a planning application?"
what is this called
"What's a bottleneck?" · "What's the difference between responsibility and accountability?" · "What does a handover actually mean?" · "Why do people say the constraint sets the pace?"
find out why it's broken
details
where does the time go
"Where does the time go when you order a custom couch?" · "Our tickets take three days to close but the actual work is twenty minutes." · "Why does it take so long to get a medical permit approved?" · "Why does customs clearance take so long?" · "Why does a two-day approval take three weeks?" · "Why is a visa application still being vetted after four months?"
where's the hold up
"Why does my package always get stuck at the same hub?" · "Where is the hold up in an insurance claim?" · "Why do invoices pile up at approval?" · "Why does work keep accumulating at quality inspection?" · "Why does the kitchen fall apart every Friday around seven?"
why does it keep going wrong
"Where did a lost customer order actually fail?" · "Why do jobs get dropped between sales and installation?" · "Why do the same mistakes keep reaching the customer?" · "What are the common mistakes when setting up a returns process?" · "Nobody knows whose job this is — how do I work out who should own it?"
set one up
details
how do i set one up
"How do I set up a returns process from scratch?" · "Design a flow for getting three kids out the door in the morning." · "What do I need to start a volunteer rota for a food bank?" · "Build an onboarding process for new staff." · "How should work move from sales to production?" · "What's the best way to organize pickup for eighty customers?" · "How do I run a smooth handover between shifts?"
which way is better
"How does online grocery ordering differ from shopping in person?" · "Centralized versus decentralized kitchen prep — what does each cost?" · "Should invoice matching be automated?" · "Is it faster to have one person do the whole thing or split it up?" · "Dark store fulfilment versus a normal shop — what changes?"
what happens if
"What happens if a kitchen runs out of chicken mid-shift?" · "What breaks first if order volume doubles?" · "What happens if a supplier misses a delivery window?" · "Why did our scheduling stop working once we hit forty staff?" · "Which manual steps won't survive growth?" · "What happens if two people call in sick on the same day?"
know whether it's working
details
how would i know
"How do I tell if a claims process is succeeding?" · "What should I track in a service workflow?" · "How do you evaluate whether a rota is working?" · "What are the signs that a process is failing?" · "What queue length should trigger intervention?" · "What would tell me this is about to fall over?"
iv. what does ops do?
Eleven things, and most of them are not what people expect from a workflow tool.
details
follows a process end to end
- Traces work from what starts it through to what comes out, naming the people, systems, materials and decisions involved at each stage.
traces a delay back to where it starts
- Separates where a problem becomes visible from where it actually began, which are frequently several steps apart.
finds what is limiting the whole thing
- Finds the one step the whole thing is waiting on, and says what would actually move if you fixed it.
says what a step is waiting on
- Names the approval, delivery, decision, result or resource that has to arrive before work can move, and what happens while it doesn't.
examines where work changes hands
- Traces what passes between people, teams or systems, what has to travel with it, and what gets dropped when it doesn't.
works out who should own what
- Separates who does the work, who decides, who approves, who supports and who remains accountable, and shows where that is currently unclear.
sets up a flow that doesn't exist yet
- Designs the sequence, the owners, the handoffs, the checkpoints and the capacity needed to get from a starting point to a required result.
tests what happens under pressure
- Works out what breaks first when volume rises, staff are missing, a supplier fails or conditions change, and where the problem moves to next.
compares two ways of doing the same work
- Explains how two structures differ, what each is good at, and what each one costs.
identifies what to watch
- Says what would reveal whether a process is working, what to measure, and what should trigger attention before a problem becomes a crisis.
explains the vocabulary from inside the example
- Names what something is called once you can see what it refers to, rather than the other way round.
v. what does ops return?
The answer follows the question. How something works gets you the sequence. A problem gets a diagnosis. A flow you need built gets built. What to watch gets the measures. Short question, short answer.
details
Depending on what you ask, Ops can:
- walk you through the sequence, stage by stage, in the order it actually happens;
- name who does each part, and who is accountable when it goes wrong;
- show where work changes hands and what has to travel with it;
- explain what each stage is waiting on before it can proceed;
- show where decisions send the work down different paths, including the exceptions;
- trace a delay or a failure back to where it started;
- identify where work piles up and what is limiting the whole thing;
- set up a flow from scratch, with owners, checkpoints and a finish line;
- compare two ways of organizing the same work and what each one costs;
- say what to measure, and what would warn you early;
- work out what breaks first as volume grows.
Not every answer contains all of this. The question determines what is relevant.
vi. what does ops know?
The reasoning underneath the answers. Why the slow step is usually innocent, why waiting costs more than working, and what actually breaks a process.
details
the step that looks slow is usually innocent
Where a problem becomes visible and where it began are usually different places. A stage that looks slow may be waiting on something that hasn't arrived, receiving incomplete work, or absorbing the effects of a constraint several steps upstream. Diagnosis follows the path backward before anything at the visible stage gets changed.
most of the elapsed time is waiting, not working
A task needing fifteen minutes of actual work can spend three days in queues, approvals and handovers. Measuring how long the work takes misses almost all of the delay, because the cost is hiding in the gaps between people rather than inside their fifteen minutes.
a bottleneck is whatever sets the pace for everything else
A bottleneck is defined by what comes out of the system, not by which task feels slowest. Speeding up a stage that isn't the constraint improves that stage and changes nothing about the final result — the extra work simply piles up in front of whatever the real limit is.
a handover has to carry three things
Work, context and responsibility have to move together. Lose any one of the three and the transfer fails even when everyone involved does their own job correctly. Processes break at the seams far more often than they break at the stages.
capacity on paper is not capacity in practice
Demand does not arrive evenly. Two people who can each handle twenty an hour have capacity for forty, so thirty-eight looks comfortable — but arrivals cluster, and at high utilization there is no slack to absorb a cluster before the next one lands. Waiting times rise sharply rather than gradually. The same arithmetic governs an emergency room, a support queue and a motorway.
automating an uncorrected process makes it fail faster
Automation preserves whatever it automates. An unnecessary approval, a duplicated task or a badly designed handover runs at higher speed and produces the same structural problem more often. Fixing the process and automating it are two separate decisions, in that order.
growth removes the invisible coordination small teams run on
A small operation runs on memory, informal conversation and people noticing what needs doing. That shared context disappears as the operation grows, which is why ownership, explicit triggers, standard handovers and visible status start to matter at a size where they previously didn't. Scale changes the nature of the coordination problem, not just its size.
exceptions are part of the system
A process designed only for the standard path is incomplete. Missing information, failed checks, unusual cases, unavailable resources and rejected approvals all need defined routes of their own, or they become the work that falls outside the process and gets handled by whoever happens to notice.
the written process and the real one drift apart
A shortcut gets taken once under pressure. Nothing goes wrong, so it gets taken again. Eventually the version being lived is a different system from the one written down, and nobody experienced that as a decision. Honest analysis starts from how work actually moves rather than how the manual says it should.
every piece of work should have a known next state
At any point, work should have a next action, an owner, something it is waiting on, or a finish line. Ambiguity about what happens next is not a gap in documentation — it is an operational defect, and it is where things quietly stall.
improving one part can damage the whole
A warehouse that speeds up picking overwhelms packing. A farm that doubles harvest speed overwhelms storage sized for the old pace. A call centre that shortens call time pushes more repeat calls onto escalations. Nobody did anything wrong in isolation; the problem only appears when the system is judged as a whole.
efficiency and resilience pull against each other
A system squeezed for maximum efficiency has no slack, and a system with no slack has no room to absorb a surprise. Abnormal days are not rare — they are unpredictable. The real question is how much of one an operation is trading for the other, deliberately, given what it costs on the day it fails.