- How long does the assessment take, and what does it cost?
- Two weeks is the standard shape. It is a fixed-scope, fixed-price engagement, and the evaluation harness and thin slice are yours regardless of the recommendation — including when the recommendation is not to build it. We would rather be paid once for an honest no than three times for a slow yes.
- We already picked a model and a vendor. Is this still useful?
- Usually more useful, not less. Almost none of the seven criteria are about the model. Data access, error tolerance, workflow fit, evaluation, unit economics and accountability all sit outside the vendor decision, and they are where projects actually fail. If your model choice turns out to be wrong we will say so, but that is rarely the finding.
- Does this apply to agents, or only to retrieval and chat features?
- Both, and agents raise the stakes on every criterion. An agent taking actions rather than producing text turns error tolerance and accountability from paperwork into the core design problem — what it is allowed to do unsupervised, what requires a human, what gets logged, and how it is stopped. We evaluate agents against the same seven questions with a much harder line on the last two.
- What if we fail the assessment?
- There is no pass mark — there is a scored picture of where the risk actually sits. Most projects come out amber somewhere, and the useful output is knowing which two things to fix first. The three answers we listed above are the genuine stoppers, and in every case the honest move is to fix the underlying condition rather than build around it.
- Can you just build it? We're confident about the case.
- Yes — if you can answer the seven questions, we will happily skip the assessment and start building. We are not attached to selling a discovery phase. We are attached to not spending your budget on something the data was never going to support.