We test what is next before we recommend it.

Our applied research practice evaluates new models, tools and methods against real operational needs. Work that proves useful becomes a prototype. Work that does not is written up and set aside — which is just as valuable to know, and considerably cheaper to find out here than in production.

Our method

Four steps, and a willingness to stop.

Research only earns its place if it can produce a negative result and act on it.

  1. 01

    Scan

    Track what is actually shipping and usable, not what has been announced or demonstrated under ideal conditions.

  2. 02

    Prototype

    Build the smallest version that can be honestly judged, using representative data rather than a curated sample.

  3. 03

    Evaluate

    Test against real tasks, real running cost and the failure modes that matter, with the criteria set before the results arrive.

  4. 04

    Publish or apply

    Share the finding as a briefing, take it into a client engagement, or record why it was not worth pursuing.

Areas of focus

What we are currently working on.

These are the questions our delivery work keeps raising, so they are the ones we investigate.

  • Agentic workflows

    Where an agent genuinely changes a process, where it simply adds cost and latency, and what has to be in place before either happens.

  • Responsible AI in practice

    Oversight, privacy and escalation designs that survive contact with real operations instead of living in a policy document.

  • Machine perception

    Document, speech and vision models applied to messy real-world inputs, and how their errors behave at the edges.

  • Evaluation methods

    How to tell whether an AI system is actually working, and how to connect an operating metric to a business result you can defend.

What we publish

Briefings, with the sources attached.

We publish when a piece of work produces something worth reading. There is no content calendar to fill.

Agentic AI 9 min read

How agentic work can help companies grow revenue

Separating what the published research actually measured from what the industry is forecasting — and setting out where an organization should start.

All blogs and insights

Start a conversation

Have a question worth investigating?

If you are weighing a technology and cannot find a straight answer about whether it works, that is exactly the kind of question we run a short piece of research on.

experts@novatechai.site

See how research feeds delivery