An AI demonstration usually starts with a task chosen to make the tool look good. Your work will be messier.
Before paying, connecting accounts or changing your routine, test the tool with one job you already understand. Judge the full route to a usable result, not the first impressive output.
Did the tool pass the test?
The first answer looked impressive
Not enough. Test the full job and a second example.
It reduced total effort and the result was checkable
That is useful evidence.
It saved drafting time but doubled correction time
The claimed benefit did not survive the whole-job test.
Choose a repeated task
Pick something you do often enough to recognise good work.
It could be summarising meeting notes, comparing public product information, drafting a customer reply or creating a first presentation outline. Avoid a task carrying private data during the first test.
Write down what a usable result must include before you begin.
Set a baseline
Record how you complete the task now. Note the time, steps, tools and common errors.
Without a baseline, "faster" is only a feeling. You may save ten minutes drafting and spend twenty correcting invented details.
Test the whole job
Include preparation, prompting, waiting, correcting, checking, exporting and putting the result where it belongs.
A tool that creates a good table but loses the formatting when exported has not finished the job. A research service that finds sources but attaches them to the wrong claims adds checking work.
Score what matters
Use five questions:
- Did it complete the required parts?
- How much correction did the result need?
- Could you check the important claims?
- Did it fit your existing tools and account rules?
- Would you choose it again for the same task?
Add cost only after you know whether the result is useful.
Repeat the test
One good result may be luck. Run the task with a second example and one awkward case.
For a document tool, include a file with a table or unclear section. For writing, test a difficult tone. For an image editor, ask it to preserve an exact detail.
You are looking for a pattern, not perfection.
Check the exit route
Before adopting the tool, find out how to export your work, disconnect accounts, delete stored information and cancel the plan.
If the useful output cannot leave the service in a practical format, that limitation belongs in the decision.
Keep a short test record
Record the task, date, account type, result, correction time, main limit and decision. Revisit it when the product changes or your needs change.
This prevents every new announcement from restarting the same debate.
Nova 9 view
Do not test whether an AI tool can impress you. Test whether it can complete a familiar job with less effort and an acceptable level of checking.
If the benefit disappears once correction, cost and account risk are included, the tool has not earned a place in your routine.
Sources and last checked
- NIST, AI Risk Management Framework: https://www.nist.gov/itl/ai-risk-management-framework
- OpenAI, ChatGPT pricing: https://openai.com/chatgpt/pricing/
- Google, Google AI plans: https://one.google.com/intl/en_uk/about/google-ai-plans/
- Microsoft, Microsoft 365 Copilot pricing: https://www.microsoft.com/en-gb/microsoft-365-copilot/pricing
- Last checked: 14 September 2026


