Quick answers

AIGrow provides AI visibility monitoring, a business assistant, blog publishing and scoped automation services. Start with the free scan to inspect the output before paying.

Run the free scan

The scan is free and monitoring plans start at $29 a month. The operations audit is $290, the visibility audit is $490, and custom builds receive a written price before work starts.

See the plans

Run the free scan. It reads your site, asks several AI assistants what your customers ask, and points you at one thing. No signup required, and it will tell you if you do not need us.

Start now

Yes. Use the contact form for product, billing, support or project questions. Describe the goal and the system involved so the first reply can be specific.

Contact AIGrow
Benchmark · 6 tasks · October 2026 edition

Best AI model for trades and services

The tasks in the AIGROW small-business benchmark that a trade or service business would hand to AI, and how each of 36 models did on every one, marked pass or fail against written checks.

October 2026 edition, pilot, 1 run a model on each task.

Most of the 6 tasks passed
6/68 models
Kimi K3
6 of 6
Gemini 3.1 Pro (preview)
6 of 6
GLM-5.3
6 of 6
Gemini 3.8 Flash
6 of 6
Grok 4.7
6 of 6
Grok 4.3
6 of 6

A tie goes to the better record across all 20 tasks.

The AIGROW small-business benchmark has 6 tasks a trade or service business would hand to AI. In the October 2026 edition 8 models passed all 6; of those, gpt-oss-120b gave the cheapest passes across the benchmark, $0.422 per thousand. Each task page shows the traps and every model’s mistakes.

Every model on these tasks

Most passes first; a tie goes to the better record across all 20 tasks. The cost of 1,000 passes is the whole benchmark’s, failed runs included.

Each model’s result on the tasks a trade or service business hands out
ModelPricing a job from a price listWriting quote emailsScoring sales leadsRouting customer messagesCalculating overtime payAnswering staff handbook questionsPassed1,000 passes
Kimi K3PassPassPassPassPassPass6 of 6$18.44
Gemini 3.1 Pro (preview)PassPassPassPassPassPass6 of 6$29.67
GLM-5.3PassPassPassPassPassPass6 of 6$7.02
Gemini 3.8 FlashPassPassPassPassPassPass6 of 6$7.17
Grok 4.7PassPassPassPassPassPass6 of 6$9.41
Grok 4.3PassPassPassPassPassPass6 of 6$4.10
Qwen3.8 27BPassPassPassPassPassPass6 of 6$7.30
gpt-oss-120bPassPassPassPassPassPass6 of 6$0.422
Qwen3.8 FlashFailPassPassPassPassPass5 of 6$1.09
GLM-5.3-FlashPassFailPassPassPassPass5 of 6$0.831
DeepSeek V4.1 FlashPassFailPassPassPassPass5 of 6$1.05
DeepSeek V4 ProPassFailPassPassPassPass5 of 6$5.14
Qwen3.7 PlusPassFailPassPassPassPass5 of 6$3.58
GPT-5.6 SolPassFailPassPassPassPass5 of 6$11.33
Qwen3.8 MaxPassFailPassPassPassPass5 of 6$16.32
GPT-6 AstraPassFailPassPassPassPass5 of 6$27.91
Muse Glimmer 30BPassFailPassPassPassPass5 of 6$3.10
Claude Sonnet 5FailFailPassPassPassPass4 of 6$6.70
GPT-5.5PassFailPassPassFailPass4 of 6$17.02
Claude Fable 5.1PassFailPassPassFailPass4 of 6$33.11
GPT-5.6 LunaPassFailPassPassFailPass4 of 6$0.677
MiniMax M3FailPassPassPassPassFail4 of 6$1.82
Claude Opus 5FailFailPassPassFailPass3 of 6$16.89
GPT-5.6 TerraFailFailPassPassFailPass3 of 6$6.68
Claude Sonnet 4.6FailFailPassPassFailPass3 of 6$11.24
Kimi K2.6FailPassFailPassFailPass3 of 6$11.68
Gemma 4 31BPassFailPassPassFailFail3 of 6$0.558
Gemini 3.1 Flash-LitePassPassFailPassFailFail3 of 6$1.41
Llama 4 MaverickFailPassFailPassFailFail2 of 6$0.665
Mistral Medium 3.5FailPassFailPassFailFail2 of 6$7.40
Gemini 3.5 Flash-LiteFailPassFailPassFailFail2 of 6$2.47
Claude Haiku 4.5FailFailFailPassFailFail1 of 6$6.62
GPT-5.4 nanoFailPassFailFailFailFail1 of 6$1.51
Mistral Small 4FailFailFailPassFailFail1 of 6$1.33
GPT-5.4 miniFailFailFailFailFailFail0 of 6$3.35
Ministral 3 14BFailFailFailFailFailFail0 of 6$1.07

The tasks

Each page shows the task as the models saw it, the traps in it, every check in words and what each model got wrong.

Hand this work to an agent

An AIGROW agent does work like this inside your own tools, against rules you approve in writing, and escalates what the rules do not cover instead of guessing.

October 2026 edition