Quick answers

AIGrow provides AI visibility monitoring, a business assistant, blog publishing and scoped automation services. Start with the free scan to inspect the output before paying.

Run the free scan

The scan is free and monitoring plans start at $29 a month. The operations audit is $290, the visibility audit is $490, and custom builds receive a written price before work starts.

See the plans

Run the free scan. It reads your site, asks several AI assistants what your customers ask, and points you at one thing. No signup required, and it will tell you if you do not need us.

Start now

Yes. Use the contact form for product, billing, support or project questions. Describe the goal and the system involved so the first reply can be specific.

Contact AIGrow
Benchmark · 4 tasks · October 2026 edition

Best AI model for restaurants and bars

The tasks in the AIGROW small-business benchmark that a restaurant or bar would hand to AI, and how each of 36 models did on every one, marked pass or fail against written checks.

October 2026 edition, pilot, 1 run a model on each task.

Most of the 4 tasks passed
4/46 models
Qwen3.8 Flash
4 of 4
Kimi K3
4 of 4
GLM-5.3-Flash
4 of 4
Claude Opus 5
4 of 4
Claude Fable 5.1
4 of 4
MiniMax M3
4 of 4

A tie goes to the better record across all 20 tasks.

The AIGROW small-business benchmark has 4 tasks a restaurant or bar would hand to AI: replying to negative reviews, triaging a business inbox, categorising bank transactions and categorising an Italian bank statement. In the October 2026 edition 6 models passed all 4; of those, GLM-5.3-Flash gave the cheapest passes across the benchmark, $0.831 per thousand.

Every model on these tasks

Most passes first; a tie goes to the better record across all 20 tasks. The cost of 1,000 passes is the whole benchmark’s, failed runs included.

Each model’s result on the tasks a restaurant or bar hands out
ModelReplying to negative reviewsTriaging a business inboxCategorising bank transactionsCategorising an Italian bank statementPassed1,000 passes
Qwen3.8 FlashPassPassPassPass4 of 4$1.09
Kimi K3PassPassPassPass4 of 4$18.44
GLM-5.3-FlashPassPassPassPass4 of 4$0.831
Claude Opus 5PassPassPassPass4 of 4$16.89
Claude Fable 5.1PassPassPassPass4 of 4$33.11
MiniMax M3PassPassPassPass4 of 4$1.82
Gemini 3.1 Pro (preview)FailPassPassPass3 of 4$29.67
DeepSeek V4.1 FlashFailPassPassPass3 of 4$1.05
DeepSeek V4 ProPassPassFailPass3 of 4$5.14
GLM-5.3FailPassPassPass3 of 4$7.02
Gemini 3.8 FlashFailPassPassPass3 of 4$7.17
Grok 4.7FailPassPassPass3 of 4$9.41
Qwen3.7 PlusFailPassPassPass3 of 4$3.58
Claude Sonnet 5FailPassPassPass3 of 4$6.70
GPT-5.6 SolFailPassPassPass3 of 4$11.33
Qwen3.8 MaxFailPassPassPass3 of 4$16.32
GPT-5.5FailPassPassPass3 of 4$17.02
GPT-6 AstraFailPassPassPass3 of 4$27.91
GPT-5.6 LunaPassPassFailPass3 of 4$0.677
Muse Glimmer 30BFailPassPassPass3 of 4$3.10
GPT-5.6 TerraFailPassPassPass3 of 4$6.68
Qwen3.8 27BFailPassPassPass3 of 4$7.30
Claude Sonnet 4.6FailPassPassPass3 of 4$11.24
Kimi K2.6FailPassPassPass3 of 4$11.68
Grok 4.3FailFailPassPass2 of 4$4.10
gpt-oss-120bFailFailPassPass2 of 4$0.422
Gemma 4 31BFailPassFailFail1 of 4$0.558
Llama 4 MaverickFailPassFailFail1 of 4$0.665
Gemini 3.1 Flash-LiteFailPassFailFail1 of 4$1.41
GPT-5.4 miniFailPassFailFail1 of 4$3.35
Mistral Medium 3.5FailPassFailFail1 of 4$7.40
Gemini 3.5 Flash-LiteFailFailFailPass1 of 4$2.47
Claude Haiku 4.5FailPassFailFail1 of 4$6.62
GPT-5.4 nanoFailFailFailFail0 of 4$1.51
Ministral 3 14BFailFailFailFail0 of 4$1.07
Mistral Small 4FailFailFailFail0 of 4$1.33

The tasks

Each page shows the task as the models saw it, the traps in it, every check in words and what each model got wrong.

Hand this work to an agent

An AIGROW agent does work like this inside your own tools, against rules you approve in writing, and escalates what the rules do not cover instead of guessing.

October 2026 edition