Hotel Step 3 Automation: Local Qwen vs Budget Cloud
A controlled Hampton benchmark compares local Qwen 3.8 with Alibaba Qwen 3.8 Flash and DeepSeek V4 Pro for reusable, hash-validated Step 3 hotel scoring.
A controlled Hampton benchmark compares local Qwen 3.8 with Alibaba Qwen 3.8 Flash and DeepSeek V4 Pro for reusable, hash-validated Step 3 hotel scoring.
A complete evidence table from a ten-hotel local-first extraction experiment: what the system looked for, the supported claim, exact source wording, criterion gaps, and telemetry.
A grid comparing every installed non-coding, non-vision text model on the bounded hotel source-selection task, plus paired local-versus-remote speed trials.
The exact first-wave local-model discovery cards for three hotel classes, including complete prompts, limits, dependencies, outputs, receipts, and stop conditions.
Experimental capability test report — this documents a non-production model workflow, including failures and incomplete telemetry; linked reviews may contain errors.