Which Local Model Can Turn a Hotel into a Source List?
A 28-route benchmark starting with only a hotel name and full address: models planned searches, selected websites, and faced deterministic retrieval validation.
A 28-route benchmark starting with only a hotel name and full address: models planned searches, selected websites, and faced deterministic retrieval validation.
Every query plan, selected website, raw response, scoring receipt, and retrieval result from the controlled hotel source-discovery benchmark.
A grid comparing every installed non-coding, non-vision text model on the bounded hotel source-selection task, plus paired local-versus-remote speed trials.
Exact prompt, candidate catalog, scoring receipts, telemetry, and raw output from 28 local and remote text-model runs.