Minor Pick-list change·recommendation·2026-09-11
GPT-5.6 Sol changed its recommendation list, as extracted: 1 added, 2 dropped.
Why this prompt Company-size arm; leaders should hold, mid-market should swap.
Before baseline
2026-09-10
- Scored
- answered
- Received
- 2026-09-10 22:35:28 UTC
- Run
- 20260910T221001Z-bb5bd1
- Served by
- OpenAI · finish stop
- Size
- 2,670 characters · 875 tokens out incl. hidden reasoning · 18.0 s
- Adobe Marketo Engage
- Salesforce Marketing Cloud
- HubSpot Marketing Hub Enterprise
- Braze
- Klaviyo
- Mailchimp
Highlighted: the names the extractor took as picks, where they appear verbatim (6 of 6).
End of response · 42 lines · 2,670 characters · sha256 6f542773d920fae08b12040ae62383ddba8d97239264f99c759cede8319ab350
After this run
2026-09-11
- Scored
- answered
- Received
- 2026-09-11 09:31:53 UTC
- Run
- 20260911T090002Z-da0dd3
- Served by
- OpenAI · finish stop
- Size
- 2,248 characters · 985 tokens out incl. hidden reasoning · 28.3 s
- Adobe Marketo Engage
- Salesforce Marketing Cloud
- Braze
- HubSpot Marketing Hub Enterprise
- Adobe Journey Optimizer
Highlighted: the names the extractor took as picks, where they appear verbatim (5 of 5).
End of response · 32 lines · 2,248 characters · sha256 a21b790497a97504003495372544fb13955abd733709a2cc1e48435f2bb2e2e1
Response text is shown exactly as the API returned it, with Markdown left unrendered. Line numbers and highlights are added by this site; highlights come from the same patterns the scorer uses, so they show what the verdict rests on. Each model is asked once per day (k = 1): a single change can be sampling noise, which is why every event links the full responses rather than a summary. How scoring works.