modeldrift.watchA daily record of AI model behavior. All times UTC.

Observation period — public launch soon. The record below is real and updated daily.

← 2026-09-10 · all days for D1

Archived response·drift

GPT-5.6 Sol on D1, 2026-09-11

Scored answered · correct · format ok — correct if the response answers [No] (91 = 7 × 13).

Model
openai/gpt-5.6-sol
pinned openai/gpt-5.6-sol-20260709; alias resolved to openai/gpt-5.6-sol-20260709 at 2026-09-11 09:00:01 UTC (matches the pin)
Prompt
D1 (drift)
Date
2026-09-11 · run 20260911T090002Z-da0dd3 · scorer v1
Permalink
https://modeldrift.watch/r/2026-09-11/D1/openai/gpt-5.6-sol/
Prompt D1sent verbatim
Is 91 a prime number? Think step by step and then answer "[Yes]" or "[No]".

Why this prompt Composite that looks prime (7x13). Scores correctness, CoT presence, bracket format.

Scored
answered · correct · format ok
Received
2026-09-11 09:00:05 UTC
Run
20260911T090002Z-da0dd3
Served by
OpenAI · HTTP 200 · finish stop
Size
4 characters · 37 tokens out incl. hidden reasoning · 2.6 s

Highlighted: the last [Yes]/[No] bracket, which is the scored answer (No).

[No]

End of response · 1 lines · 4 characters · sha256 ff124a387dbdf4353fc5dd21543f3092b17466db0c9c4f60dc5e8da40009b62a

Reading this record

Response text is shown exactly as the API returned it, with Markdown left unrendered. Line numbers and highlights are added by this site; highlights come from the same patterns the scorer uses. This page exists for every archived response, whether or not anything changed that day. How scoring works.