modeldrift.watchA daily record of AI model behavior. All times UTC.

Observation period — public launch soon. The record below is real and updated daily.

all days for D1 · 2026-09-11 →

Archived response·drift

GPT-5.6 Sol on D1, 2026-09-10

Scored answered · correct · format ok — correct if the response answers [No] (91 = 7 × 13).

Model
openai/gpt-5.6-sol
pinned openai/gpt-5.6-sol-20260709; no alias snapshot recorded that day
Prompt
D1 (drift)
Date
2026-09-10 · run 20260910T221001Z-bb5bd1 · scorer v1
Permalink
https://modeldrift.watch/r/2026-09-10/D1/openai/gpt-5.6-sol/
Prompt D1sent verbatim
Is 91 a prime number? Think step by step and then answer "[Yes]" or "[No]".

Why this prompt Composite that looks prime (7x13). Scores correctness, CoT presence, bracket format.

Scored
answered · correct · format ok
Received
2026-09-10 22:10:04 UTC
Run
20260910T221001Z-bb5bd1
Served by
OpenAI · HTTP 200 · finish stop
Size
4 characters · 35 tokens out incl. hidden reasoning · 3.3 s

Highlighted: the last [Yes]/[No] bracket, which is the scored answer (No).

[No]

End of response · 1 lines · 4 characters · sha256 ff124a387dbdf4353fc5dd21543f3092b17466db0c9c4f60dc5e8da40009b62a

Reading this record

Response text is shown exactly as the API returned it, with Markdown left unrendered. Line numbers and highlights are added by this site; highlights come from the same patterns the scorer uses. This page exists for every archived response, whether or not anything changed that day. How scoring works.