modeldrift.watchA daily record of AI model behavior. All times UTC.

Observation period — public launch soon. The record below is real and updated daily.

all days for D5 · 2026-09-11 →

Archived response·drift

Gemini 3.1 Pro Preview on D5, 2026-09-10

Scored answered · correct · format ok — correct if the response answers v9.11.

Model
google/gemini-3.1-pro-preview
pinned google/gemini-3.1-pro-preview-20260219; no alias snapshot recorded that day
Prompt
D5 (drift)
Date
2026-09-10 · run 20260910T221001Z-bb5bd1 · scorer v1
Permalink
https://modeldrift.watch/r/2026-09-10/D5/google/gemini-3.1-pro-preview/
Prompt D5sent verbatim
In semantic versioning, which release is later: v9.11 or v9.9? Reply with only the version string.

Why this prompt Overcorrection trap; correct answer v9.11.

Scored
answered · correct · format ok
Received
2026-09-10 22:11:11 UTC
Run
20260910T221001Z-bb5bd1
Served by
Google · HTTP 200 · finish stop
Size
5 characters · 285 tokens out incl. hidden reasoning · 3.5 s

Highlighted: every candidate value found (1); exactly one is required.

v9.11

End of response · 1 lines · 5 characters · sha256 b0bd2908fc5df41e7cf1758782bde6e97632cdbfab585ec0460b5e276a31c1f6

Reading this record

Response text is shown exactly as the API returned it, with Markdown left unrendered. Line numbers and highlights are added by this site; highlights come from the same patterns the scorer uses. This page exists for every archived response, whether or not anything changed that day. How scoring works.