Type: AI model or model family
GPT-5.1 appears in the radar stream as a model or model family. This page is a living index of dated mentions and sources — open Recent Updates and Backlinks for context, not a full product brief.
Recent Updates
- 2026-07-01: Calibration, Not Compilation: Evaluating LLM-Written Probabilistic Programs (cs.LG updates on arXiv.org) · arxiv.org — LLM Evals Claude
- 2026-07-04: Agentic coding can fabricate convincing but false repros (Hacker News) · danluu.com — Code Agents LLM Evals GPT-5.0
- 2026-07-08: Opinionated notes on AI-assisted testing and agentic coding (Hacker News) · danluu.com — GPT-5.0
- 2026-07-09: Institutional red-teaming tests how deployment rules change multi-agent safety (cs.AI updates on arXiv.org) · arxiv.org — Agents LLM Evals
- 2026-07-11: OmniFood-Bench evaluates VLM nutrient reasoning and health advice (cs.AI updates on arXiv.org) · arxiv.org — LLM Evals Gemini 3 Flash · Qwen3-VL-8B
- 2026-07-13: Token pricing is not comparable without tokenizer counts (Hacker News) · playcode.io — Anthropic · OpenAI · Google · xAI · Claude Opus 4.6 · Claude Opus 4.8 · GPT 5.5 · GPT-5.6 Sol · Grok
- 2026-07-14: Agentic context learning appears to depend on specification acquisition, not just retrieval (cs.AI updates on arXiv.org) · arxiv.org — Context Engineering LLM Evals Long Context Qwen3.5-27B · Gemini 3 Pro
- 2026-07-15: Habr experiment: benchmarking Karpathy-style LLM-wiki against simple vector RAG (Все статьи подряд / Искусственный интеллект / Хабр) · habr.com — RAG RAG Evaluation LLM Evals Agents Code Agents Chunking Embeddings Andrej Karpathy Gemma4-31B
- 2026-07-19: Last Week in AI: China, Compression, and the Open-Model Race (TheSequence) · thesequence.substack.com — Open Source LLMs Long Context LLM Evals TheSequence · Thinking Machines Lab · Moonshot AI · PrismML · OpenAI · NVIDIA · Google Shanghai World AI Conference Xi Jinping · Inkling · Kimi K3 · Bonsai 27B GPT-Red · GPT-5.6
- 2026-07-22: BatchDAG Plans and Executes Analytical Work as a Typed DAG (cs.AI updates on arXiv.org) · arxiv.org — Agents Tool Use Context Engineering LLM Evals
- 2026-07-26: A personal report on agentic coding, testing, and model variance (Hacker News) · danluu.com — Code Agents LLM Evals OpenAI · Mastodon · Playwright · GPT-5.0
Incident radar
GROUNDING’s Hallucination Incident Index has flagged 1 incident naming GPT-5.1 — a heuristic proxy over the AI-news radar’s daily journal (hallucination/jailbreak/refusal/bias keyword matches), not a verified incident registry.
- Severity: S3 1
- Category: Hallucination 1
Most recent:
- A personal report on agentic coding, testing, and model variance source (2026-07-26)
Full breakdown: Hallucination Incident Index · Subscribe: incident-index RSS feed
FAQ
What is GPT-5.1?
GPT-5.1 appears in the radar stream as a model or model family. This page is a living index of dated mentions and sources — open Recent Updates and Backlinks for context, not a full product brief.
What does this page track?
Dated radar mentions, source links, related concepts, and builder-relevant context for GPT-5.1, collected automatically by GROUNDING.
When was GPT-5.1 last mentioned?
GPT-5.1 was most recently mentioned in a radar update dated 2026-07-26.
Category: Text / Language Models