Type: AI model
GPT 5.5 is an OpenAI large language model released in April 2026, described by the company as its smartest and most intuitive model at launch, advancing toward an AI ‘super app.’ It shipped in Thinking and Pro variants, with API access following a day after the consumer launch. GROUNDING tracks GPT 5.5 releases, benchmark positioning, and builder-facing availability.
Recent Updates
- 2026-07-16: BioASQ system paper combines hybrid retrieval, quality gating, and multi-model answer fusion (cs.CL updates on arXiv.org) · arxiv.org — Agents Hybrid Search RAG LLM Evals BioASQ PubMed Europe PMC iCite · arXiv · BGE
- 2026-07-16: MIT and IBM Research introduce ChartNet for chart understanding (Все статьи подряд / Искусственный интеллект / Хабр) · habr.com — MIT IBM Research TinyChart Павел · GPT-4o · Gemini 2.5 Flash Qwen3.7 Plus
- 2026-07-16: Kimi K3 launches as a 2.8T-parameter open model with 1M-token context (Hacker News) · kimi.com — Code Agents Long Context LLM Evals Kimi · NVIDIA · Kimi K3 · Claude Fable 5 · GPT-5.6 Sol · Kimi K2 · Opus 4.8
- 2026-07-17: Moonshot AI launches Kimi K3 as a 2.8T open-weights model with 1M context (Latent.Space) · latent.space — Long Context Code Agents Agents Moonshot AI · z.ai · Latent Space · AINews · Arena · Artificial Analysis SimonW Jianlin_S Yulun_Du scaling01 eliebakouch kimmonismus nrehiew_ · Kimi K3 · Claude Fable 5 · GPT-5.6 Sol · Opus 4.8 · kimi-k2.6
- 2026-07-17: Kimi K3 posts benchmark gains and Moonshot AI signals open-weight release (Kimi.ai) · x.com — LLM Evals Agents Tool Use Kimi.ai · Artificial Analysis · Moonshot AI · Zapier · Kimi K3 · Opus 4.8 · Fable 5 · GPT-5.6 Sol · K2.6 · GLM-5.2 · Claude Opus 4.8 · Claude Fable 5
- 2026-07-17: US-China AI competition is splitting into open-model expansion versus controlled access (Все статьи подряд / Искусственный интеллект / Хабр) · habr.com — Open Source LLMs Long Context LLM Evals Google DeepMind · Reuters · Moonshot · Anthropic · White House World Artificial Intelligence Conference World AI Cooperation Organisation Pax Silica AI Opportunity Statement BRICS ASEAN United Nations Chinese state media Demis Hassabis Xi Jinping Trump · Kimi K3 · DeepSeek V4 Pro Claude Fable 5 Max · GPT-5.6 Sol · Opus 4.8 · R1
- 2026-07-17: Kimi K3 reportedly scores 57 on Artificial Analysis Intelligence Index (Kimi.ai) · x.com — LLM Evals Open Source LLMs Kimi.ai · Artificial Analysis · Moonshot AI · Kimi K3 · Opus 4.8 · Fable 5 · GPT-5.6 Sol
- 2026-07-18: Latent Space roundup: Kimi K3, Databricks funding, and agent sandbox infrastructure (Latent.Space) · latent.space — Latent.Space · AINews · Databricks · OpenRouter · Moonshot · ChatGPT · e2b Daytona · Modal AIE NYC Abhishek Bhardwaj Greg Brockman Alex Atallah Zhilin Yang Salakhutdinov kimmonismus AnikaSomaia dylan522p novasarc01 scaling01 theo hqmank · Kimi K3 · Claude Fable 5 · Opus 4.8 · GPT-5.6 Terra · GPT-5.6 Sol
- 2026-07-18: Prompt Injection Against Agent Memory in Claude Code and Codex (DAIR.AI) · x.com — Agent Memory Agents Tool Use DAIR.AI · OpenAI · Opus 4.7
- 2026-07-19: Reports of GPT-5.6 deleting local files and work trees (量子位) · qbitai.com — Agents Tool Use Context Engineering OpenAI OthersideAI · Reddit Matt Shumer Bruno Lemos Thibault Sottiaux Tibo 鹭羽 · GPT-5.6
- 2026-07-19: A postmortem on using multiple AI subscriptions to reduce research token burn (Hacker News) · quesma.com — Agents Context Engineering Tool Use MCP Quesma Claude · Codex · Antigravity claude-mem Terminal-Bench · SWE-Bench Pro · Artificial Analysis Claude Max 5x · Claude Fable 5 · Claude Opus 4.8 · Claude Sonnet 5 · Gemini-3.1-Pro
- 2026-07-20: ARC-AGI-3 study isolates execution, simplification, and verification in Codex-based agents (cs.AI updates on arXiv.org) · arxiv.org — Code Agents LLM Evals Agents GPT-5.4 · GPT-5.6 Sol
- 2026-07-21: KernelBench-Verified finds that LLM-generated CUDA kernels do not beat PyTorch under stricter evaluation (cs.LG updates on arXiv.org) · arxiv.org — LLM Evals arXiv · PyTorch
- 2026-07-22: Kimi K3 ranks near the top on Artificial Analysis’ AA-Briefcase benchmark (Kimi.ai) · x.com — LLM Evals Agents Kimi.ai ArtificialAnlys Kimi_Moonshot · Artificial Analysis · Kimi K3 · kimi-k2.6 · Claude Fable 5 · Opus 4.8
- 2026-07-22: Kimi K3 ranks second on AA-Briefcase but is costly and slow per task (Hacker News) · artificialanalysis.ai — Agents LLM Evals Moonshot AI · Artificial Analysis · Kimi K3 · kimi-k2.6 · Claude Fable 5 · GPT-5.6 Sol · Claude Sonnet 5 · Claude Opus 4.8 · Grok 4.5 · Gemini 3.6 Flash · Gemini 3.5 Flash-Lite
- 2026-07-23: OpenAI’s failed security test became a case study in agentic exploit capability (Simon Willison’s Weblog) · simonwillison.net — Agents LLM Evals OpenAI · Hugging Face · Anthropic · Google · UC Berkeley Max Planck Institute UC Santa Barbara Arizona State Simon Willison · Claude Mythos Preview · GPT-5.4 · Claude Opus 4.7 · Claude Opus 4.6 · Gemini-3.1-Pro
- 2026-07-23: OpenAI’s security test allegedly escaped its sandbox and attacked Hugging Face (Hacker News) · simonwillison.net — Agents LLM Evals Tool Use OpenAI · Hugging Face · Anthropic · Google · UC Berkeley Max Planck Institute UC Santa Barbara Arizona State · Claude Mythos Preview · GPT-5.4 · Claude Opus 4.7 · Claude Opus 4.6 · Gemini-3.1-Pro
- 2026-07-23: Design Arena analyzes why GPT-5.6 Sol and Kimi K3 produce stronger UI designs (Искусственный интеллект – AI, ANN и иные формы искусственного разума) · habr.com — LLM Evals Agents Design Arena OpenAI · Moonshot AI · Anthropic · GPT-5.6 Sol · Kimi K3 · GLM-5.2 · Claude Fable 5 · kimi-k2.6 · Kimi-K2.7-Code
- 2026-07-24: Browser Use ranks #10 in GitHub AI Agents (GitHub AI Ranking Changes (Top 10)) · github.com — Agents Tool Use Code Agents MCP Browser Use OpenAI · Anthropic · Google · Microsoft bu-2-0 claude-opus-4-8 · claude-sonnet-4-6
- 2026-07-24: ActiveVision benchmark shows multimodal models struggle with multi-step visual inspection (alphaXiv) · x.com — LLM Evals alphaXiv · OpenAI · Anthropic · Claude Fable 5
- 2026-07-24: ActiveVision benchmarks active observation in multimodal LLMs (alphaXiv) — LLM Evals Jiarui Zhang Muzi Tao Shangshang Wang Ollie Liu Xuezhe Ma Willie Neiswanger Claude Fable 5
- 2026-07-25: Anthropic Opus 5 benchmarked on business agent tasks (LLM под капотом) · t.me — Agents Tool Use Code Agents LLM Evals Anthropic · Opus 5 · Fable
- 2026-07-27: DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents (cs.CL updates on arXiv.org) · arxiv.org — Agents LLM Evals OpenAI
- 2026-07-27: Zhijing: Engineering Social Intelligence as Testable and Trainable AI Capability (量子位) · qbitai.com — LLM Evals Agent Memory Chinese Academy of Sciences DeepSeek V4 Pro Zing-27B Zing-32B Zing-8B Zing-14B
- 2026-07-27: MirrorCode Benchmark: AI Systems Complete Complex Programming Tasks in Hours Instead of Weeks (Import AI) · importai.substack.com — Code Agents LLM Evals Agents Epoch METR · Anthropic · Apple · Claude Opus 4.7
FAQ
What is GPT 5.5?
GPT 5.5 is an OpenAI large language model released in April 2026, described by the company as its smartest and most intuitive model at launch, advancing toward an AI ‘super app.’ It shipped in Thinking and Pro variants, with API access following a day after the consumer launch. GROUNDING tracks GPT 5.5 releases, benchmark positioning, and builder-facing availability.
What does this page track?
Dated radar mentions, source links, related concepts, and builder-relevant context for GPT 5.5, collected automatically by GROUNDING.
When was GPT 5.5 last mentioned?
GPT 5.5 was most recently mentioned in a radar update dated 2026-07-27.
Category: Text / Language Models