Skip to content

Type: AI model

GPT-OSS 120B is the larger open-weight model in OpenAI’s gpt-oss line, distributed on Hugging Face for self-hosted deployment. It pairs with the smaller gpt-oss-20b variant. GROUNDING tracks GPT-OSS 120B’s benchmarks and use in open-model serving.

Recent Updates

  • 2026-07-01: InterFLOPBench benchmarks LLMs on floating-point error classification in code (cs.AI updates on arXiv.org) · arxiv.orgLLM Evals arXiv CCSD proxy · Connected Papers · Litmaps · scite · alphaXiv · CatalyzeX · DagsHub · Gotit.pub · Hugging Face · ScienceCast · CORE · arXivLabs Lisa Taldir Qwen 3 32b · Gemini 2.5 Flash Phi-4-Reasoning DeepSeek R1T2 · gpt-oss-20b
  • 2026-07-02: GRACE-RAG proposes a graph-augmented retrieval architecture for institutional QA (cs.AI updates on arXiv.org) · arxiv.orgRAG arXiv · Mistral · OpenAI · Google · Hugging Face · alphaXiv · CatalyzeX · DagsHub · Gotit.pub · ScienceCast · Connected Papers · Litmaps · scite · CORE Mistral 24B · Gemini 2.5 Flash
  • 2026-07-07: LLMs show mechanism-level routing failures on Lean-verified algebraic structures (cs.CL updates on arXiv.org) · arxiv.orgLLM Evals Manuel Israel Cazares Llama-3.3-70b
  • 2026-07-07: HASE claims to co-evolve task solutions and the evaluation harness in one agentic loop (cs.AI updates on arXiv.org) · arxiv.orgAgents LLM Evals Qwen3-8B
  • 2026-07-09: LiveOIBench introduces a competitive programming benchmark for evaluating LLMs (cs.CL updates on arXiv.org) · arxiv.orgLLM Evals GPT-5
  • 2026-07-09: SmartHomeSecure uses constrained LLM repair for Home Assistant YAML errors (cs.AI updates on arXiv.org) · arxiv.orgLLM Evals home-assistant · gpt-oss-20b · Llama-3.1-8B · Llama-3.3-70b
  • 2026-07-10: Hands-On with the AMD Ryzen AI Halo (Hacker News) · microcenter.comLong Context AMD Micro Center Dan Ackerman Jacob Bobo Qwen3-Coder-30B · Gemma 4 31B Z-Image-Turbo
  • 2026-07-11: AegisDx proposes a safety-oriented framework for AI-assisted differential diagnosis (cs.AI updates on arXiv.org) · arxiv.orgAgents Tool Use RAG Context Engineering Yale New Haven Health System GPT-5
  • 2026-07-13: AutoWorldBuilder paper on multi-agent worldbuilding with context compression and iterative review (cs.AI updates on arXiv.org) · arxiv.orgAgents Context Engineering LLM Evals DeepSeek-V3.2
  • 2026-07-14: Benchmarking LLM Legal Citation Fabrication in GDPR and Saudi PDPL (cs.CL updates on arXiv.org) · arxiv.orgLLM Evals arXiv · Gemini 2.5 Flash Nemotron-3-Super-120B
  • 2026-07-16: GSM-Plus-BN benchmarks Bangla mathematical reasoning in LLMs (cs.CL updates on arXiv.org) · arxiv.orgLLM Evals Open Source LLMs Qwen3-32B Llama-3.1-8b-instant llama-3.3-70b-versatile Llama-4-Scout-17B-16E-Instruct · gpt-oss-20b
  • 2026-07-16: Diagnosing and mitigating context rot in long-horizon search (@askalphaxiv) · x.comAgents Context Engineering Long Context alphaXiv Fudan University Shanghai Jiao Tong University Shijie Xia Yikun Wang Zhen Huang Pengfei Liu · Qwen3.5-397B-A17B · GLM-4.7 GLM-5.0 · MiniMax-M2.5
  • 2026-07-20: Reviewer Precision Does Not Guarantee Better Multi-Agent Math Reasoning (cs.AI updates on arXiv.org) · arxiv.orgAgents LLM Evals Context Engineering
  • 2026-07-23: Poolside releases Laguna S 2.1, an open coding model with 1M-token context (Искусственный интеллект – AI, ANN и иные формы искусственного разума) · habr.comCode Agents Open Source LLMs Long Context LLM Evals Poolside · NVIDIA · Hugging Face · OpenAI · DeepSeek Thinking Machines · OpenRouter · Xiaomi · Tencent · Z ai · Moonshot · Anthropic · GitHub · Axios · White House US Department of Commerce Jason Warner David Sacks · Laguna S 2.1 DeepSeek V4 lite · Nemotron 3 Ultra · Inkling · Kimi K3

FAQ

What is GPT-OSS 120B?

GPT-OSS 120B is the larger open-weight model in OpenAI’s gpt-oss line, distributed on Hugging Face for self-hosted deployment. It pairs with the smaller gpt-oss-20b variant. GROUNDING tracks GPT-OSS 120B’s benchmarks and use in open-model serving.

What does this page track?

Dated radar mentions, source links, related concepts, and builder-relevant context for GPT-OSS 120B, collected automatically by GROUNDING.

When was GPT-OSS 120B last mentioned?

GPT-OSS 120B was most recently mentioned in a radar update dated 2026-07-23.