Skip to main content

What model-makers publish

From the Labs

System cards, usage studies, and safety research published by the AI labs themselves — primary evidence of how frontier models behave and how their makers measure it.

61 entries, newest first

7 Oct 2026 Anthropic Lab publication

Lab publication

System Card: Claude Haiku 5.5

System card for Claude Haiku 5.5, Anthropic's small-model release of October 2026, condensed relative to its frontier cards. It reports Responsible Scaling Policy and cyber evaluations, harmlessness…

7 Oct 2026 OpenAI Lab publication

Lab publication

GPT-6 Sol and GPT-6 Luna: October 2026 update

System card for the October 2026 release of GPT-6 Sol and GPT-6 Luna, which replace GPT-5.6 Sol and GPT-5.6 Luna across all free and paid ChatGPT plans globally (Codex and ChatGPT Work keep the Septe…

7 Oct 2026 OpenAI Lab publication

Lab publication

Helping teens learn, plan, and shape the future of AI

Product post in which OpenAI reports early usage figures for ChatGPT for Teens, the under-18 experience launched on 2026-08-18, and announces a College Planner feature, flashcards and quizzes, and th…

28 Sept 2026 Anthropic Lab publication

Lab publication

System Card: Claude Sonnet 5.5

System card for Claude Sonnet 5.5, released 28 September 2026, reporting pre-deployment safety, alignment and capability evaluations. Its safeguards chapter reports single-turn and multi-turn results…

22 Sept 2026 Anthropic Lab publication

Lab publication

Claude Opus 5.5 System Card

230-page system card for Claude Opus 5.5, the first model in the Claude 5.5 family, released 22 September 2026. Alongside RSP, cyber and agentic-safety sections it reports harmful-request evaluations…

21 Sept 2026 xAI (styled SpaceXAI in the card) Lab publication

Lab publication

Model Card: Grok 4.7

Model card for Grok 4.7, released on 21 September 2026 as xAI's (now styled SpaceXAI) frontier coding and knowledge-work model. Alongside capability benchmarks, the 30-page card reports the company's…

18 Sept 2026 OpenAI Lab publication

Lab publication

An Australian Youth Safety Blueprint

Seven-page policy document in which OpenAI sets out six pillars it proposes for Australian youth-AI policy: recognising positive uses and AI literacy in education; privacy-preserving age assurance; u…

10 Sept 2026 Anthropic (Threat Intelligence) Lab publication

Lab publication

Detecting and countering misuse of AI: September 2026

Anthropic's periodic threat-intelligence report covering December 2025 to August 2026 across seven harm areas. One case study, GTG-15001, documents a China-based app studio that used Claude to build…

8 Sept 2026 OpenAI (OpenAI Group PBC) Lab publication

Lab publication

OpenAI's AI and Teen Development Research Grant Program

A funding call from OpenAI for independent research on how generative AI use relates to adolescent development, including safeguards and design choices. The programme opened 8 September 2026 with a d…

3 Sept 2026 Character.AI (Character Technologies) Lab publication

Lab publication

Continuing To Build Upon Our Safety Priorities

A first-party safety update from Character.AI describing safeguards in operation on its platform as of September 2026. It states that self-harm safeguards consider the surrounding conversation, inclu…

3 Sept 2026 OpenAI Lab publication

Lab publication

GPT-6 Astra System Card

System card for GPT-6 Astra, published 2026-09-03. Most of the document concerns cyber capabilities at OpenAI's Preparedness 'Critical' threshold, alignment and chain-of-thought monitorability. The p…

1 Sept 2026 Anthropic Lab publication

Lab publication

System Card: Claude Fable 5.1 & Claude Mythos 5.1

Anthropic's 212-page system card for Claude Fable 5.1 and Claude Mythos 5.1, two safeguard configurations of the same frontier model, released 1 September 2026. Alongside Responsible Scaling Policy,…

1 Sept 2026 Microsoft (Office of Responsible AI) Lab publication

Lab publication

2026 Responsible AI Transparency Report

Microsoft's third annual responsible-AI transparency report, covering July 2025 to June 2026. One of its four 2026 trends is people turning to conversational AI for personal advice, health questions…

28 Aug 2026 Anthropic Lab publication

Lab publication

Automated Researchers Can Reliably Mitigate Alignment Failures

Anthropic study of whether automated alignment researchers — Claude autonomously proposing and running post-training interventions — can mitigate ten benchmark-measurable alignment failures including…

26 Aug 2026 Anthropic (Societal Impacts), with Stanford SALT Lab, Oxford Human Information Processing Lab and METR Lab publication

Lab publication

Enabling independent research on how people use Claude

A report on a pilot in which three external research groups designed their own studies and ran them through Anthropic Insights, the company's privacy-preserving aggregate analysis tool, on roughly 25…