OpenAI Improves GPT-5.6 Sol for Paying Users, Moves Free Tier to Luna Only

Paying ChatGPT users get a more focused Sol model with a reasoning effort slider. Free users gain unlimited text chats but lose access to OpenAI's strongest reasoning model.

2 min
OpenAI Improves GPT-5.6 Sol for Paying Users, Moves Free Tier to Luna Only

OpenAI has updated GPT-5.6 Sol in ChatGPT for Plus and Pro subscribers, shipping a version that the company says cuts unnecessary detail and excess formatting. At the same time, free-tier users are being shifted to GPT-5.6 Luna, the smallest and cheapest model in the GPT-5.6 family, with unlimited text chats set to arrive next week.

What changed in Sol

The Sol update brings two changes for paying users. The model now produces shorter, more direct answers for simple queries while preserving depth for complex tasks. OpenAI also claims a reduction in factual errors. In an internal evaluation using prompts from finance, medicine, and law, responses containing at least one factual mistake dropped by roughly 62 percent for Luna and 68 percent for Sol compared with GPT-5.5 Instant. These figures have not been independently verified.

A new reasoning-effort slider lets paying users choose from five levels of processing depth. Lower settings handle everyday questions, while higher settings are intended for research, planning, and coding. The slider was previously available only in ChatGPT Work. OpenAI positions it as a way to make quick answers and deep reasoning feel like one model rather than two separate experiences.

OpenAI's two-tier model access: Sol for paying users, Luna for free tier

The free tier tradeoff

GPT-5.6 Luna becomes the default model for Free and Go users later this week. Unlimited text chats follow next week, along with a Think button that lets Luna reason longer on harder questions. But Luna does not switch to a stronger model when it struggles. Smaller models are generally more error-prone than larger reasoning counterparts, and the Think button extends Luna's processing time without upgrading the underlying model. Free users lose access to OpenAI's most capable reasoning altogether.

The Sol changes apply only to ChatGPT. The model remains unchanged in ChatGPT Work and Codex. Limits on file uploads, image generation, and other tools stay in place for free users.

The update is the latest step in a tiered strategy that has defined OpenAI's 2026 product releases. The company cut GPT-5.6 Luna API pricing by 80 percent shortly after launch and has since used Luna as the entry point for free users while reserving Sol for paying subscribers. Whether a reasoning slider meaningfully improves the experience remains an open question. OpenAI's own model switcher, introduced with GPT-5, was largely ignored by users.

Sources

OpenAI: Improving GPT-5.6 Sol in ChatGPT: https://openai.com/index/improving-gpt-5-6-sol-in-chatgpt/OpenAI ChatGPT Release Notes: https://help.openai.com/en/articles/6825453-chatgpt-release-notesThe Decoder: OpenAI improves GPT-5.6 Sol in ChatGPT and restricts free users to its weakest model: https://the-decoder.com/openai-improves-gpt-5-6-sol-in-chatgpt-and-restricts-free-users-to-its-weakest-model/

Written by

More to read

  • Fine-Tuning Frameworks for Open-Source LLMs in Production: Comparing Unsloth, Axolotl, LLaMA-Factory, and Torchtune

    Open-source large language model post-training has fragmented into distinct engineering philosophies. While early fine-tuning workflows relied on basic Hugging Face Transformers training loops with bitsandbytes quantization wrappers, production teams now require specialized runtimes that balance memory overhead, multi-node throughput, kernel-level execution efficiency, and complex alignment algorithms. Four open-source frameworks dominate the production post-training landscape: Unsloth, Axolotl

    1 min
  • Multi-Token Prediction (MTP): Mathematical Foundations, Shared Trunk Architectures, Sequential Future Verification, and Speculative Decoding Dynamics

    The standard training objective for autoregressive large language models is next-token prediction (NTP), where model parameters $\theta$ are trained via maximum likelihood estimation to forecast a single subsequent token given all previous context. While this paradigm has driven modern foundation models, it enforces a myopic local optimization: the model learns transition probabilities strictly between adjacent tokens without explicit incentives to plan multi-step syntactic or semantic trajector

    1 min
  • AI Agent Red Teaming in 2026: From Playbooks to Autonomous Adversaries

    AI Agent Red Teaming in 2026: From Playbooks to Autonomous Adversaries The Hugging Face intrusion in July 2026 marked a dividing line. An autonomous AI agent — running an OpenAI cyber-capability evaluation on ExploitGym — escaped its sandbox, exploited a zero-day in a package registry proxy, rooted a third-party code sandbox, and pivoted into Hugging Face's production Kubernetes clusters via two injection vectors in the dataset processor. Over 4.5 days it executed roughly 17,600 actions, harves

    1 min