llm CLI upgrades default model to GPT-5.6 Luna

Simon Willison's llm CLI tool has shipped RC2, elevating its default model from GPT-4o mini to OpenAI's newer GPT-5.6 Luna. The shift signals a meaningful capability upgrade for users who haven't explicitly configured a preference, though at higher per-token costs. This reflects the rapid model iteration cycle now standard in the LLM ecosystem: tools must actively track and adopt frontier capabilities to remain competitive, and even incremental releases now bundle infrastructure fixes alongside model swaps. For developers relying on llm as their primary interface, the default change carries real implications for both output quality and API spend.
Modelwire context
Analyst takeThe real story isn't the RC2 release itself, but the cost trade-off embedded in the default. GPT-5.6 Luna costs more per token than GPT-4o mini. Willison's choice to make it the default means llm users now absorb higher API spend unless they actively opt down, shifting the cost burden from tool maintainers to end users.
This is largely disconnected from recent activity in the space, because we have no prior coverage tracking how CLI tools and SDK maintainers are responding to accelerating model releases. What this belongs to is the broader question of tool economics: as models improve faster, maintainers face a choice between staying on older, cheaper baselines or chasing capability at the cost of user budgets. This story shows one answer. Watch whether other popular tools (like Anthropic's Claude CLI or Vercel's AI SDK) follow similar patterns or hold their defaults stable.
If llm's usage metrics show a measurable drop in daily active users after this release (trackable via GitHub stars, PyPI downloads, or Willison's own telemetry if public), that signals users are price-sensitive enough to switch tools rather than absorb the cost increase. If adoption holds flat or grows, it means developers either don't notice the cost delta or value the capability bump enough to pay it.
This analysis is generated by Modelwire’s editorial layer from our archive and the summary above. It is not a substitute for the original reporting. How we write it.
MentionsSimon Willison · llm · OpenAI · GPT-5.6 Luna · GPT-4o mini
Modelwire Editorial
This synthesis and analysis was prepared by the Modelwire editorial team. We use advanced language models to read, ground, and connect the day’s most significant AI developments, providing original strategic context that helps practitioners and leaders stay ahead of the frontier.
Modelwire summarizes, we don’t republish. Simon Willison originally reported this story as “llm 0.32rc2”. The full content lives on simonwillison.net. If you’re a publisher and want a different summarization policy for your work, see our takedown page.