AI & LLMs · Guide · AI & Prompt Tools
What Changed in GPT-5
GPT-5's reasoning router, 400k context, pricing drops, mini/nano tiers, Atlas + Operator + Sora 2. What got better in practice and what didn't.
By FreeToolArena Staff · Updated June 2026 · 6 min read
GPT-5 (released August 2025) is the biggest practical leap from GPT-4o. Beyond the headline benchmarks, the real changes that affect day-to-day use are the reasoning router, voice + multimodal, and the API restructure.
Advertisement
The headline changes
- Reasoning router: GPT-5 picks fast vs slow thinking automatically. You don’t need to flip a model.
- 400k context: 4x GPT-4o, smaller than Claude (1M) and Gemini (2M) but plenty for most.
- Pricing drop: $2.50/$10 per 1M tokens vs GPT-4o’s same price — but with reasoning routed automatically.
- Mini + nano tiers: $0.25/$2 and $0.05/$0.40 dramatically expand cheap-tier usage.
- Atlas + Operator: agentic browsing built in.
- Sora 2: video generation rolled into ChatGPT for Plus / Pro.
What got better in practice
- Fewer obvious hallucinations on factual queries.
- More consistent instruction-following on long prompts.
- Voice mode (Advanced Voice) now feels conversational, not turn-based.
- Vision: better at reading dense text (PDFs, screenshots).
What didn’t change as much as marketing implied
- Coding: still trails Claude Sonnet/Opus on SWE-bench. Solid for autocomplete, second to Claude on agents.
- Memory: still has cross-session leakage if not pruned.
- Long context: 400k is generous but not always reliable past ~200k.
Run the cost math at the Gemini vs ChatGPT cost calculator. Compare to Claude at Claude vs ChatGPT.
Use these while you read
Tools that pair with this guide
- AI Feature Comparison MatrixVision, audio, video, tool use, web search, code interpreter, file upload, voice mode, memory, agents — across ChatGPT, Claude, Gemini, Perplexity, and 6 more.AI & Prompt Tools
- Frontier AI Model TrackerLive tracker of every frontier AI model: Claude 4.x, GPT-5, Gemini 3 Pro, DeepSeek R1/V3.2, Kimi K2, Grok 4, Llama 4, Qwen 3.5, Mistral Large 3.AI & Prompt Tools
- AI Prompt GeneratorTurn a vague idea into a structured prompt. Pick role, task, context, constraints, and output format. Works with ChatGPT, Claude, and Gemini.AI & Prompt Tools
- AI Prompt LibraryBrowse a curated catalog of prompt templates for writing, coding, marketing, and research. One click to copy.AI & Prompt Tools
Advertisement
Continue reading
- AI & LLMsGitHub Copilot Pricing and ComparisonCompare free vs paid GitHub Copilot tiers and analyze it against ChatGPT, Cursor, and Tabnine. Find the best value plan instantly with this free online guide.
- AI & LLMsGitHub Copilot Features and CapabilitiesTest what Copilot really does — code accuracy, scope limits, debugging, web dev, legacy code, tests, docs, team customization. Free guide, no sign-up.
- AI & LLMsGitHub Copilot Security and Data HandlingAudit where your code goes, who sees it, training-data policy, network needs, and what happens when Copilot suggests broken code. Free, no sign-up.
- AI & LLMsAI Fluency SkillsThe 8 sub-skills of AI fluency: prompt structure, model selection, tool use, quality calibration, iteration, context management, cost awareness, privacy.
- AI & LLMsAnthropic Skills ExplainedSkills as Anthropic's answer to Custom GPTs — markdown-defined, version-controlled in git, work in terminal. Anatomy + Skills vs Custom GPTs.
- AI & LLMsKimi K2 vs DeepSeek V3Two open-weight Chinese flagships. Kimi K2 = 1M context, DeepSeek V3.2 = top-tier reasoning + coding. Pick by use case.