P.01OpenAI Cut GPT-5.6 API Prices 80%: What Changes
OpenAI cut GPT-5.6 Luna's API price 80% and Terra's 20% three weeks after launch. What moved, why so fast, and whether to switch tiers rather than coast.
Tag
10 articles tagged #OpenAI.
P.01OpenAI cut GPT-5.6 Luna's API price 80% and Terra's 20% three weeks after launch. What moved, why so fast, and whether to switch tiers rather than coast.
P.02OpenAI's Presence deploys voice and chat agents with access controls, simulation testing, and Codex improvement loops. When building in-house still wins.
P.03OpenAI's gpt-realtime-2.1-mini brings reasoning and tool use down-market at 25% lower latency. What changed, what it costs, and when to skip the flagship.
P.04Models that think before they answer are reshaping AI engineering. We break down how extended thinking, reasoning budgets, and chain-of-thought inference work across Claude, OpenAI o3, Gemini, and DeepSeek-R1 — and when you should actually use them.
Anthropic used Super Bowl ads to pledge Claude will never show ads, contrasting itself with OpenAI's move to add sponsored suggestions to ChatGPT.
OpenAI is retiring GPT-4o, GPT-4.1, GPT-4.1 mini, and o4-mini from ChatGPT on February 13, 2026. Only 0.1% of users still choose GPT-4o daily, but the model's retirement marks the end of the GPT-4 generation.
The big tech AI arms race hit $350B+ in 2026 spending across Meta, Microsoft, Alphabet, and Amazon. Here's where the money is going and what it means for you.
AI infrastructure runs on tight power and chip supply limits in 2026, shaping API pricing, latency, and model availability for developers building on top of it.
Architecture patterns, prompt engineering, cost control, and a production checklist for AI applications. The parts that outlast any given model.
P.10Meta Compute, OpenAI's 750MW deal, and a projected $3 trillion investment in AI infrastructure. The biggest story in tech isn't about models—it's about who controls the compute.