Monday, 15 December 2025
25.7 C
Singapore
23.5 C
Thailand
20.4 C
Indonesia
26.5 C
Philippines

OpenAI introduces Flex processing to cut AI costs for slower tasks

OpenAI launches Flex processing, cutting AI usage costs by 50% for non-urgent tasks using o3 and o4-mini models with slower response times.

OpenAI has just rolled out a new way to save on AI usage costs. It’s called Flex processing, offering lower prices in return for slightly slower performance and occasional access issues. This change is part of OpenAI’s effort to keep up with strong competition from companies like Google, pushing cheaper and faster AI models.

Flexible processing is needed to handle less urgent tasks such as testing models, enriching data, or running background processes. It’s in beta and available for two of OpenAI’s latest reasoning models, o3 and o4-mini.

Half the price for non-urgent use

Flex processing cuts your API costs exactly in half. That means you’ll pay US$5 per million input tokens (about 750,000 words) and US$20 per million output tokens if you use the o3 model. Usually, the standard prices are US$10 and US$40, respectively. For the o4-mini model, the Flex rate drops to US$0.55 per million input tokens and US$2.20 per million output tokens, compared to the regular prices of US$1.10 and US$4.40.

Of course, these savings come with a trade-off. When using Flex processing, you notice slower response times and some delays if resources are temporarily unavailable. But if your work doesn’t need instant results, this option could help you stay within budget while accessing advanced AI tools.

Why Flex is launching now

The timing of this launch isn’t random. As AI development becomes more expensive, companies seek ways to offer affordable tools. Just recently, Google introduced Gemini 2.5 Flash — a budget-friendly model that performs just as well, if not better, than competitors like DeepSeek’s R1. Input tokens also come at a lower cost.

By offering Flex processing, OpenAI makes it easier to run experiments or process large batches of data without paying premium prices. This move is aimed at users who want access to robust AI but don’t need real-time speed.

New ID checks for certain users

In the same update, OpenAI told users that some developers must now complete an ID verification process. This applies to those in tiers 1 through 3, which are based on how much you’ve spent on OpenAI services. If you fall into one of these categories and want to access the o3 model — or specific features like reasoning summaries or streaming API — you’ll need to verify your identity first.

OpenAI says this step helps prevent misuse of its tools and ensures that users follow its policies. It’s another sign that as AI grows more powerful, companies are taking extra care to manage who can use these systems and how.

Whether running a business, managing a project, or experimenting with AI for fun, Flex processing could be a cost-effective choice — especially if your tasks don’t require instant replies. With half the price and more flexibility, it might be just the right fit for your next AI-powered idea.

Hot this week

Adobe integrates Photoshop, Acrobat and Adobe Express into ChatGPT

Adobe brings Photoshop, Acrobat and Adobe Express to ChatGPT, allowing users to edit and create via natural language prompts.

Pudu Robotics unveils new robot dog as it expands global presence

Pudu Robotics unveils its new D5 robot dog in Tokyo as part of its global push into service and industrial robotics.

Grab signs partnership with Charge+ to expand EV charging network in Vietnam

Grab and Charge+ partner to expand Vietnam’s EV charging network and support the country’s shift towards green mobility.

Kaspersky uncovers macOS malware campaign abusing ChatGPT chat-sharing feature

Kaspersky reports a macOS malware campaign using ChatGPT’s chat-sharing feature to spread the AMOS infostealer.

Enterprise AI adoption accelerates as organisations deepen workflow integration

A new OpenAI report shows rapid global growth in enterprise AI, rising productivity gains, and a widening gap between leading and lagging adopters.

Tiiny AI unveils pocket-sized AI supercomputer verified by Guinness World Records

Tiiny AI reveals a Guinness-verified pocket-sized AI supercomputer designed to run massive models locally without relying on the cloud.

Samsung Galaxy Z TriFold sells out first batch, second waitlist opens in Singapore

Samsung’s Galaxy Z TriFold sells out its first batch in Singapore, with a second waitlist now open for the premium tri-fold phone.

PlayStation introduces limited edition Genshin Impact DualSense controller

PlayStation announces a limited edition Genshin Impact DualSense controller for PS5, launching in Singapore on 21 January 2026.

PGL brings Counter-Strike 2 Major to Singapore in November 2026

PGL confirms the Counter-Strike 2 Major is coming to Singapore in November 2026, marking the first CS2 Major in Southeast Asia.

Related Articles

Popular Categories