Saturday, 19 April 2025
29.1 C
Singapore
33.7 C
Thailand
22.8 C
Indonesia
29.3 C
Philippines

OpenAI introduces Flex processing to cut AI costs for slower tasks

OpenAI launches Flex processing, cutting AI usage costs by 50% for non-urgent tasks using o3 and o4-mini models with slower response times.

OpenAI has just rolled out a new way to save on AI usage costs. It’s called Flex processing, offering lower prices in return for slightly slower performance and occasional access issues. This change is part of OpenAI’s effort to keep up with strong competition from companies like Google, pushing cheaper and faster AI models.

Flexible processing is needed to handle less urgent tasks such as testing models, enriching data, or running background processes. It’s in beta and available for two of OpenAI’s latest reasoning models, o3 and o4-mini.

Half the price for non-urgent use

Flex processing cuts your API costs exactly in half. That means you’ll pay US$5 per million input tokens (about 750,000 words) and US$20 per million output tokens if you use the o3 model. Usually, the standard prices are US$10 and US$40, respectively. For the o4-mini model, the Flex rate drops to US$0.55 per million input tokens and US$2.20 per million output tokens, compared to the regular prices of US$1.10 and US$4.40.

Of course, these savings come with a trade-off. When using Flex processing, you notice slower response times and some delays if resources are temporarily unavailable. But if your work doesn’t need instant results, this option could help you stay within budget while accessing advanced AI tools.

Why Flex is launching now

The timing of this launch isn’t random. As AI development becomes more expensive, companies seek ways to offer affordable tools. Just recently, Google introduced Gemini 2.5 Flash — a budget-friendly model that performs just as well, if not better, than competitors like DeepSeek’s R1. Input tokens also come at a lower cost.

By offering Flex processing, OpenAI makes it easier to run experiments or process large batches of data without paying premium prices. This move is aimed at users who want access to robust AI but don’t need real-time speed.

New ID checks for certain users

In the same update, OpenAI told users that some developers must now complete an ID verification process. This applies to those in tiers 1 through 3, which are based on how much you’ve spent on OpenAI services. If you fall into one of these categories and want to access the o3 model — or specific features like reasoning summaries or streaming API — you’ll need to verify your identity first.

OpenAI says this step helps prevent misuse of its tools and ensures that users follow its policies. It’s another sign that as AI grows more powerful, companies are taking extra care to manage who can use these systems and how.

Whether running a business, managing a project, or experimenting with AI for fun, Flex processing could be a cost-effective choice — especially if your tasks don’t require instant replies. With half the price and more flexibility, it might be just the right fit for your next AI-powered idea.

Hot this week

OpenAI’s latest reasoning AI models are more prone to making mistakes

OpenAI’s new o3 and o4-mini AI models perform better in some areas but hallucinate more often than their predecessors, raising concerns.

Identity theft and document forgery pose rising fraud risks for APAC businesses

Businesses in APAC face rising fraud threats, with identity theft and document forgeries driving up costs and urging investment in smarter ID tools.

Trump leaves smartphones and computers out of new tariff hike

Trump exempts phones, laptops, and chips from new tariffs, easing price fears but keeping pressure on China with other duties.

YouTrip adds a Malaysian Ringgit wallet to help you save more on JB trips

YouTrip now lets you store MYR and offers free JB shuttles and cashback to celebrate, making your trips across the Causeway more rewarding.

AI is reshaping tech infrastructure as Seagate urges balance between cost and carbon

Seagate’s new global report urges data centre operators to balance sustainability with cost as AI-driven data demands surge.

OpenAI’s latest reasoning AI models are more prone to making mistakes

OpenAI’s new o3 and o4-mini AI models perform better in some areas but hallucinate more often than their predecessors, raising concerns.

Google removes over 5 billion ads in 2024 as AI boosts enforcement against online scams

Google’s Ads Safety Report 2024 shows how AI helped remove over 5.1 billion ads and block 700,000 scam accounts from its platform.

Microsoft highlights growing AI-assisted scams and offers advice on how to stay safe

Microsoft’s latest report warns of rising AI-driven scams and outlines new tools and tips to help users stay safe online.

Qualcomm unveils new Snapdragon 8s Gen 4 with high-end features for less

Qualcomm quietly unveils the Snapdragon 8s Gen 4 with high-end features and strong performance for next-gen smartphones at a lower price.

Related Articles

Popular Categories