No Result
View All Result
  • Login
Wednesday, August 26, 2026
theadvisertimes.com
  • Home
  • Business
  • Financial Planning
  • Personal Finance
  • Investing
  • Money
  • Economy
  • Markets
  • Stocks
  • Trading
  • Home
  • Business
  • Financial Planning
  • Personal Finance
  • Investing
  • Money
  • Economy
  • Markets
  • Stocks
  • Trading
No Result
View All Result
theadvisertimes.com
No Result
View All Result
Home Cryptocurrency

GPT-5.4 Pro jumps to 150 IQ on MESNA Norway test as OpenAI breaks its own record

by theadvisertimes.com
5 months ago
in Cryptocurrency
Reading Time: 8 mins read
A A
0
GPT-5.4 Pro jumps to 150 IQ on MESNA Norway test as OpenAI breaks its own record
Share on FacebookShare on TwitterShare on LInkedIn


Make CryptoSlate preferred on

OpenAI’s latest GPT-5.4 Pro model has now achieved an IQ score higher than 99.96% of all human beings, giving markets a fresh signal that AI capability gains are starting to outpace the usual product-cycle noise.

OpenAI’s GPT-5.4 Pro touches 150 on public IQ benchmark as markets enter another macro-heavy week

TrackingAI’s public leaderboard now places OpenAI GPT-5.4 Pro at an IQ score of 150, a sharp step up from the 136 score that OpenAI’s o3 posted on the Mensa Norway test last year.

The jump arrives at a moment when market attention has narrowed around Iran, energy, labor softness, and the next inflation print. That creates a different question for the week ahead: how quickly is machine intelligence compounding, and when will that acceleration begin to overlap with economic positioning?

Why this matters: A move from 136 to 150 on a widely understood benchmark compresses a complex capability shift into a simple signal. For businesses, that signal feeds directly into decisions around automation, software budgets, and headcount planning. For markets, it adds another variable alongside rates, inflation, and growth expectations.

OpenAI introduced GPT-5.4 as its most capable and efficient frontier model for professional work, with stronger coding, tool use, and computer use, and a context window of up to 1 million tokens. In the same release, OpenAI said GPT-5.4 achieved a new state of the art on GDPval and exceeded human performance on OSWorld-Verified.

Those benchmarks are separate from a public IQ test, yet the direction of travel aligns. Capability is rising across separate measurement systems, and that rise is becoming fast enough to influence budgeting, hiring plans, workflow design, and software spend.

A score of 150 on a public IQ-style benchmark compresses a broader capability move into a single, portable signal. The number is easy to understand even before the methodology is debated.

The earlier o3 Mensa result established the benchmark and its limits. GPT-4.1’s one-million-token context window showed how OpenAI was extending model utility across long-horizon code and document tasks, while our analysis of OpenAI’s expanding capital loop linked model progress to hardware expansion, financing loops, and infrastructure demand.

Taken together, those developments place the latest IQ score within a broader commercial and economic context. A move from 136 to 150 on a public benchmark is striking on its own. A move from 136 to 150 while OpenAI is pushing deeper into tool use, computer use, enterprise productivity, and capital-intensive infrastructure carries broader implications.

Public IQ benchmarks are limited, but the capability curve is still moving higher

Public IQ-style tests remain imperfect instruments for measuring frontier models. TrackingAI runs a public Mensa-style benchmark and also maintains a harder private offline test.

IQ-style tests compress a narrow slice of cognitive performance into a single number, obscuring variation across reasoning types, context handling, creativity, and real-world problem-solving.

For AI and humans alike, scores are sensitive to test design, training exposure, and pattern familiarity, which makes them a noisy proxy for general capability.

An IQ of 150 sits at the extreme upper tail of the distribution, often associated with individuals such as Albert Einstein or Richard Feynman. In practical terms, it implies very fast abstraction, strong pattern recognition, and the ability to navigate complex, multi-step problems with limited guidance.

The platform reports scores as rolling averages across recent completions, and the methodology raises familiar questions around prompt structure, reproducibility, training-set contamination, and format familiarity. Those concerns were already visible when o3 reached 136, and they remain active now that GPT-5.4 Pro sits at 150.

OpenAI’s o3 scores 136 on Mensa Norway test, surpassing 98% of human populationOpenAI’s o3 scores 136 on Mensa Norway test, surpassing 98% of human population
Related Reading

OpenAI’s o3 scores 136 on Mensa Norway test, surpassing 98% of human population

OpenAI’s o3 model reaches Mensa-Level IQ in independent testing.

Apr 17, 2025 · Liam ‘Akiba’ Wright

Even with those limits, the broader pattern has become harder to dismiss. One isolated benchmark result can be explained away as a quirk. A cluster of gains across public IQ-style testing, coding, browser use, desktop navigation, and knowledge-work performance carries more analytical weight.

TrackingAI’s latest leaderboard places GPT-5.4 Pro at the top of its public IQ board ahead of all Cluade, Gemini, Qwen, and Grok models, offering an external, legible public benchmark that maps quickly onto the broader capability debate.

Few people need a detailed understanding of benchmark design to grasp that 150 sits in a rare range and investors do not need to accept every premise behind an IQ-style test to recognize that a jump of this size suggests acceleration rather than drift.

Chart titled “AI IQ Test Results” showing average Mensa Norway IQ scores for major AI models on a bell curve, with OpenAI’s GPT-5.4 variants plotted near the top end of the range.Chart titled “AI IQ Test Results” showing average Mensa Norway IQ scores for major AI models on a bell curve, with OpenAI’s GPT-5.4 variants plotted near the top end of the range.
Chart titled “AI IQ Test Results” showing average Mensa Norway IQ scores for major AI models on a bell curve, with OpenAI’s GPT-5.4 variants plotted near the top end of the range.

Enterprise buyers also do not need to believe that IQ equals general intelligence to see that systems with stronger pattern recognition, stronger tool use, and stronger long-horizon task handling are moving toward economically useful territory, extending far beyond puzzle-solving.

This points toward systems that can search, plan, verify, navigate, and produce real work across extended contexts. In that setting, the IQ score functions less as a novelty number and more as a signal of the density of frontier reasoning.

There is also competitive value in the leaderboard itself. A leadership position on a public benchmark reinforces OpenAI’s standing in the race for visible capability leadership, especially at a moment when model differentiation is becoming harder to discern from architecture notes alone.

Benchmark leadership compresses complexity into a simple hierarchy. It offers developers a signal, enterprise buyers a narrative handle, and investors another proxy for where the capability frontier currently sits.

CryptoSlate Daily Brief

Daily signals, zero noise.

Market-moving headlines and context delivered every morning in one tight read.

5-minute digest 100k+ readers

Free. No spam. Unsubscribe any time.

Whoops, looks like there was a problem. Please try again.

You’re subscribed. Welcome aboard.

OpenAI’s benchmark climb is beginning to overlap with the economic week ahead

The week ahead still runs through macro. The Bureau of Labor Statistics calendar clearly lays out the next key releases: the FOMC minutes from the March 17 to 18 meeting, due on April 8; the March Consumer Price Index, due on April 10; and the March Producer Price Index, due on April 14.

That schedule keeps rates, inflation, and growth anxiety in the foreground, but beneath that surface, a second economic track is taking shape, and OpenAI sits near its center.

Capability growth in frontier AI increasingly intersects with capital allocation. A model that pushes higher on public reasoning tests while also improving in coding, search, and computer use changes how businesses think about workflow redesign. It changes what software buyers expect from copilots and agents. It changes how quickly enterprises move from experimentation toward deployment.

Jack Dorsey recently posted that Block is moving “from hierarchy to intelligence,” using AI to take over coordination work once handled by management layers as the company reorganizes around individual contributors, directly responsible individuals, and player-coaches

Capability growth also changes which tasks can be carved out of labor cost structures and reassigned to software. These effects move through narrower channels first, including document workflows, spreadsheet workflows, customer support, research tasks, browser automation, internal operations, code generation, and verification loops.

OpenAI’s commercial direction reinforces that interpretation. In its GPT-5.4 launch materials, the company described stronger performance in professional work, stronger tool search, native computer use, and gains in benchmarked knowledge work across occupations that map directly onto the U.S. economy.

That places AI capability growth inside a familiar market question, where spending flows next if these systems continue improving at this pace.

The answer extends beyond model subscription revenue into cloud demand, chips, data centers, networking, power, software licenses, and labor productivity assumptions. OpenAI’s expanding capital loop already reflects part of that structure, and the benchmark gain adds a simpler public-facing signal on top of it.

That overlap is what gives the latest result broader relevance during a macro-heavy week. Markets already know the CPI setup. Markets already know oil prices can feed into inflation expectations. Markets already know the Fed minutes will be parsed for policy tone.

But is the growth in intelligence itself beginning to behave like a macro variable? Faster capability gains can alter enterprise spending plans, tighten competitive pressure across white-collar functions, support higher infrastructure outlays, and strengthen the case for AI-linked capital expenditure even in a slower nominal growth environment.

When TrackingAI shows GPT-5.4 Pro at 150, the number falls within a market that already views OpenAI as more than a lab. It is a platform company, a deployment company, an infrastructure customer, and a signal generator for adjacent sectors.

The next test sits in two places at once. One is methodological; public IQ-style benchmarks will keep drawing scrutiny, and they should. The other is economic; markets will decide, step by step, whether capability jumps of this size deserve to be priced alongside labor data, rate expectations, and capital spending trends.

OpenAI’s latest benchmark climb pushes that decision closer. The score is compact, legible, and easy to circulate. Its deeper relevance comes from the same place as the company’s broader product push; the frontier is still climbing, and the economic footprint of that climb is becoming harder to keep in a separate category.

Mentioned in this article



Source link

Tags: BreaksGPT5.4jumpsMESNANorwayOpenAIProrecordtest
ShareTweetShare
Previous Post

Seniors 62+ Can Take College Classes Tuition‑Free at Public Universities

Next Post

What’s Open, Closed on Easter? See Hours for Restaurants, Stores, More

Related Posts

XRP Price Forecast as Binance Whale Activity Remains Elevated Amid XRPL Tokenization Push

XRP Price Forecast as Binance Whale Activity Remains Elevated Amid XRPL Tokenization Push

by theadvisertimes.com
August 8, 2026
0

Ripple (XRP) price is up by 0.46% today, August 8, to trade at $1.04 at the time of writing. XRP...

Bitcoin outlook: ,000 or ,000 this weekend

Bitcoin outlook: $70,000 or $60,000 this weekend

by theadvisertimes.com
August 8, 2026
0

Bitcoin traded near $65,000 heading into the weekend, sitting at the center of two macro forces pulling in opposite directions.The...

Lummis Warns US Crypto Rules Remain Broken as CLARITY Fight Stalls

Lummis Warns US Crypto Rules Remain Broken as CLARITY Fight Stalls

by theadvisertimes.com
August 7, 2026
0

Key TakeawaysLummis vows to keep pushing the CLARITY Act despite its stalled progress.The CLARITY Act would strengthen crypto safeguards and...

Bybit Uses Tokenised Equities as Underlyings for Structured Yield

Bybit Uses Tokenised Equities as Underlyings for Structured Yield

by theadvisertimes.com
August 7, 2026
0

Bybit is expanding the role of tokenised equities on its platform by using more xStocks as underlyings for its Dual...

Reform UK Chair Calls for Probe into SBF-Linked Donation: Report

Reform UK Chair Calls for Probe into SBF-Linked Donation: Report

by theadvisertimes.com
August 7, 2026
0

The chairman of the UK’s Reform party has called for an investigation following reports of a $50,000 political donation linked...

CFTC Warns Polymarket, Kalshi Against American-Style Gambling Odds Amid State Scrutiny

CFTC Warns Polymarket, Kalshi Against American-Style Gambling Odds Amid State Scrutiny

by theadvisertimes.com
August 7, 2026
0

The U.S. Commodity Futures Trading Commission (CFTC) has issued a warning to its regulated entities that offer prediction markets over...

Next Post
What’s Open, Closed on Easter? See Hours for Restaurants, Stores, More

What’s Open, Closed on Easter? See Hours for Restaurants, Stores, More

Kevin Warsh Fed Chair Nomination Hearing Set for April 16

Kevin Warsh Fed Chair Nomination Hearing Set for April 16

  • Trending
  • Comments
  • Latest
Wealth management has got junior advisors’ first 90 days covered. What happens on day 91?

Wealth management has got junior advisors’ first 90 days covered. What happens on day 91?

August 7, 2026
Biggerpockets Pro Members Can Now Turn Home Equity Into a Flexible Line of Credit With Aven

Biggerpockets Pro Members Can Now Turn Home Equity Into a Flexible Line of Credit With Aven

August 3, 2026
The 19 Largest Global Startup Funding Rounds of June 2026 – AlleyWatch

The 19 Largest Global Startup Funding Rounds of June 2026 – AlleyWatch

July 27, 2026
Fourth of July 2026 Freebies and Deals

Fourth of July 2026 Freebies and Deals

July 3, 2026
Kellogg’s Back to School Snacks Instant Savings: Save  off  Purchase + Deal Scenario!

Kellogg’s Back to School Snacks Instant Savings: Save $10 off $35 Purchase + Deal Scenario!

August 6, 2026
CVS Deals Under  This Week

CVS Deals Under $1 This Week

July 27, 2026
XRP Price Forecast as Binance Whale Activity Remains Elevated Amid XRPL Tokenization Push

XRP Price Forecast as Binance Whale Activity Remains Elevated Amid XRPL Tokenization Push

0
Wealth management has got junior advisors’ first 90 days covered. What happens on day 91?

Wealth management has got junior advisors’ first 90 days covered. What happens on day 91?

0
Is Lettuce Safe to Eat Now? What to Know Amid Cyclospora Outbreak

Is Lettuce Safe to Eat Now? What to Know Amid Cyclospora Outbreak

0
The Unwinnable Iran War | Armstrong Economics

The Unwinnable Iran War | Armstrong Economics

0
Bezeq declares NIS 515m dividend

Bezeq declares NIS 515m dividend

0
AI Will Clarify What Asset Managers Are Paid For

AI Will Clarify What Asset Managers Are Paid For

0
XRP Price Forecast as Binance Whale Activity Remains Elevated Amid XRPL Tokenization Push

XRP Price Forecast as Binance Whale Activity Remains Elevated Amid XRPL Tokenization Push

August 8, 2026
Is Lettuce Safe to Eat Now? What to Know Amid Cyclospora Outbreak

Is Lettuce Safe to Eat Now? What to Know Amid Cyclospora Outbreak

August 8, 2026
Bitcoin outlook: ,000 or ,000 this weekend

Bitcoin outlook: $70,000 or $60,000 this weekend

August 8, 2026
Mortgage and refinance interest rates today, Saturday, August 8, 2026: Rates mixed this weekend

Mortgage and refinance interest rates today, Saturday, August 8, 2026: Rates mixed this weekend

August 8, 2026
CEO of the world’s largest workspace provider says commuting will be extinct by 2040

CEO of the world’s largest workspace provider says commuting will be extinct by 2040

August 8, 2026
F&O Talk: Smallcaps look strong on charts, says Sudeep Shah; outlines Trent, Swiggy, Kalyan Jewellers strategy

F&O Talk: Smallcaps look strong on charts, says Sudeep Shah; outlines Trent, Swiggy, Kalyan Jewellers strategy

August 8, 2026
theadvisertimes.com

Get the latest news and follow the coverage of Business & Financial News, Stock Market Updates, Analysis, and more from the trusted sources.

CATEGORIES

  • Business
  • Cryptocurrency
  • Economy
  • Financial Planning
  • Investing
  • Market Analysis
  • Markets
  • Money
  • Personal Finance
  • Startups
  • Stock Market
  • Trading

LATEST UPDATES

  • XRP Price Forecast as Binance Whale Activity Remains Elevated Amid XRPL Tokenization Push
  • Is Lettuce Safe to Eat Now? What to Know Amid Cyclospora Outbreak
  • Bitcoin outlook: $70,000 or $60,000 this weekend
  • Our Great Privacy Policy
  • Terms of Use, Legal Notices & Disclosures
  • About Us
  • Contact Us

© Copyright 2024 All Rights Reserved
See articles for original source and related links to external sites.

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Business
  • Financial Planning
  • Personal Finance
  • Investing
  • Money
  • Economy
  • Markets
  • Stocks
  • Trading

© Copyright 2024 All Rights Reserved
See articles for original source and related links to external sites.