GPT-5.6's IQ Breaks 130 Genius Threshold for the First Time, Outsmarting 99% of Humans

marsbitОпубликовано 2026-07-16Обновлено 2026-07-16

Введение

GPT-5.6 has reportedly achieved an IQ score of 136 on Tracking AI's proprietary offline test, surpassing the human "genius" threshold of 130 for the first time. This places it above an estimated 99% of humans in this specific metric. The test is designed to prevent memorization by using a private question bank. Multiple GPT-5.6 variants, including the vision model, consistently scored 136, leading competitors like Claude-5 Fable (130). User anecdotes suggest practical superiority over rivals in real-world coding and problem-solving tasks, such as building a physics simulation or a customer service app from a single prompt. While some speculate this approaches AGI for most users, the article notes IQ tests only measure a narrow slice of cognitive ability like pattern recognition. The significance lies in GPT-5.6's apparent ability to translate high test scores into effective task performance on novel, real-world problems.

Today, 99% of the global human population is actually outperformed by an AI in terms of IQ.

In Tracking AI's latest offline IQ test, multiple versions of the GPT-5.6 "full suite" soared to a score of 136.

This is the first time an LLM has pushed its IQ beyond the 130 mark.

In the distribution of human intelligence, 130 is the starting line for "genius," a level only about 1% of the global population can reach.

In other words, GPT-5.6 is smarter than 99% of humans.

GPT-5.6 Racks Up 136 Points, IQ Breaks "Genius Line" for the First Time

How credible is this "IQ"?

In fact, Tracking AI uses two sets of questions.

One is a public Mensa Norway-style test, available online for anyone to take, which models have already scored over 140 on.

The other is its own curated "offline question bank." It's not public, prevents leaks, and is specifically designed to block the loophole of "models memorizing answers in advance."

The 136 points GPT-5.6 achieved this time was on this most difficult, anti-cheating offline test.

On this offline leaderboard, the various variants of GPT-5.6 (including the vision version) collectively surged to 136 points, leaving all competitors far behind.

Close behind is Claude-5 Fable, with 130 points.

Further down, names like GPT-5.6 LUNA Max and Claude-4.8 Opus are still hovering between 117 and 123 points.

It's important to note that this 130-point threshold had never been crossed before.

Over the past year, wave after wave of models, from o3 to various flagship models, surged forward, all getting stuck at the 130-point door, with none truly stepping into the "genius range."

GPT-5.6 is the first to kick that door open.

And it didn't achieve this score alone; the entire SOL, TERRA family collectively soared to 136, with even the vision version keeping pace.

On Reddit, a developer conducted a hands-on test and concluded that GPT-5.6's intelligence feels significantly higher than GPT-5.5's.

In the following test questions, GPT-5.6 achieved outstanding results in the shortest possible time.

One test score might not be convincing enough, so what does GPT-5.6 look like when taken out of the exam room and put to real work?

More Than Just a Score: Putting GPT-5.6 to Work

Developer Amir Bohlooli fed the same physics simulation prompt to both Fable 5 and GPT-5.6 Sol, expecting to be crushed by Fable, but ended up being amazed by GPT.

It chose particle fluid simulation, with physics progressing in real-time rather than blindly running fixed calculations per frame, cramming CSS, interface, and rendering all into a single HTML file, and automatically hosting it as a shareable webpage. In short, a finished product.

Similarly, Ramanpal Singh used a single prompt to create a RAG-based customer service ticketing system.

Four roles, an admin backend, embeddable components, and it can automatically categorize complaints, recognize sentiment, and draft replies.

It built 5 such apps in one go, at a cost that was only a fraction of what Fable 5 would require.

The most vivid story is from Claire Vo.

A few days ago, she was stuck on a bug, thinking her own code was broken. After switching to GPT-5.6 Sol, she just threw out the line, "I just don't believe I can't fix this."

Sol fixed it in one attempt and even managed to get it running on other models.

Her assessment hit the nail on the head: Fable gets bogged down in technical absolute precision, becoming its own trap, while Sol's pragmatic approach gets the job done.

It has to be said, there's an entire real-world project between an AI that can solve test problems and an AI that can save the day.

Does This Count as AGI?

Some netizens have said, "For 99% of people, this is already AGI."

Looking at it calmly, this 136 score was achieved on a specific offline / Mensa Norway-style test by Tracking AI.

What it measures is mainly "standardized cognition" like abstract pattern recognition and logical reasoning.

The problem is: IQ tests were never designed for large models.

A Mensa exam paper can't measure a model's factual reliability, its tool-calling ability, or how dependable it is in real professional scenarios.

It only slices off one thin layer of "intelligence" and tells you how bright that slice is.

However, hands-on testing by users provides the other half of the answer: GPT-5.6 seems to be slowly merging the two capabilities of "solving test problems" and "getting things done."

The questions in standardized tests are ones models have likely seen thousands of times in their training data; the real test of skill is with those new problems they've never encountered and have no answers to copy from.

Whoever can hold steady there truly deserves the word "intelligence."

References:

https://x.com/davidpattersonx/status/2077049232490672458

https://trackingai.org/

This article is from the WeChat public account "新智元" (New AI Era), author: ASI Revelation

Трендовые криптовалюты

Связанные с этим вопросы

QAccording to the article, what was the significant achievement of GPT-5.6 in the Tracking AI offline IQ test?

AGPT-5.6 achieved a score of 136 on the private, offline IQ test, which is the first time a large language model has crossed the 130-point 'genius' threshold.

QHow does the article describe the difference between the two sets of IQ tests used by Tracking AI?

ATracking AI uses two sets of tests: a publicly available Mensa Norway-style test that models have already scored highly on, and a private, offline question bank designed to prevent models from having seen the questions before, which is considered more difficult and cheat-proof.

QWhat practical examples are given in the article to demonstrate GPT-5.6's capabilities beyond test scores?

AThe article provides examples where GPT-5.6 successfully created a particle fluid simulation HTML file, built a RAG-based customer service ticket system with multiple features, and efficiently debugged a coding problem that other models failed to solve.

QWhat caution does the article mention about interpreting the IQ score of GPT-5.6?

AThe article cautions that the IQ test only measures a specific slice of intelligence, like abstract pattern recognition and logical reasoning, and does not assess a model's factual reliability, tool-use ability, or performance in real-world professional scenarios.

QWhat was a key distinction made between Claude-5 Fable and GPT-5.6 Sol in their approach to solving problems, according to developer feedback cited in the article?

AAccording to developer feedback, Claude-5 Fable was described as being overly focused on technical perfection, which could hinder practical problem-solving, while GPT-5.6 Sol was praised for its pragmatic approach that successfully got the job done.

Похожее

Впервые в истории: ETF, торгующий биткоинами на спотовом рынке, решил закрыться!

Компания Hashdex объявила о закрытии и ликвидации первого в США спотового биткоин-ETF — Hashdex Bitcoin (DEFI), торговавшегося на NYSE Arca. По состоянию на 30 июля 2026 года активы фонда составляли около $14,7 млн. Решение принято из-за низких активов по сравнению с операционными расходами, что сделало его долгосрочное существование нецелесообразным. Фонд инвестировал исключительно в физический биткоин. Инвесторы могут продавать акции до 17 августа 2026 года. После этой даты компания продаст оставшиеся биткоины, а не успевшие продать акции инвесторы получат денежную выплату около 28 августа 2026 года. Сумма выплаты будет зависеть от цены продажи биткоинов и затрат на ликвидацию, при этом возможны значительные колебания цены в процессе закрытия фонда.

cryptonews.ru8 мин. назад

Впервые в истории: ETF, торгующий биткоинами на спотовом рынке, решил закрыться!

cryptonews.ru8 мин. назад

Кук из ФРС заявила, что поддержит повышение ставок, если дезинфляция остановится

Член Совета управляющих Федеральной резервной системы США Лиза Кук заявила, что готова поддержать повышение процентных ставок, если инфляция в США не продолжит замедление. Выступая перед экономическими представителями Анкориджа, она подчеркнула, что, хотя некоторые дефляционные силы действуют, она «готова действовать», если процесс снижения инфляции остановится. Кук отметила, что инфляция остается слишком высокой, и риски для цели по инфляции в настоящее время перевешивают риски для сферы занятости. ФРС нацелена на долгосрочную инфляцию в 2%. По данным Trading Economics, годовая инфляция в июне 2026 года снизилась до 3,5%, что стало первым снижением за пять месяцев. Однако Кук призвала не придавать чрезмерного значения одному показателю, указав, что индекс цен на личное потребление (PCE) вырос на 3,7% за 12 месяцев до июня, что почти вдвое превышает целевой показатель. Она предупредила, что пять лет инфляции выше целевого уровня повышают риск укоренения высокой инфляции в поведении при установлении цен и зарплат, что сделает ее более устойчивой и сложной для сдерживания. «Если я не увижу признаков продолжения дезинфляции в ближайшее время, я готова действовать», — резюмировала Кук. Такие меры могут оказать давление на криптовалюты и другие высокорисковые инвестиции.

cointelegraph1 ч. назад

Кук из ФРС заявила, что поддержит повышение ставок, если дезинфляция остановится

cointelegraph1 ч. назад

Инвесторы начинают «менять игровой стол»

Китайские венчурные инвесторы начали активно переходить на работу в технологические компании, меняя «карточный стол». На фоне перегрева сектора hard tech (искусственный интеллект, робототехника, аэрокосмос) и взрывного роста финансирования конкуренция между фондами ужесточилась, а неопределённость с выходом из инвестиций и новое регулирование усиливают давление. Вместо того чтобы бороться за проекты в переполненных топовых сделках, многие опытные инвестиционные профессионалы, включая партнёров и управляющих директоров, целенаправленно выбирают переход в компании на позиции вице-президентов, руководителей по финансированию или даже сооснователей. Это движение стало осознанным выбором, а не вынужденным шагом. Компании-лидеры в AI и робототехнике, такие как Ziyuan Robot, Moon’s Dark Side, Minimax, активно нанимают таких специалистов для управления быстрыми и крупными раундами финансирования. Со стороны инвесторов это возможность получить более стабильную карьеру, прямой доступ к индустрии и привлекательные пакеты вознаграждения. Хотя переход требует адаптации от оценки к практическому исполнению, он отражает расширение карьерных границ в венчурной экосистеме.

marsbit1 ч. назад

Инвесторы начинают «менять игровой стол»

marsbit1 ч. назад

Торговля

Спот

Популярные статьи

Как купить S

Добро пожаловать на HTX.com! Мы сделали приобретение Sonic (S) простым и удобным. Следуйте нашему пошаговому руководству и отправляйтесь в свое крипто-путешествие.Шаг 1: Создайте аккаунт на HTXИспользуйте свой адрес электронной почты или номер телефона, чтобы зарегистрироваться и бесплатно создать аккаунт на HTX. Пройдите удобную регистрацию и откройте для себя весь функционал.Создать аккаунтШаг 2: Перейдите в Купить криптовалюту и выберите свой способ оплатыКредитная/Дебетовая Карта: Используйте свою карту Visa или Mastercard для мгновенной покупки Sonic (S).Баланс: Используйте средства с баланса вашего аккаунта HTX для простой торговли.Третьи Лица: Мы добавили популярные способы оплаты, такие как Google Pay и Apple Pay, для повышения удобства.P2P: Торгуйте напрямую с другими пользователями на HTX.Внебиржевая Торговля (OTC): Мы предлагаем индивидуальные услуги и конкурентоспособные обменные курсы для трейдеров.Шаг 3: Хранение Sonic (S)После приобретения вами Sonic (S) храните их в своем аккаунте на HTX. В качестве альтернативы вы можете отправить их куда-либо с помощью перевода в блокчейне или использовать для торговли с другими криптовалютами.Шаг 4: Торговля Sonic (S)С легкостью торгуйте Sonic (S) на спотовом рынке HTX. Просто зайдите в свой аккаунт, выберите торговую пару, совершайте сделки и следите за ними в режиме реального времени. Мы предлагаем удобный интерфейс как для начинающих, так и для опытных трейдеров.

1.9k просмотров всегоОпубликовано 2025.01.15Обновлено 2026.06.02

Как купить S

Sonic: Обновления под руководством Андре Кронье – новая звезда Layer-1 на фоне спада рынка

Он решает проблемы масштабируемости, совместимости между блокчейнами и стимулов для разработчиков с помощью технологических инноваций.

2.4k просмотров всегоОпубликовано 2025.04.09Обновлено 2025.04.09

Sonic: Обновления под руководством Андре Кронье – новая звезда Layer-1 на фоне спада рынка

HTX Learn: Пройдите обучение по "Sonic" и разделите 1000 USDT

HTX Learn — ваш проводник в мир перспективных проектов, и мы запускаем специальное мероприятие "Учитесь и Зарабатывайте", посвящённое этим проектам. Наше новое направление .

1.9k просмотров всегоОпубликовано 2025.04.10Обновлено 2025.04.10

HTX Learn: Пройдите обучение по "Sonic" и разделите 1000 USDT

Обсуждения

Добро пожаловать в Сообщество HTX. Здесь вы сможете быть в курсе последних новостей о развитии платформы и получить доступ к профессиональной аналитической информации о рынке. Мнения пользователей о цене на S (S) представлены ниже.

活动图片