Tuesday, September 22, 2026
No Result
View All Result
Blockchain 24hrs
  • Home
  • Bitcoin
  • Crypto Updates
    • General
    • Altcoins
    • Ethereum
    • Crypto Exchanges
  • Blockchain
  • NFT
  • DeFi
  • Metaverse
  • Web3
  • Blockchain Justice
  • Analysis
Crypto Marketcap
  • Home
  • Bitcoin
  • Crypto Updates
    • General
    • Altcoins
    • Ethereum
    • Crypto Exchanges
  • Blockchain
  • NFT
  • DeFi
  • Metaverse
  • Web3
  • Blockchain Justice
  • Analysis
No Result
View All Result
Blockchain 24hrs
No Result
View All Result

xAI Ships Grok 4.7 With New Safeguard Stack: Independent Benchmarks Confirm Gains, Flag Doubled Token Consumption

Home Metaverse
Share on FacebookShare on Twitter


by
Alisa Davidson


Printed: September 22, 2026 at 5:45 am Up to date: September 22, 2026 at 5:45 am

by Anastasiia O


Edited and fact-checked:
September 22, 2026 at 5:45 am

To enhance your local-language expertise, typically we make use of an auto-translation plugin. Please observe auto-translation might not be correct, so learn unique article for exact info.

In Temporary

xAI launches Grok 4.7 with frontier coding and agentic work efficiency, a brand new security stack, and $2/$6 per million token pricing. Impartial benchmarks verify the features however flag doubled token use.

xAI Ships Grok 4.7 With New Safeguard Stack: Independent Benchmarks Confirm Gains, Flag Doubled Token Consumption

xAI has launched Grok 4.7, its most succesful mannequin so far for coding and information work, positioning it as a frontier providing at a aggressive value level. The mannequin is constructed on a brand new, bigger base than its predecessor, Grok 4.6, and was educated with an prolonged reinforcement studying run on a more durable mixture of duties weighted towards issues that require many hours to finish. In keeping with the corporate, these enhancements improve the mannequin’s code technology, long-context dealing with, and self-verification capabilities, whereas native understanding of the Grok Bot harness makes it simpler at conversational duties and normal information work.

Benchmark outcomes illustrate the place the features are most pronounced. On CursorBench 4.0, which stresses longer-running coding duties, Grok 4.7 posted a rating of 46.3%, up from 40.4% on Grok 4.6 and forward of GPT-5.6 Sol’s 41.7%, whereas remaining behind Fable 5.1’s 51.8%. The mannequin additionally recorded a high-effort rating of 71.0% on DeepSWE v1.1, trailing GPT-5.6 Sol’s 72.7% however edging Fable 5.1’s 70.0%. 

In skilled information work, Grok 4.7 improved on its predecessor throughout GDPval and AA Briefcase, the latter of which evaluates multi-hour workplace duties carried out by professionals similar to attorneys, nurses, and monetary analysts, the place it scored 1,657, shut behind Fable 5.1’s 1,678. Essentially the most placing bounce got here on Terminal-Bench 4.0, a measure of multi-hour terminal work, the place the mannequin almost doubled Grok 4.6’s efficiency, rising from 20.3 to 38.0%. On HackerBench v0.3, which covers dangerous cyber duties, the mannequin allowed solely 3.3% of harmful dual-use prompts by means of whereas not often blocking professional safety work.

A distinguishing ingredient of the discharge is Grok 4.7’s security stack, described by the corporate as completely new and the strongest it has examined on refusals and jailbreak resistance. The mannequin leads on LatchBio’s biosafety benchmark at 62.4%, balancing utility for benign duties with secure refusal on harmful ones in dual-use domains similar to cybersecurity and organic analysis. xAI has additionally begun granting choose cybersecurity companions invite-only entry to the mannequin’s red-team capabilities for protection analysis.

Pricing is a central a part of the launch technique. Grok 4.7 is served on the similar charges as Grok 4.6, two {dollars} per million enter tokens and 6 {dollars} per million output tokens, undercutting GPT-5.6 Sol (4 and twenty {dollars} respectively) and Fable 5.1 (ten and fifty {dollars}). A quick variant providing twice the output pace is on the market at twice the value. On CursorBench’s price-performance frontier, this locations Grok 4.7 among the many most cost-efficient choices for prolonged coding workloads.

The mannequin is on the market instantly in Cursor and Grok Construct, with entry additionally supplied by means of the Grok API, third-party coding harnesses, and mannequin routers and cloud platforms. The discharge intensifies competitors within the frontier mannequin phase, the place distributors are more and more differentiating on long-duration agentic duties, security calibration, and price per accomplished activity relatively than uncooked benchmark scores alone. For builders and enterprises evaluating coding-oriented fashions, Grok 4.7 presents a mix of pricing stability, improved multi-hour activity efficiency, and hardened safeguards that might make it a viable different to costlier frontier choices.

Impartial Evaluation Confirms Frontier Standing, With a Token-Consumption Caveat

Impartial benchmarking agency Synthetic Evaluation has corroborated xAI’s claims, scoring Grok 4.7 at 46 on its Intelligence Index, a two-point achieve over Grok 4.6, a end result that locations SpaceXAI among the many prime 4 AI labs. The agency’s analysis, carried out at xhigh reasoning effort, discovered the mannequin’s clearest advance in long-horizon agentic information work: on its non-public AA-Briefcase benchmark, Grok 4.7 gained 111 Elo over its predecessor to achieve 1,657, rating alongside Claude Opus 5 and Claude Fable 5.1 on the frontier. The advance was pushed primarily by analytical high quality, which rose sharply to 1,994 Elo from 1,690, whereas presentation high quality held roughly regular. On GDPval-AA, which measures sensible work merchandise similar to paperwork, spreadsheets, and slides, the mannequin scored 1,695 Elo, up 90 factors.

Grok 4.7 scores 46 on the Synthetic Evaluation Intelligence Index to convey SpaceXAI into the highest 4 AI labs. Coding Agent Index efficiency has additionally improved, overtaking GPT-5.6 Sol

Grok 4.7 scores +2 factors over Grok 4.6 on the Intelligence Index, with robust efficiency on… pic.twitter.com/8p4KC6cgZy

— Synthetic Evaluation (@ArtificialAnlys) September 21, 2026

Efficiency features got here at a price in compute consumption. Synthetic Evaluation measured roughly 81,000 output tokens per Intelligence Index activity for Grok 4.7 at xhigh, greater than double the 36,000 utilized by Grok 4.6 and almost triple the 27,000 consumed by GPT-6 Astra at max effort. The agency additionally famous incremental modifications elsewhere: modest enhancements on Terminal-Bench 4.0 and GDP.pdf, alongside small regressions on AA-LCR and AutomationBench-AA. Reliability metrics improved modestly, with the AA-Omniscience hallucination charge falling to 29% from 34%, whereas accuracy held near-unchanged at 47%. Technical specs stay in keeping with the earlier technology, together with an unchanged 500,000-token context window and cache-hit pricing discounted to $0.50 per million tokens.

Disclaimer

In step with the Belief Venture pointers, please observe that the knowledge supplied on this web page just isn’t meant to be and shouldn’t be interpreted as authorized, tax, funding, monetary, or another type of recommendation. You will need to solely make investments what you may afford to lose and to hunt unbiased monetary recommendation if in case you have any doubts. For additional info, we recommend referring to the phrases and situations in addition to the assistance and help pages supplied by the issuer or advertiser. MetaversePost is dedicated to correct, unbiased reporting, however market situations are topic to vary with out discover.

About The Writer


Alisa, a devoted journalist on the MPost, focuses on crypto, AI, investments, and the expansive realm of Web3. With a eager eye for rising developments and applied sciences, she delivers complete protection to tell and have interaction readers within the ever-evolving panorama of digital finance.

Extra articles


Alisa, a devoted journalist on the MPost, focuses on crypto, AI, investments, and the expansive realm of Web3. With a eager eye for rising developments and applied sciences, she delivers complete protection to tell and have interaction readers within the ever-evolving panorama of digital finance.








Extra articles





Source link

Tags: BenchmarksConfirmconsumptionDoubledFlaggainsGrokindependentsafeguardShipsStacktokenXai
Previous Post

Polymarket Ignored Warnings During $10M Fraud Attack

Next Post

Capital.com Appears to Be Making a Move for the UK’s Crypto Market

Related Posts

Gate Update: CoinGecko Confirms Liquidity Lead as Gate Launches Stock Event Contracts and Deepens Multi-Asset Infrastructure
Metaverse

Gate Update: CoinGecko Confirms Liquidity Lead as Gate Launches Stock Event Contracts and Deepens Multi-Asset Infrastructure

September 20, 2026
Canopy Ships Banyan Protocol Upgrade Live, With No Downtime
Metaverse

Canopy Ships Banyan Protocol Upgrade Live, With No Downtime

September 17, 2026
Top 10 DePIN Projects Bringing Real-World Infrastructure Onchain
Metaverse

Top 10 DePIN Projects Bringing Real-World Infrastructure Onchain

September 16, 2026
OpenSim stats all up with the fall season – Hypergrid Business
Metaverse

OpenSim stats all up with the fall season – Hypergrid Business

September 17, 2026
How Tokenization Is Reshaping Music, IP, And Licensing Markets
Metaverse

How Tokenization Is Reshaping Music, IP, And Licensing Markets

September 13, 2026
New Meta VR glasses leaked – Hypergrid Business
Metaverse

New Meta VR glasses leaked – Hypergrid Business

September 13, 2026
Next Post
Capital.com Appears to Be Making a Move for the UK’s Crypto Market

Capital.com Appears to Be Making a Move for the UK's Crypto Market

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Facebook Twitter Instagram Youtube RSS
Blockchain 24hrs

Blockchain 24hrs delivers the latest cryptocurrency and blockchain technology news, expert analysis, and market trends. Stay informed with round-the-clock updates and insights from the world of digital currencies.

CATEGORIES

  • Altcoins
  • Analysis
  • Bitcoin
  • Blockchain
  • Blockchain Justice
  • Crypto Exchanges
  • Crypto Updates
  • DeFi
  • Ethereum
  • Metaverse
  • NFT
  • Regulations
  • Web3

SITEMAP

  • About Us
  • Advertise With Us
  • Disclaimer
  • Privacy Policy
  • DMCA
  • Cookie Privacy Policy
  • Terms and Conditions
  • Contact Us

Copyright © 2024 Blockchain 24hrs.
Blockchain 24hrs is not responsible for the content of external sites.

  • bitcoinBitcoin(BTC)$86,118.001.08%
  • ethereumEthereum(ETH)$2,757.620.86%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$790.09-0.13%
  • rippleXRP(XRP)$1.543.68%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$117.610.30%
  • tronTRON(TRX)$0.3454840.38%
  • zcashZcash(ZEC)$1,518.65-1.42%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.021.39%
No Result
View All Result
  • Home
  • Bitcoin
  • Crypto Updates
    • General
    • Altcoins
    • Ethereum
    • Crypto Exchanges
  • Blockchain
  • NFT
  • DeFi
  • Metaverse
  • Web3
  • Blockchain Justice
  • Analysis
Crypto Marketcap

Copyright © 2024 Blockchain 24hrs.
Blockchain 24hrs is not responsible for the content of external sites.