GetChain News
中简 中繁 EN
GetChain News
Toggle sidebar

Online/Update

News linked to both this project and an event.

Bank of America: Anthropic leads frontier AI model rankings, token prices ease but GPU rents remain firm

According to TechFlow Research, the frontier AI data tracking report released by Bank of America Securities on August 17 shows that Anthropic leads comprehensively in three major AI benchmarks, with Claude Opus 5 ranking first in the Intelligence Index, Agent Index, and Coding Agent Index, GPT-5.6 Sol following closely behind, Meta MuseSpark 1.2 entering the top ten, and Google Gemini 3.6 Flash ranking outside the top ten. In terms of usage, DeepSeek leads with approximately 30% of the Vercel platform token share, Anthropic accounts for 25% and OpenAI accounts for 16%; but in terms of payment amount, Anthropic leads far ahead with 65%, while OpenAI accounts for only 11%. In terms of pricing, the AI Token Price Index decreased 9% month-over-month in August to $2.21, but still increased 87% year-over-year; GPU rental rates remain strong, with H100 increasing 33% year-over-year to $2.77/hour, DRAM increasing 483% year-over-year, and NAND increasing 432% year-over-year. The research report judges that AI infrastructure demand remains healthy, with open-source model usage growing but payment share still highly concentrated on top closed-source models. BofA believes that the Meta "Watermelon" and Google Gemini 4 releases, token pricing trends and GPU rental trends are

DeepSeek V4-Flash Dominates Competitors in Cost-Performance, Operating Costs Only 1/105 of Claude's

According to Reuters, the latest V4-Flash API model officially released by Chinese AI startup DeepSeek on July 31 incurred operating costs of only 1/105 that of Anthropic Claude Fable 5 in benchmark tests conducted by AI performance analysis agency Artificial Analysis. Regarding specific pricing, V4-Flash input token costs are $0.14 per million, and output token costs are $0.28 per million, with an average cost per test of approximately 3 cents, far lower than Wenxin Kimi K3 (86 cents), OpenAI GPT-5.6 Sol ($1.86), and Claude Fable 5 ($3.15). In terms of performance, V4-Flash scored 50 points on the Comprehensive Intelligence Index, tying with Google Gemini 3.6 Flash, but still more than 9 points lower than leading models such as Claude Opus 5 and GPT-5.6. It is worth noting that a low listed price does not equate to low actual costs—if the model consumes more inference and output tokens when generating responses, actual expenses may rise significantly.

Bitgo CEO deposits ~$6.3M in BTC, challenges Claude to move the funds

: Bitgo CEO Mike Belshe deposited 100 BTC into a public Bitcoin address on August 1, worth approximately $6.3 million at the time, and invited Anthropic's Claude model to attempt to move the funds out of the address. On-chain records show the wallet received the funds on July 31, and the balance had not been transferred out as of August 2. Anthropic previously disclosed that during 141,006 cybersecurity assessment runs, 3 incidents were found, with 6 evaluation sessions involving 3 models inadvertently interacting with real organizational systems. The models involved include Claude Opus 4.7, Claude Mythos 5, and an unreleased internal research model. The cause was a configuration error by third-party testing partner Irregular, which led to the test environment being connected to the internet. Anthropic stated that Claude Opus 4.7, during one evaluation, located a real website with the same name as a simulated company, exploited weak passwords and exposed services to recover infrastructure credentials, and accessed a production database containing hundreds of records. The company said the model was attempting to complete assigned tasks, not actively breaking constraints or pursuing independent goals. Belshe's challenge involves Bitgo's institutional custody platform, which uses multi-signature or multi-party computation technology to distribute signing authority across multiple independent keys. As of August 2, Anthropic had not publicly responded to the challenge.

DeepSeek V4-Flash-0731 Public Beta: Code Agent Capabilities Significantly Enhanced, Price Unchanged

DeepSeek officially launches the public beta of version V4-Flash-0731. Through post-training and Agent framework optimization, DeepSWE benchmark scores have improved from 7.3 to 54.4, Cybergym scores have doubled to 76.7, and performance on some hard benchmarks approaches Claude Opus 4.8. The new version does not increase parameters, prices remain unchanged, it natively supports the Responses API, and existing user APIs will automatically switch to the new version. Currently, only the Flash API is updated; the V4 Pro official version is still under development.

Anthropic discloses three cybersecurity incidents involving Claude, where the model accessed real systems and obtained unauthorized permissions

Odaily News, July 30 – Anthropic released a report stating that during a review of cybersecurity assessment records, three incidents were discovered in which the Claude model accessed the internet in a third-party evaluation environment and further obtained unauthorized access to three real organizations' systems.Anthropic stated that the review covered 141,000 evaluation runs that could have potentially gained network access, and a total of three related incidents were found. All incidents occurred during Capture The Flag (CTF) cybersecurity tests, where the model was told the environment was a simulation with no internet access; however, due to configuration errors by the evaluation partner, the actual environment had internet connectivity.Among these, Claude Opus 4.7 accessed real company infrastructure during one test and obtained database permissions containing hundreds of production data records; Claude Mythos 5 built a malicious Python package and uploaded it to PyPI, resulting in the package being downloaded and run on 15 real systems; another internal research test model scanned approximately 9,000 targets and accessed a company's internet application through a public vulnerability.Anthropic stated that these incidents were not cases of the model actively seeking to escape or pursue its own goals, but rather the model mistakenly believed the real systems were within the test scope and continued executing the assigned cyberattack tasks. Notably, the newer internal research model stopped attacking after identifying that the targets might be real systems.Anthropic stated that these incidents primarily reflect issues with evaluation environment isolation and operational processes, rather than model alignment failures. The company has suspended related cybersecurity assessments, strengthened security controls in evaluation environments, continuously monitored test records, and will collaborate with third-party organizations to conduct further reviews.

Perplexity CEO: GLM 700B Parameter Model Performance Close to Opus Level

Perplexity CEO Aravind Srinivas tweeted that GLM is a severely underrated model, demonstrating performance close to Opus-level at the 700B parameter scale with extremely high operational efficiency. Previously, Moonshot AI just released the open-source model Kimi K3, and industry attention on the GLM series models under Zhipu AI is rising.

Kimi K3 Released, Gap Between Open-Source and Closed-Source Models Narrows to 4 Points

Moonshot AI launches open-source model Kimi K3, scoring 57 points on the Artificial Analysis Intelligence Index, becoming the third highest-scoring model, second only to Anthropic's Claude Opus 5 (61 points), Claude Fable 5 (60 points), and OpenAI's GPT-5.6 Sol (59 points). The index shows that the gap between leading closed-source models and open-weight models has narrowed to 4 points, the smallest gap since the release of GLM-5 in February. This signifies that open-source models are rapidly catching up to closed-source models in performance.

Claude Opus 5 Officially Available on B.AI API, Dual-Channel Access Balances Performance and Cost

The B.AI API platform officially launches the Claude Opus 5 model. This model provides a context window of up to 1 million tokens and a single output limit of 128,000 tokens, achieving significant performance leaps in scenarios such as complex reasoning, long code engineering development, large-scale document analysis, and knowledge-intensive workflows. To meet diverse deployment needs, this release simultaneously opens a dual-channel API access solution on the B.AI platform: in addition to the official standard interface, a new "Designated Service Provider" cooperation channel is added, through which enterprise users can obtain cost discounts of up to 40%, facilitating flexible balancing of performance and budget expenditure based on actual business loads. Developers can log in to the official B.AI platform starting today to experience the exceptional capabilities of Claude Opus 5 first.

Claude Opus 5 Tops Artificial Analysis Intelligence Index v4.1

Anthropic's Claude Opus 5 (max and xhigh versions) achieved the highest intelligence score on the Artificial Analysis Intelligence Index v4.1. The index integrates 9 evaluation benchmarks including GDPval-AA v2, GPQA Diamond, and Humanity's Last Exam. Previously, Opus 5 has demonstrated leading performance in multiple benchmark tests, including scoring twice that of the previous generation on Frontier-Bench, topping the AA-Briefcase agent benchmark, and outperforming Fable 5 in cost-performance ratio.

Anthropic Launches Opus 5 AI Model, Performance Close to Fable 5, Price Halved

Anthropic has officially released the Opus 5 AI model; the company states that its performance is close to the frontier model Fable 5, but the price is only half that of the latter.

Polymarket: "Next Claude Opus model will be released on July 23, 2026" probability currently at 61%, up 51% in 24 hours

PPP Prediction Market Tool monitoring shows that on Polymarket, the probability of "the next Claude Opus model will be released on July 23, 2026" is currently at 61%, up 51% in 24 hours; additionally, the probability of "release before July 24" is at 9%; the probability of "release before July 25" is at 5%;This event will be settled based on the date (Eastern Time) when Anthropic's next Claude Opus model becomes available to the public. Only models explicitly named "Opus" by Anthropic are valid, and they must be accessible to the public (including public beta testing or open waitlist). Other models such as Sonnet and Haiku are not counted. The final settlement will primarily rely on official Anthropic information.Join the PPP Signal Push Community to stay ahead and seize the opportunity.

月之暗面将发布新模型 Kimi K3,或超越 Opus 4.8

据两名知情人士透露,月之暗面计划在未来几天内发布 Kimi K3,该模型参数规模达 2 万亿至 3 万亿,将成为中国迄今最大的 AI 模型。Anthropic 尚未披露旗下模型的参数规模,但业内人士普遍推测,Opus 4.8 拥有约 1.5 万亿至 2 万亿个参数。

SpaceXAI and Cursor Plan to Launch First Joint AI Model

According to Reuters citing The Information, SpaceXAI and Cursor plan to release their first jointly developed artificial intelligence model as early as Wednesday. The model was originally scheduled for launch early this week but was delayed for efficiency optimization. The report noted that the new model is expected to possess strong rapid information processing capabilities, with some performance aspects potentially competitive with Anthropic's Opus 4.8 and OpenAI's GPT 5.5.

GPT-5.6 rumored to be publicly available as early as July 7, Gemini 3.5 Pro may launch on July 17

Tech blogger Leo has revealed that OpenAI may make GPT-5.6 available to the public between July 7 and July 9, with the earliest possible date being July 7. The new model's plan usage quotas are said to be more generous, and OpenAI is further strengthening its safety strategies ahead of the launch.Additionally, according to sources, Google DeepMind has tentatively scheduled the release of Gemini 3.5 Pro for July 17. Another tech blogger, Astro Polo, stated that Gemini 3.5 Pro will support a 2 million token context window, doubling the 1 million token context window currently supported by Claude Sonnet 5, Claude Opus 4.8, and Claude Fable 5. This makes it more suitable for handling large codebases, lengthy documents, and long conversations.Note: The above information is based on market rumors and has not yet been officially confirmed by OpenAI or Google DeepMind.

Meituan Releases Trillion-Parameter Large Model LongCat-2.0, the First Trillion-Parameter Model to Complete Full-Process Training on a Domestic Computing Cluster

According to Meituan's official release, Meituan has officially launched the new generation large model LongCat-2.0 and open-sourced it simultaneously. The model features a total of 1.6T parameters, making it the industry's first trillion-parameter model to complete full-process training and inference on a domestic computing cluster of 50,000 cards. It natively supports 1M ultra-long context and focuses primarily on code understanding, generation, and execution in Agentic Coding scenarios. Technically, LongCat-2.0 adopts the LongCat Sparse Attention (LSA) sparse attention mechanism, reducing long text computation complexity from quadratic to linear; achieves token-level dynamic activation (33B~56B) via a zero-computation expert mechanism; and introduces the MOPD architecture to fuse three sets of expert capabilities: Agent, Reasoning, and Interaction. In terms of training efficiency, the team spent three years overcoming challenges in adapting to domestic computing power, reducing the monthly average daily failure rate by over 70%, increasing training MFU by 1.5 times, and achieving steady-state daily throughput exceeding 1T tokens/day. In terms of performance evaluation, LongCat-2.0 achieved a score of 59.5 on SWE-bench Pro, surpassing Gemini 3.1 Pro (54.2), GPT-5.5 (58.6), and Claude Opus 4.6 (57.3); and achieved a score of 79.9 on BrowseComp.

Polymarket probability of "Claude Fable 5 restored for US customers before July 1" rises to 73%, up 33% in 24H

Monitoring by Odaily Seer Prophet Channel shows that the probability of "Claude Fable 5 restored for US customers before July 1" on Polymarket has risen to 73%, up 33% in 24H.If Anthropic reopens Claude Fable 5 (or Claude Mythos, or a version confirmed to be the same model) to the US public before the specified date, this event will settle as "Yes"; otherwise, it will settle as "No". Qualifying methods of restoration include public beta or public waitlist; closed testing and private access do not count. Other models such as Haiku, Sonnet, and Opus are not included by default unless they are confirmed to be the same model as Claude Fable 5. Settlement will primarily be based on official announcements from Anthropic, supplemented by consensus reports from mainstream media.Due to export controls imposed by the US government on national security grounds, Anthropic urgently suspended global access to the Claude Fable 5 model just days after its release on June 9. Although the restrictions were primarily aimed at foreign users, due to the difficulty of real-time user identity screening, Anthropic ultimately closed access for all users, including those in the US, and switched requests to less capable models such as Opus 4.8. Anthropic is currently engaged in high-level discussions with the White House on issues including model capabilities and potential jailbreak risks. The company has publicly stated that the ban may have stemmed from a misunderstanding and is actively pushing to restore access.Odaily Seer Prophet Channel continues to monitor the prediction market, seeing changes before they are priced in.

Zhipu Founder Tang Jie: Open-Source Release of GLM-5.2 Will Gradually Narrow the Performance Gap with OpenAI and Anthropic

: Zhipu AI founder Tang Jie posted on the X platform, stating that since the official open-source release of GLM-5.2, it has achieved leading results in multiple international authoritative evaluations and competitive rankings. In the Artificial Analysis Intelligence Index comprehensive evaluation, GLM-5.2 scored 51 points, placing it in the same range as Anthropic's Claude Opus 4.8. In the Code Arena front-end code generation adversarial test, it ranked 2nd globally with an Elo of 1595, and in the DesignArena design and code integration scenario, it scored 1360 points, ranking 1st.Overall, Zhipu GLM-5.2 continues to rank among the top globally in real-world scenario evaluations across multiple areas, including front-end development, design generation, and software engineering. It is steadily narrowing the performance gap with cutting-edge models from OpenAI and Anthropic and will continue to push the upper limits of model capabilities.Previously, in response to Musk's statement that Chinese large models might reach Anthropic's Fable level by the first quarter of next year, Zhipu AI founder Tang Jie replied, "It won't take that long."

Anthropic Launches Phase Two of Project Fetch: Claude Opus 4.7 Achieves 10x Speedup in Robotics Tasks

Anthropic has released the results of Phase Two of “Project Fetch,” evaluating the enhanced capabilities of its latest model in real-world robotic manipulation. Conducted in August 2025, the experiment tasked non-robotics-expert Anthropic employees with completing a series of complex tasks using off-the-shelf quadruped robots. Performance was compared between two conditions: “using Claude models for assistance” versus “relying solely on humans and the internet.” Results show that, under fully autonomous operation by the latest model—Claude Opus 4.7—the robot achieved significantly faster average completion times across all successfully executed tasks than the human team, with execution speed improved by at least 10×.

Zcash Founder Says Claude Mythos Audit Found No Critical Vulnerabilities

Odaily Zcash founder Zooko Wilcox posted on X stating that a security audit conducted by Anthropic's Claude Mythos AI model did not find any "more severe vulnerabilities" in the Zcash protocol. The audit was commissioned by Shielded Labs, a Swiss non-profit organization supporting Zcash development. On June 3, Zcash developers temporarily paused Orchard transactions after discovering a vulnerability in the shielded pool, restoring functionality through an emergency upgrade the same day. The issue stemmed from a four-year-old forging vulnerability in the Orchard shielded pool, identified by security researcher Taylor Hornby with the assistance of Anthropic's Claude Opus 4.8 model. The Zcash Foundation stated there is no evidence that the vulnerability was exploited, nor was any unauthorized value creation detected, and user privacy remained unaffected.Anthropic released the first public version of the Claude Mythos model, Fable 5, on Tuesday, and stated on Friday that it has suspended access to the Fable 5 and Mythos 5 AI models due to export control directives issued by the U.S. government citing national security concerns. (Cointelegraph)

Claude Announces Suspension of Fable 5 Access Due to U.S. Government Ban; Developers Urged to Switch Models Promptly

Claude announced that, due to U.S. government-related restrictions, access to the Fable 5 model will be suspended for all users. New sessions will default to the user’s configured default model or Opus 4.8; existing Fable 5 sessions will terminate immediately with an error. Additionally, API requests to Fable 5 on the Claude Platform will also return errors. Developers must promptly migrate their applications and integrations to other Claude models.