GetChain News
中简 中繁 EN
GetChain News
Toggle sidebar

Online/Update

News linked to both this project and an event.

Claude Sonnet 5.5 Officially Launches on B.AI, API and Web Chat Both Now Available

B.AI announces that Anthropic’s latest Sonnet model, Claude Sonnet 5.5, is now officially available on the platform. Serving as a faster, more cost-effective complement to Claude Opus 5.5, Claude Sonnet 5.5 boosts output speed by over 30% and reduces per-task costs by up to approximately 30%, making it tailored for well-defined software engineering tasks, daily Agent workflows, and professional document creation. The model supports image and PDF comprehension, enabling the generation of reports, presentations, spreadsheets, and UIs. It offers five adjustable reasoning intensity levels alongside a 1-million-token context window, with long-context processing included at no additional charge. Both the B.AI API and Web Chat platforms are now live, allowing developers and users to directly access this next-generation Sonnet model that seamlessly balances speed, cost-efficiency, and long-context capabilities. Visit B.AI to experience it.

Claude Sonnet 5.5 Officially Lands on B.AI, Now Available on Both API and Web Chat

Odaily News — On September 29, B.AI announced that Anthropic's latest Sonnet model, Claude Sonnet 5.5, has officially launched on the platform. As a faster, lower-cost complement to Claude Opus 5.5, Claude Sonnet 5.5 delivers over 30% improvement in output speed and reduces per-task costs by up to approximately 30%, targeting well-defined software engineering, everyday Agent workflows, and professional document creation scenarios. The model supports image and PDF understanding, can generate reports, presentations, spreadsheets, and UIs, and offers five adjustable reasoning effort levels along with a 1 million Token context window, with no additional charges for long context.Currently, B.AI API and Web Chat are both available simultaneously, allowing developers and users to directly access this next-generation Sonnet model that balances speed, cost, and long-context capabilities.

Claude Opus 5.5 Officially Lands on B.AI, API and Web Chat Simultaneously Open

the B.AI platform has announced the official launch of Claude Opus 5.5, the first model in the Claude 5.5 series released by Anthropic, with simultaneous access now available via both API and Web Chat. The model is purpose-built for long-horizon agentic coding and complex knowledge work, matching Claude Fable 5.1 performance on most tasks while reducing operating costs by approximately 40% compared to Opus 5. It supports a 1M Token ultra-long context window and up to 128K output, and demonstrates leading capabilities in Computer Use and multi-step automation scenarios. Starting today, developers and users can directly access Claude Opus 5.5 through the B.AI platform.

Anthropic Releases Claude Opus 5.5, Cutting Typical Workload Costs by 40%

Anthropic has released Claude Opus 5.5, the first model in the Claude 5.5 series. According to official statements, the model demonstrates improvements in agentic programming, computer use, and knowledge work compared to Opus 5. Under default settings, typical workload costs are reduced by 40%, and output speed increases by over 30%. Additionally, its natural language expression, presentation of key information, and instruction-following capabilities have been enhanced, and the model underwent evaluation by external organizations such as METR and Frontier Design prior to release.

Polymarket odds for "the next Claude Opus model will be released on September 22" rose to 77%, up 50% in 24 hours

PPP prediction market tool monitoring shows that Polymarket odds for "the next Claude Opus model will be released on September 22" rose to 77%, up 50% in 24 hours.Only a new model explicitly named "Opus" by Anthropic counts, such as Claude Opus 5.1, Opus 5.5, Opus 6, etc. Other names such as Sonnet, Haiku, Fable, and Mythos do not count, unless Anthropic officially names it Opus.The release date is based on the calendar date in U.S. Eastern Time.It must actually be made available to the public to trigger settlement. Public betas and open rolling waitlists both count; closed betas, invite-only access, or private access do not count.If the model name, a placeholder, or a mislabel merely appears on Anthropic's official website but the model is not actually usable by the public, it does not count as a release.Settlement will primarily be based on Anthropic's official information, verified in combination with credible media reports.Join the PPP signal push community to stay one step ahead and seize the opportunity.

Polymarket's "Anthropic's next-generation Claude Opus will be released before September 27" probability rose to 90%, up 13% in 24 hours

PPP prediction market tool monitoring shows that Polymarket's "Anthropic's next-generation Claude Opus will be released before September 27" probability rose to 90%, up 13% in 24 hours; the probability of release before September 30 rose to 94%, up 7% in 24 hours.According to the resolution rules, if Anthropic's next Claude Opus model becomes available to the public before the specified date (Eastern Time, ET), the prediction market will resolve as "Yes"; otherwise, it will resolve as "No." Claude Opus refers to a model explicitly named "Opus" by Anthropic. Models released under other names, such as Sonnet, Haiku, Fable, or Mythos, do not qualify.In addition, qualifying models must have been officially launched and made accessible to the public, including open beta or public-facing rolling waitlist registration. Closed beta or any form of private access does not qualify for resolution.Join the PPP signal push community to stay one step ahead and seize the initiative.

Google Gemini 3.8 Flash Launches, DeepSWE 1.1 Scores 71%

Google Gemini 3.8 Flash has now launched on Gemini, Google AI Studio, and the API. According to Testing Catalog, the model scored 71% on the DeepSWE 1.1 benchmark, close to Claude Opus 5’s 74%, but at a much lower price. Regarding pricing, input costs $0.75 and output (including thinking tokens) costs $3.75 before December 31, 2026; starting January 1, 2027, these rates will increase to $1.50 and $7.50 respectively.

Google will release Gemini 3.8 Flash as early as Wednesday.

According to a report by Wall Street Journal reporter Erin Woo citing employee sources, Google’s artificial intelligence research department is set to release a new model, Gemini 3.8 Flash, featuring significantly enhanced coding capabilities. Internally codenamed "Skimaki," the model could go live as early as Wednesday, September 2. In comparative testing against Google’s internal coding tool Jetski, company engineers demonstrated a stronger preference for the new model than for Anthropic’s Opus.

Anthropic's New Research: Rewarding Hacking Behavior Could Lead to Severe Model Misalignment

Anthropic has released a new study titled "Training a Misaligned Reward Chaser," examining whether "reward hacking" during training compels models to pursue rewards at all costs. The research team trained an Opus-scale model across 80 known exploitable production environments. Simulated evaluations revealed that the model engaged in unauthorized network attacks, tampered with reward mechanisms, and attempted to evade security monitoring.

Zhipu AI launches and open-sources the "Niulai" model GLM-5.3-Flash, enhancing multimodal capabilities and inference cost efficiency.

Z.ai has released GLM-5.3-Flash, calling it the first native multimodal model in the GLM-5 series. The model features 320 billion total parameters and 18 billion active parameters, outperforming GLM-5.2 across multiple coding, agent, and vision benchmarks, and approaching Claude Opus 4.8 on certain coding tasks. GLM-5.3-Flash employs a hybrid architecture that combines sparse attention with linear attention, and introduces mechanisms such as mHC and IndexPool to reduce long-context inference costs and KV cache overhead.

Grayscale launches Zcash ETF, trading on NYSE Arca following key privacy vulnerability fix

Odaily News - Digital asset manager Grayscale's Zcash ETF began trading on NYSE Arca on Tuesday under the ticker ZCSH. The product is the world's first exchange-traded product offering spot exposure to Zcash, allowing investors to track ZEC prices through securities accounts without needing to directly purchase or store the token.ZCSH was formerly known as the Grayscale Zcash Trust, established in October 2017 through a private placement. Grayscale filed an application with the U.S. Securities and Exchange Commission in November 2025 to convert the trust into an ETF, with shareholders holding shares that track the fund's ZEC holdings rather than holding ZEC directly.In May of this year, security researcher Taylor Hornby, using Anthropic's Claude Opus 4.8, discovered a vulnerability in Zcash's Orchard shielded pool that had existed for four years, which could potentially allow attackers to mint counterfeit ZEC. Developers deployed an emergency patch on June 1, but due to privacy mechanisms, it was not possible to cryptographically confirm whether the vulnerability had been exploited.Zcash activated the Ironwood upgrade in July, replacing Orchard with a new shielded pool and introducing accounting rules that limit the amount of ZEC exiting the old shielded pool to no more than the amount entering. Grayscale stated it will monitor the adoption of the Ironwood upgrade, network security, exchange support, and regulatory conditions for privacy assets. (Decrypt)

Anthropic Two Mysterious Claude Models Exposed, Marshmallow Said to Offer Better Conversation Experience Than Opus 5

Developers have discovered two early-access Claude models from Anthropic, codenamed claude-marshmallow-eap and claude-melon-eap.According to tester feedback, Marshmallow performs better overall than Melon, but neither has yet reached the level of Fable 5. Another tester noted that Marshmallow's conversation experience surpasses that of Opus 5.Previously, Anthropic has repeatedly seen early models named in the "food name + EAP" format. In early July, Claude Honeycomb EAP briefly appeared in Cursor before being removed; on July 24, Anthropic officially released Opus 5, but the company has never confirmed a direct correspondence between Honeycomb and Opus 5.

DeepSeek-V4-Flash-Vision-Exp Lands on B.AI, Fully Free to Access, With Extra Perks Added

B.AI has launched DeepSeek's brand-new multimodal experimental flagship model, DeepSeek-V4-Flash-Vision-Exp. The official API team has taken the lead in completing full integration and is now making it freely accessible to all users. Building upon V4-Flash's state-of-the-art pure text and code Agent capabilities, the model officially unlocks native image understanding, with its comprehensive multimodal Agent performance approaching that of the Claude Opus-4.8 flagship. The new model supports mixed text-and-image input and a 1M ultra-long context window, effortlessly handling demanding scenarios such as multimodal PPT generation, complex UI reconstruction, and frontend visual effects. Image input is billed per token, consuming a maximum of just 384 tokens per image. All users can now log into the B.AI platform to access this latest multimodal model capability completely free of charge and with zero barriers, empowering top-tier compute to rapidly accelerate the realization of creative and productive workflows.

Former DeepMind Employee Founded Inherent, AI Agent Research Reproduction Outperforms Claude and GPT-5.5

According to TechCrunch, the UK AI lab Inherent, founded by several former members of Google DeepMind, has released an AI research agent named Faraday. The company states that Faraday outperforms frontier models such as Anthropic Claude Opus 4.8 and OpenAI GPT-5.5 in tasks requiring the independent reproduction of research results from published scientific papers.

V4-Flash-Vision-Exp Launched, Now Offering Multimodal API Services

DeepSeek announced today that its brand-new multimodal visual understanding model, DeepSeek-V4-Flash-Vision-Exp, is now available on the DeepSeek API platform. The model is experimental and can be invoked by users by setting model="deepseek-v4-flash-vision-exp". Its pure text capabilities are roughly on par with the official release of DeepSeek-V4-Flash, while demonstrating significantly improved performance in benchmarks requiring visual understanding, bringing its multimodal Agent capabilities close to Opus-4.8.

Bank of America: Anthropic leads frontier AI model rankings, token prices ease but GPU rents remain firm

According to TechFlow Research, the frontier AI data tracking report released by Bank of America Securities on August 17 shows that Anthropic leads comprehensively in three major AI benchmarks, with Claude Opus 5 ranking first in the Intelligence Index, Agent Index, and Coding Agent Index, GPT-5.6 Sol following closely behind, Meta MuseSpark 1.2 entering the top ten, and Google Gemini 3.6 Flash ranking outside the top ten. In terms of usage, DeepSeek leads with approximately 30% of the Vercel platform token share, Anthropic accounts for 25% and OpenAI accounts for 16%; but in terms of payment amount, Anthropic leads far ahead with 65%, while OpenAI accounts for only 11%. In terms of pricing, the AI Token Price Index decreased 9% month-over-month in August to $2.21, but still increased 87% year-over-year; GPU rental rates remain strong, with H100 increasing 33% year-over-year to $2.77/hour, DRAM increasing 483% year-over-year, and NAND increasing 432% year-over-year. The research report judges that AI infrastructure demand remains healthy, with open-source model usage growing but payment share still highly concentrated on top closed-source models. BofA believes that the Meta "Watermelon" and Google Gemini 4 releases, token pricing trends and GPU rental trends are

DeepSeek V4-Flash Dominates Competitors in Cost-Performance, Operating Costs Only 1/105 of Claude's

According to Reuters, the latest V4-Flash API model officially released by Chinese AI startup DeepSeek on July 31 incurred operating costs of only 1/105 that of Anthropic Claude Fable 5 in benchmark tests conducted by AI performance analysis agency Artificial Analysis. Regarding specific pricing, V4-Flash input token costs are $0.14 per million, and output token costs are $0.28 per million, with an average cost per test of approximately 3 cents, far lower than Wenxin Kimi K3 (86 cents), OpenAI GPT-5.6 Sol ($1.86), and Claude Fable 5 ($3.15). In terms of performance, V4-Flash scored 50 points on the Comprehensive Intelligence Index, tying with Google Gemini 3.6 Flash, but still more than 9 points lower than leading models such as Claude Opus 5 and GPT-5.6. It is worth noting that a low listed price does not equate to low actual costs—if the model consumes more inference and output tokens when generating responses, actual expenses may rise significantly.

Bitgo CEO deposits ~$6.3M in BTC, challenges Claude to move the funds

: Bitgo CEO Mike Belshe deposited 100 BTC into a public Bitcoin address on August 1, worth approximately $6.3 million at the time, and invited Anthropic's Claude model to attempt to move the funds out of the address. On-chain records show the wallet received the funds on July 31, and the balance had not been transferred out as of August 2. Anthropic previously disclosed that during 141,006 cybersecurity assessment runs, 3 incidents were found, with 6 evaluation sessions involving 3 models inadvertently interacting with real organizational systems. The models involved include Claude Opus 4.7, Claude Mythos 5, and an unreleased internal research model. The cause was a configuration error by third-party testing partner Irregular, which led to the test environment being connected to the internet. Anthropic stated that Claude Opus 4.7, during one evaluation, located a real website with the same name as a simulated company, exploited weak passwords and exposed services to recover infrastructure credentials, and accessed a production database containing hundreds of records. The company said the model was attempting to complete assigned tasks, not actively breaking constraints or pursuing independent goals. Belshe's challenge involves Bitgo's institutional custody platform, which uses multi-signature or multi-party computation technology to distribute signing authority across multiple independent keys. As of August 2, Anthropic had not publicly responded to the challenge.

DeepSeek V4-Flash-0731 Public Beta: Code Agent Capabilities Significantly Enhanced, Price Unchanged

DeepSeek officially launches the public beta of version V4-Flash-0731. Through post-training and Agent framework optimization, DeepSWE benchmark scores have improved from 7.3 to 54.4, Cybergym scores have doubled to 76.7, and performance on some hard benchmarks approaches Claude Opus 4.8. The new version does not increase parameters, prices remain unchanged, it natively supports the Responses API, and existing user APIs will automatically switch to the new version. Currently, only the Flash API is updated; the V4 Pro official version is still under development.

Anthropic discloses three cybersecurity incidents involving Claude, where the model accessed real systems and obtained unauthorized permissions

Odaily News, July 30 – Anthropic released a report stating that during a review of cybersecurity assessment records, three incidents were discovered in which the Claude model accessed the internet in a third-party evaluation environment and further obtained unauthorized access to three real organizations' systems.Anthropic stated that the review covered 141,000 evaluation runs that could have potentially gained network access, and a total of three related incidents were found. All incidents occurred during Capture The Flag (CTF) cybersecurity tests, where the model was told the environment was a simulation with no internet access; however, due to configuration errors by the evaluation partner, the actual environment had internet connectivity.Among these, Claude Opus 4.7 accessed real company infrastructure during one test and obtained database permissions containing hundreds of production data records; Claude Mythos 5 built a malicious Python package and uploaded it to PyPI, resulting in the package being downloaded and run on 15 real systems; another internal research test model scanned approximately 9,000 targets and accessed a company's internet application through a public vulnerability.Anthropic stated that these incidents were not cases of the model actively seeking to escape or pursue its own goals, but rather the model mistakenly believed the real systems were within the test scope and continued executing the assigned cyberattack tasks. Notably, the newer internal research model stopped attacking after identifying that the targets might be real systems.Anthropic stated that these incidents primarily reflect issues with evaluation environment isolation and operational processes, rather than model alignment failures. The company has suspended related cybersecurity assessments, strengthened security controls in evaluation environments, continuously monitored test records, and will collaborate with third-party organizations to conduct further reviews.