GetChain News
中简 中繁 EN
GetChain News
Toggle sidebar

Online/Update

News linked to both this project and an event.

B.AI Popular Model Perks Update: Starting September 25, MiMo-V2.6 Series Takes Over Special Offers

Starting September 25 at 15:00 (SGT), B.AI will roll out a new round of upgrades to its popular model benefits: the brand-new MiMo-V2.6-Flash takes over with 90% off pricing, while the flagship MiMo-V2.6-Pro enjoys a half-price launch special. Currently, DeepSeek-V4.1-Flash, GLM-5.3-Flash, Qwen3.8-Flash, Hy3, and MiMo-V2.5 are enjoying their final 90% discounts. Developers are urged to take advantage of this final window to top up credits and make calls in advance to ensure seamless workflow continuity. Log in to chat.b.ai/chat to experience it now.

B. AI Responses API Adds Support for GLM and Kimi Series Models

B.AI announces that its Responses API officially supports the GLM and Kimi model series. Using a B.AI API Key, developers can directly call these two model series within the Codex client. Combined with the previously supported GPT and DeepSeek series, it now covers four mainstream model lineups. After completing the streamlined configuration of the API Key and Base URL, you can seamlessly integrate the reasoning, code, and text processing capabilities of these models into your daily development workflow, flexibly meeting diverse coding and business requirements. Log in to b.ai to experience it now.

GLM-5.3-FlashX officially launches on B.AI, with simultaneous access to both the API and Web Chat.

The B.AI platform announced that Z.AI's speed-optimized native multimodal large model GLM-5.3-FlashX has officially launched, with both the API and Web Chat now open for public testing.

Zhipu officially launches GLM-5.3-FlashX

Zhipu has officially announced the launch of GLM-5.3-FlashX. The model was previously made available to global developers under the name "Ox Alpha". According to Zhipu, leveraging inference compute from 100,000 domestic chips, the company has further increased infrastructure investment and optimized inference performance, achieving speeds of up to 200 Tokens/s. The GLM-5.3-FlashX API is now live, with the model identifier set to "GLM-5.3-FlashX".

Zhipu GLM Unveils Recursive Self-Improvement Practice, GLM-5.3-Flash Released

The Zhipu GLM team has disclosed its first engineering implementation of recursive self-improvement. An Infra Agent driven by GLM-5.3 completed the design, debugging, and optimization of the GLM-5.3-Flash inference infrastructure. The practice was executed on a domestic 100,000-GPU cluster, which the team describes as the first recursive self-improvement practice for large models in China.

Just 10% of the official price: B.AI’s Top 5 Most Popular Models Now Fully Live.

B.AI officially announces the arrival of five popular models now available at just 10% of their standard price. Starting September 16 at 17:00 (SGT), Qwen3.8-Flash, Hy3, and MiMo-V2.5 officially join the 10% pricing tier, teaming up with DeepSeek-V4.1-Flash and GLM-5.3-Flash—already available at this rate—to form B.AI’s exclusive five-model discounted bundle. This means developers can seamlessly switch across diverse workloads including inference, coding, multimodal processing, Chinese-language applications, and high-frequency APIs by paying only 10% of the official standard rate, empowering them to build cost-efficient AI infrastructure. Whether for independent development or large-scale team deployments, B.AI makes top-tier compute power accessible, truly delivering a low-cost, highly flexible model calling experience. Start your trial now: https://chat.b.ai/chat

B.AI Model Benefit Update: Two Flagship Models Now Fully Live at 10% Rate

B.AI announces that a new round of benefits for its popular models is now officially in effect: exclusive 10% discounted API call privileges for high-concurrency flagship models DeepSeek-V4.1-Flash and GLM-5.3-Flash are now fully available, allowing users to invoke them at just 10% of the original price to meet high-throughput scenario demands with exceptional cost-efficiency. Simultaneously, Qwen3.8-Flash, Hy3, and MiMo-V2.5 continue to offer fully free API calls, providing developers with richer zero-cost options. This benefit adjustment delivers maximum value for high-throughput needs while leveraging a multi-model matrix to lower the exploration threshold for developers. B.AI continues to empower developers to build low-cost AI infrastructure through a diverse model lineup and inclusive pricing, eliminating prohibitive compute expenses and accelerating the rapid deployment of innovations.

GLM-5.3-Flash Official Price Adjustments Take Effect Soon; B.AI Free Tier Extended Until September 12

Zhipu's official GLM-5.3-Flash will undergo a pricing update shortly. To provide global developers with a more ample preparation window, B.AI has announced a limited-time extension of its free access—until September 12 at 09:59 (SGT), GLM-5.3-Flash will remain completely free on the B.AI platform. Leading in platform usage volume, GLM-5.3-Flash features 320B total parameters and 18B activated parameters, supports a 1M ultra-long context window, and seamlessly combines rapid response with powerful reasoning capabilities, making it a highly cost-effective choice for high-frequency coding, massive data processing, and complex Agent workflows. Additionally, other popular models such as Qwen3.8 Flash, Hy3, and MiMo V2.5 continue to be 100% free on the B.AI platform. B.AI remains committed to supporting global developers with inclusive computing resources, ensuring that frontier model capabilities are truly within everyone's reach.

OpenAI CFO: Enterprise revenue grew 32% in a single month; Luna usage surged approximately 10x after an 80% price cut

Odaily News: OpenAI CFO Sarah Friar stated that the company is accelerating the expansion of AI into specialized fields such as chip design, life sciences, and financial services, and is experimenting with pricing based on business outcomes rather than usage. OpenAI's enterprise business revenue grew 32% from June to July this year, while the company's overall annualized revenue increased approximately 20% during the same period. As of mid-year, revenue from enterprise and consumer businesses had split roughly evenly.Additionally, OpenAI recently cut the price of its low-cost Luna model by 80%, which was followed by an approximate 10-fold increase in usage. Friar noted that in cloud deployment scenarios, Luna's cost is even lower than Z.ai's GLM 5.3. OpenAI's coding tool Codex has now reached 25 million users. The company is also leveraging its own AI models to assist in developing the Jalapeno chip and completed chip design tape-out within nine months. (Reuters)

B.AI continues to provide free calls to GLM-5.3-Flash, maintaining zero-cost development.

B.AI has announced that after Zhipu's official limited-time 50% discount expires at 24:00 on September 9, GLM-5.3-Flash will continue to provide zero-cost API calls for global developers, requiring no changes to usage habits due to upstream price adjustments.

GLM-5.3-Flash Tops B.AI Model API Call Volume Rankings, Cumulative Throughput Surpasses 2.41 Trillion Tokens

The GLM-5.3-Flash model has become the most frequently invoked and popular model on the B.AI platform, with cumulative token throughput exceeding 2.41 trillion. As the first native multimodal model in the GLM-5 series, GLM-5.3-Flash features 320B total parameters and 18B active parameters. It employs a hybrid architecture combining sparse and linear attention mechanisms, supports 1M ultra-long context windows, and balances rapid response, powerful reasoning capabilities, and high cost-effectiveness. Starting today, developers can still invoke this model for free via the B.AI platform, covering diverse scenarios such as high-frequency APIs, coding, complex Agents, and ultra-long document processing. Try it now: chat.b.ai/chat

B.AI Cumulative Token Throughput Surpasses 10.9 Trillion, Free Promotions Continue to Release Compute Dividends

The AI Agent infrastructure platform B.AI announced that since the launch of its free campaign, the platform's cumulative token throughput has exceeded 10.9 trillion, with total API calls reaching 89.56 million, attracting over 239,000 new registered users (including 235,000+ API developers). The platform infrastructure has consistently demonstrated its stability and capacity under the stress of massive high-concurrency scenarios and intensive Agent workflows. In terms of the model lineup, GLM-5.3-Flash (Ox Alpha), Qwen3.8-Flash, Tencent Hy3, and Xiaomi MiMo-V2.5 remain fully free, while DeepSeek-V4-Flash and Vision-Exp versions are now available at a 50% discount, striking a balance between zero-threshold access and exceptional cost efficiency. Going forward, B.AI will continue to deliver more efficient, reliable, and cost-effective AI compute services to developers and enterprise teams, accelerating the real-world deployment of AI applications. Visit chat.b.ai/chat to start deploying your efficient workflows today.

B.AI's Full Free Access Campaign Continues, Cumulative Throughput Surpasses 5.74 Trillion, Setting Another Milestone

Odaily News, August 31 — B.AI platform's full free access campaign for cutting-edge large models continues to operate at high intensity, with the platform's cumulative token throughput officially surpassing the historic milestone of 5.74 trillion (5T+). Under the rigorous demands of massive concurrent requests and heavy Agent tasks, B.AI's industrial-grade infrastructure has demonstrated exceptionally stable carrying capacity. Currently, the platform has assembled six top-tier models, including GLM-5.3-Flash (Ox Alpha), Qwen3.8-Flash, DeepSeek-V4-Flash, DeepSeek-V4-Flash-Vision-Exp, Tencent Hy3, and Xiaomi MiMo-V2.5, all available to users with zero barriers and unlimited free access, covering diverse scenarios such as code generation, ultra-long text processing, multimodal visual understanding, and Agent workflow deployment.From now on, users who log in to the B.AI platform can access the full suite of cutting-edge models at zero cost, with seamless integration of unlimited computing power, allowing every developer and AI enthusiast to truly enjoy "computing freedom."

Newly Released Model GLM-5.3-Flash Is Now Live, B.AI Adds More Free Benefits

The all-in-one AGI infrastructure platform B.AI has announced that Zhipu's GLM-5.3-Flash (320B-A18B) has officially launched on the platform and is now fully accessible at zero cost through the official API group. As the first native all-modal model in the GLM-5 series, GLM-5.3-Flash features 320B total parameters and 18B active parameters, powered by purely domestic AI chips. It introduces a hybrid architecture of sparse and linear attention for the first time, supports a 1M ultra-long context window, delivers formidable performance alongside exceptional cost-efficiency, and stands as a leading powerhouse among current domestic models. Effective immediately, developers can invoke this advanced tool at zero cost via the B.AI API official group, effortlessly tackling complex reasoning, long-text processing, and multimodal tasks. B.AI continues to lower the barrier to AI development through democratized computing power, enabling more developers to access top-tier domestic models.

Zhipu Open-Sources GLM-5.3-Flash, US PCE Beats Estimates, Nvidia Gains After Hours

Zhipu announced the launch and open-source release of the GLM-5.3-Flash model. The year-over-year rate of the US July PCE price index reached 3.7%, exceeding expectations, while expectations for Federal Reserve rate hikes intensified; Nvidia disclosed 70% growth guidance for fiscal year 2028, boosting its after-hours stock price.

Zhipu AI launches and open-sources the "Niulai" model GLM-5.3-Flash, enhancing multimodal capabilities and inference cost efficiency.

Z.ai has released GLM-5.3-Flash, calling it the first native multimodal model in the GLM-5 series. The model features 320 billion total parameters and 18 billion active parameters, outperforming GLM-5.2 across multiple coding, agent, and vision benchmarks, and approaching Claude Opus 4.8 on certain coding tasks. GLM-5.3-Flash employs a hybrid architecture that combines sparse attention with linear attention, and introduces mechanisms such as mHC and IndexPool to reduce long-context inference costs and KV cache overhead.

Zhipu AI confirms Ox Alpha as the new iteration of the GLM series; model weights will be released tonight.

According to Bloomberg, China-based AI company Z.AI (Zhipu) confirmed on Wednesday that the mysterious AI model Ox Alpha, which has recently risen rapidly to the top of online usage charts, is a new iteration of its GLM series. Responding to a Bloomberg news inquiry, the company stated it would release the model weights for Ox Alpha tonight.

Zhipu AI Releases GLM-5.3: Coding and Long-Horizon Task Capabilities Significantly Improved, Model Weights to Be Released in Two Weeks

Z.ai releases GLM-5.3, based on the same foundation model as GLM-5.2, achieving capability improvements through expanded post-training. According to the official announcement, GLM-5.3 improves by 50% over GLM-5.2 on the internal Z.ai Code Bench coding benchmark, and reaches a leading level among open models in public benchmarks such as Terminal Bench 3.0 and Agents' Last Exam. In terms of cybersecurity, GLM-5.3 achieved a score of 84.5% in the CyberGym vulnerability discovery test, and significantly improved compared to the previous generation in exploit chain-related tests such as ExploitBench and ExploitGym.

GLM-5.3 即将发布,聚焦编程与网络安全能力

Z.ai 发文预告 GLM-5.3 即将推出,该模型聚焦编程能力,并面向网络防御挑战。官方称,GLM-5.3 通过对 743B 基础模型进行后训练,实现了顶尖的编程和智能代理能力,并在网络安全领域取得重要进展,为开源模型树立新标准。

B.AI Platform Launches GLM-5.2 Flagship Model with Limited-Time 40% Off Promotion

The B.AI platform is now launching a limited-time 40% discount on the GLM-5.2 model for all users. During the event, all calls enjoy a 40% discount on the settlement price, whether deployed via API for engineering purposes or used for daily interaction on the web interface. The discounted prices are: Input 0.84, Cache Write 0.84, Cache Read 0.168, Output 2.64 (Unit: Credits/Token). As a new generation open-source flagship model, GLM-5.2 features a 1M ultra-long context window and is designed for high-performance scenarios such as large-scale code development, complex reasoning, and agent tasks. Users can access it directly via the official API standard channel or the web interface, enabling seamless integration across all-scenario workflows. This event aims to unleash top-tier AI productivity at a lower cost. Starting today, users can log in to the B.AI platform to experience it.