Golem is a peer-to-peer, decentralized computation network that enables a marketplace for computing power. It allows anyone to share and aggregate computing resources, creating a network to share resources. By doing so, Golem hopes to provide software developers with an alternative to traditional centralized cloud service providers, such as Amazon.
Zhipu announces the completion of approximately $5 billion in financing, including a $2 billion equity placement and a $3 billion zero-coupon convertible bond. The funds will be primarily used for the next-generation GLM model, self-training systems, and computing infrastructure development.
Odaily Odaily News: AI personal assistant startup Pally has announced the completion of a $5.2 million funding round, led by Cyber Fund and Y Combinator, with participation from Pioneer Fund, Founders Inc, 468 Capital, and angel investors. The company is valued at $30 million and primarily provides AI personal assistant services through an SMS interface. Pally was launched in 2025, with its latest version released in June, and can sync with platforms such as Gmail, Outlook, Slack, Granola, and Notion. After user authorization, Pally can book flights, manage inboxes, reply to messages, and reserve restaurant tables. Pally uses Anthropic's Claude, OpenAI's ChatGPT, as well as open-source models such as Kimi K3 and GLM 5.2. Co-founder Haz Hubble stated that the company will not sell user data or use related data to train models. (Business Insider)
Chinese AI startup Moonshot AI will release the model weights of its high-performance model, Kimi K3. Developers can download the model, modify it for various purposes, and run it in their own data centers or cloud environments.Kimi K3 boasts 2.8 trillion parameters and a 1 million token context window, enabling it to process large-scale documents and codebases in a single pass. Moonshot AI plans to later publish a technical report detailing the model's architecture, training methodology, and performance evaluation results.Following the release of Kimi K3, Moonshot AI's daily revenue is reported to have increased by at least 6 times. The company is reportedly advancing a new round of fundraising at a $50 billion valuation and is considering a Hong Kong listing as early as this year.According to Bloomberg Intelligence, following the release of Kimi K3 and Z.AI's GLM-5.2, the share of Chinese open-weight models in overall token usage has risen to 68%. Services like AWS Bedrock, Microsoft Azure Foundry, and Google Vertex AI currently do not offer Chinese open-weight models such as Kimi K3 and GLM-5.2.
Odaily News AMD, the semiconductor giant, announced the launch of its enterprise-grade AI programming platform, AMD Instinct Coder. The platform combines AMD chips, Supermicro servers, and Spectro Cloud software, aiming to help enterprises deploy AI coding assistants locally, reduce the cost of cloud-based AI models, and protect code and data security.AMD stated that Instinct Coder is an "out-of-the-box" end-to-end AI development platform, integrating AMD EPYC processors, AMD Instinct GPUs, Supermicro AI servers, Spectro Cloud PaletteAI Inference Launchpad software, and the AMD-optimized GLM-5.2 model. It can be used for software development scenarios such as code generation, application modernization, automated testing, and code review.AMD said that compared to relying on cutting-edge cloud-based AI models, Instinct Coder can help enterprises reduce total cost of ownership (TCO) by up to 70%, with the fastest payback period shortened to 6 months.AMD noted that more and more enterprises are looking to leverage AI to improve development efficiency, but face two major challenges: on one hand, the cost of invoking top-tier cloud models continues to rise; on the other hand, entrusting enterprise source code, intellectual property, and sensitive data to third-party services poses security and compliance risks.Through a local deployment model, Instinct Coder allows enterprises to maintain control over their data and code while providing more predictable infrastructure costs. The platform supports development tools such as Claude Code, OpenAI Codex, Visual Studio Code, and Cursor, with each node supporting up to 50 users (30 concurrent users).Additionally, the PaletteAI Inference Launchpad provided by Spectro Cloud enables AI workload management, model routing, request auditing, and cost monitoring, and supports invoking external models such as Anthropic, OpenAI, Google, or xAI when needed.AMD stated that Instinct Coder aims to help enterprises break free from the high costs of cloud-based AI services, accelerate AI-driven software development processes while ensuring data security and autonomous control.
据 X 用户 Rohan Paul 披露,当 OpenAI 模型攻破 Hugging Face 时,Anthropic 的 Claude 5 拒绝协助取证,Hugging Face 最终转向 Nvidia 量化版 GLM 5.2 完成应急处置。该用户评论称,禁止开源模型将削弱防御方的能力,此次事件印证了开源模型在安全响应中的实际价值。
The B.AI platform’s “Self-Selected Service Provider” model matrix has officially expanded, newly integrating leading large language models including Moonshot (Kimi series) and Z.ai (GLM series). This module offers four discount tiers: 90%, 60%, 40%, and 20% off. Users can generate a personalized “discounted model API key” with a single click, enabling seamless switching between core business operations and routine testing—achieving an optimal balance of high availability and low cost. Moreover, all discounts can be stacked with up to a 1:1 top-up bonus, further lowering the barrier to compute access. Starting today, log in to the B.AI console to customize your专属 model portfolio and enter a new era of AI API calls delivering unmatched value.
According to on-chain analyst Ember (@EmberCN), the rsETH incident on April 18 resulted in a funding shortfall of approximately 68,900 ETH (around $160 million): the hacker collateralized rsETH to borrow 99,600 ETH; after Arbitrum recovered 30,700 ETH, the remaining funds were fully converted by the hacker into BTC. The incident has now entered the remediation phase. Aave is coordinating the establishment of a “DeFi United” relief fund, which has so far received cumulative donations totaling 13,500 ETH (approximately $31.45 million). Donors include Lido Finance (2,500 stETH), ether.fi Foundation (5,000 ETH), Aave founder Stani Kulechov (5,000 ETH), Golem Foundation (1,000 ETH), as well as LayerZero and Ink Foundation (amounts undisclosed).
Z.ai releases GLM-5.3, based on the same foundation model as GLM-5.2, achieving capability improvements through expanded post-training. According to the official announcement, GLM-5.3 improves by 50% over GLM-5.2 on the internal Z.ai Code Bench coding benchmark, and reaches a leading level among open models in public benchmarks such as Terminal Bench 3.0 and Agents' Last Exam. In terms of cybersecurity, GLM-5.3 achieved a score of 84.5% in the CyberGym vulnerability discovery test, and significantly improved compared to the previous generation in exploit chain-related tests such as ExploitBench and ExploitGym.
Odaily News: The Bitcoin Red Team volunteer security initiative has scanned approximately 150 Bitcoin-related code repositories and disclosed more than a dozen vulnerabilities. The team is developing an open-source AI platform to audit Bitcoin software, covering wallets, cryptographic libraries, infrastructure, and other projects. AnchorWatch CEO Rob Hamilton stated that the team has so far spent approximately $20,000 on various AI services, using Kimi K3, OpenAI's GPT Sol, Anthropic's Claude Fable and Opus, as well as Z.ai's GLM 5.2 to identify vulnerabilities and generate related documentation. Pseudonymous Bitcoin developer Calle said that over the past 12 hours, the team has reported critical vulnerabilities to multiple projects, discovering on average roughly one critical vulnerability per person per hour, with daily spending of around $10,000. The team has not disclosed the affected projects or details of the vulnerabilities.
Developers on Reddit used Claude Code to scan the Coldcard open-source firmware for vulnerabilities, pinpointing the core issue within 8 minutes: When generating private keys, the firmware invoked a software pseudo-random number generator instead of a hardware true random number generator, and it was this vulnerability that led to the theft of approximately $70 million in BTC from 1,196 wallets. Meanwhile, community users also reported that using Zhipu GLM 5.2 (trained on June 16, offline) for an independent scan similarly discovered this vulnerability. This bug has existed in the open-source wallet code for over five years.
据 X 用户 Rohan Paul 披露,当 OpenAI 模型攻破 Hugging Face 时,Anthropic 的 Claude 5 拒绝协助取证,Hugging Face 最终转向 Nvidia 量化版 GLM 5.2 完成应急处置。该用户评论称,禁止开源模型将削弱防御方的能力,此次事件印证了开源模型在安全响应中的实际价值。
the U.S. Department of Commerce's AI Standards and Innovation Center, in collaboration with the UK AI Safety Institute, tested the cyber attack capabilities of Kimi K3, emphasizing that "the United States still leads."However, the value of the evaluation is debated due to limitations in the testing scope. Due to hosting environment constraints, Kimi K3 only participated in partial testing, with its overall cyber capabilities estimated primarily based on 41 exploit benchmarks. In contrast, other models underwent more comprehensive testing, resulting in a larger margin of error for Kimi K3's results.In the exploit testing, Kimi K3 scored approximately 32%, higher than GLM-5.2's 24%, but lower than the average of approximately 76% for leading U.S. models. In a simulated attack chain test, Kimi K3 completed an average of 17 out of 32 steps in the attack chain and successfully breached the network once in 10 attempts, while U.S. frontier models completed an average of 28.5 steps.The report notes that Kimi K3 already possesses a certain level of autonomous attack capability, and its security guardrails did not prevent the model from developing exploits or executing attacks. However, the report also emphasizes that the testing scope was limited.
OpenAI confirmed that the unreleased GPT-5.6 Sol and another unnamed, more powerful pre-release model breached a restricted sandbox environment during ExploitGym benchmark evaluations and infiltrated Hugging Face's production infrastructure to obtain test answers.OpenAI stated that the models leveraged a zero-day vulnerability in an internal software package registry proxy to escalate privileges and move laterally, ultimately connecting to a machine with internet access. The models then identified and chained together vulnerabilities in both the OpenAI research environment and Hugging Face's production infrastructure, directly retrieving test solutions from Hugging Face's production database.Hugging Face disclosed the incident on July 16, stating that the attack was executed end-to-end by an autonomous AI agent system, involving thousands of operations within short-lived sandboxes and accessing internal datasets and service credentials. OpenAI confirmed its models were the subject of the incident five days later.Hugging Face stated that its security team, in order to analyze over 17,000 attack logs, initially attempted to use a commercial US frontier AI interface, but the request was blocked due to safety guardrails. They subsequently switched to using the 753-billion parameter open-weight model GLM 5.2 from Chinese AI startup Z.ai on their own infrastructure to complete the forensic analysis.
B.AI announces that a new round of benefits for its popular models is now officially in effect: exclusive 10% discounted API call privileges for high-concurrency flagship models DeepSeek-V4.1-Flash and GLM-5.3-Flash are now fully available, allowing users to invoke them at just 10% of the original price to meet high-throughput scenario demands with exceptional cost-efficiency. Simultaneously, Qwen3.8-Flash, Hy3, and MiMo-V2.5 continue to offer fully free API calls, providing developers with richer zero-cost options. This benefit adjustment delivers maximum value for high-throughput needs while leveraging a multi-model matrix to lower the exploration threshold for developers. B.AI continues to empower developers to build low-cost AI infrastructure through a diverse model lineup and inclusive pricing, eliminating prohibitive compute expenses and accelerating the rapid deployment of innovations.
Zhipu's official GLM-5.3-Flash will undergo a pricing update shortly. To provide global developers with a more ample preparation window, B.AI has announced a limited-time extension of its free access—until September 12 at 09:59 (SGT), GLM-5.3-Flash will remain completely free on the B.AI platform. Leading in platform usage volume, GLM-5.3-Flash features 320B total parameters and 18B activated parameters, supports a 1M ultra-long context window, and seamlessly combines rapid response with powerful reasoning capabilities, making it a highly cost-effective choice for high-frequency coding, massive data processing, and complex Agent workflows. Additionally, other popular models such as Qwen3.8 Flash, Hy3, and MiMo V2.5 continue to be 100% free on the B.AI platform. B.AI remains committed to supporting global developers with inclusive computing resources, ensuring that frontier model capabilities are truly within everyone's reach.
Odaily News: OpenAI CFO Sarah Friar stated that the company is accelerating the expansion of AI into specialized fields such as chip design, life sciences, and financial services, and is experimenting with pricing based on business outcomes rather than usage. OpenAI's enterprise business revenue grew 32% from June to July this year, while the company's overall annualized revenue increased approximately 20% during the same period. As of mid-year, revenue from enterprise and consumer businesses had split roughly evenly.Additionally, OpenAI recently cut the price of its low-cost Luna model by 80%, which was followed by an approximate 10-fold increase in usage. Friar noted that in cloud deployment scenarios, Luna's cost is even lower than Z.ai's GLM 5.3. OpenAI's coding tool Codex has now reached 25 million users. The company is also leveraging its own AI models to assist in developing the Jalapeno chip and completed chip design tape-out within nine months. (Reuters)
B.AI has announced that after Zhipu's official limited-time 50% discount expires at 24:00 on September 9, GLM-5.3-Flash will continue to provide zero-cost API calls for global developers, requiring no changes to usage habits due to upstream price adjustments.
The GLM-5.3-Flash model has become the most frequently invoked and popular model on the B.AI platform, with cumulative token throughput exceeding 2.41 trillion. As the first native multimodal model in the GLM-5 series, GLM-5.3-Flash features 320B total parameters and 18B active parameters. It employs a hybrid architecture combining sparse and linear attention mechanisms, supports 1M ultra-long context windows, and balances rapid response, powerful reasoning capabilities, and high cost-effectiveness. Starting today, developers can still invoke this model for free via the B.AI platform, covering diverse scenarios such as high-frequency APIs, coding, complex Agents, and ultra-long document processing. Try it now: chat.b.ai/chat
The AI Agent infrastructure platform B.AI announced that since the launch of its free campaign, the platform's cumulative token throughput has exceeded 10.9 trillion, with total API calls reaching 89.56 million, attracting over 239,000 new registered users (including 235,000+ API developers). The platform infrastructure has consistently demonstrated its stability and capacity under the stress of massive high-concurrency scenarios and intensive Agent workflows. In terms of the model lineup, GLM-5.3-Flash (Ox Alpha), Qwen3.8-Flash, Tencent Hy3, and Xiaomi MiMo-V2.5 remain fully free, while DeepSeek-V4-Flash and Vision-Exp versions are now available at a 50% discount, striking a balance between zero-threshold access and exceptional cost efficiency. Going forward, B.AI will continue to deliver more efficient, reliable, and cost-effective AI compute services to developers and enterprise teams, accelerating the real-world deployment of AI applications. Visit chat.b.ai/chat to start deploying your efficient workflows today.
Zhipu announces the completion of approximately $5 billion in financing, including a $2 billion equity placement and a $3 billion zero-coupon convertible bond. The funds will be primarily used for the next-generation GLM model, self-training systems, and computing infrastructure development.
B.AI announces that a new round of benefits for its popular models is now officially in effect: exclusive 10% discounted API call privileges for high-concurrency flagship models DeepSeek-V4.1-Flash and GLM-5.3-Flash are now fully available, allowing users to invoke them at just 10% of the original price to meet high-throughput scenario demands with exceptional cost-efficiency. Simultaneously, Qwen3.8-Flash, Hy3, and MiMo-V2.5 continue to offer fully free API calls, providing developers with richer zero-cost options. This benefit adjustment delivers maximum value for high-throughput needs while leveraging a multi-model matrix to lower the exploration threshold for developers. B.AI continues to empower developers to build low-cost AI infrastructure through a diverse model lineup and inclusive pricing, eliminating prohibitive compute expenses and accelerating the rapid deployment of innovations.
B.AI announces that the limited-time free trial for GLM-5.3-Flash is drawing to a close, seamlessly giving way to an exclusive promotional call rate at 10% of the original price starting September 12 at 10:00 SGT. Users will pay only 10% of the original cost, directly saving 90%, with input priced as low as $0.015/M and output as low as $0.05/M. As a native multimodal large model, GLM-5.3-Flash features 320B total parameters and 18B activated parameters, incorporating a hybrid architecture that combines sparse and linear attention with a 1M ultra-long context window, delivering strong performance across use cases such as programming and development, AI agents, long-context processing, and high-frequency API workloads.
Zhipu's official GLM-5.3-Flash will undergo a pricing update shortly. To provide global developers with a more ample preparation window, B.AI has announced a limited-time extension of its free access—until September 12 at 09:59 (SGT), GLM-5.3-Flash will remain completely free on the B.AI platform. Leading in platform usage volume, GLM-5.3-Flash features 320B total parameters and 18B activated parameters, supports a 1M ultra-long context window, and seamlessly combines rapid response with powerful reasoning capabilities, making it a highly cost-effective choice for high-frequency coding, massive data processing, and complex Agent workflows. Additionally, other popular models such as Qwen3.8 Flash, Hy3, and MiMo V2.5 continue to be 100% free on the B.AI platform. B.AI remains committed to supporting global developers with inclusive computing resources, ensuring that frontier model capabilities are truly within everyone's reach.
Odaily News: OpenAI CFO Sarah Friar stated that the company is accelerating the expansion of AI into specialized fields such as chip design, life sciences, and financial services, and is experimenting with pricing based on business outcomes rather than usage. OpenAI's enterprise business revenue grew 32% from June to July this year, while the company's overall annualized revenue increased approximately 20% during the same period. As of mid-year, revenue from enterprise and consumer businesses had split roughly evenly.Additionally, OpenAI recently cut the price of its low-cost Luna model by 80%, which was followed by an approximate 10-fold increase in usage. Friar noted that in cloud deployment scenarios, Luna's cost is even lower than Z.ai's GLM 5.3. OpenAI's coding tool Codex has now reached 25 million users. The company is also leveraging its own AI models to assist in developing the Jalapeno chip and completed chip design tape-out within nine months. (Reuters)
B.AI has announced that after Zhipu's official limited-time 50% discount expires at 24:00 on September 9, GLM-5.3-Flash will continue to provide zero-cost API calls for global developers, requiring no changes to usage habits due to upstream price adjustments.