GetChain News
中简 中繁 EN
GetChain News
Toggle sidebar

Online/Update

News linked to both this project and an event.

DeepSeek V4-Flash-0731 Public Beta: Code Agent Capabilities Significantly Enhanced, Price Unchanged

DeepSeek officially launches the public beta of version V4-Flash-0731. Through post-training and Agent framework optimization, DeepSWE benchmark scores have improved from 7.3 to 54.4, Cybergym scores have doubled to 76.7, and performance on some hard benchmarks approaches Claude Opus 4.8. The new version does not increase parameters, prices remain unchanged, it natively supports the Responses API, and existing user APIs will automatically switch to the new version. Currently, only the Flash API is updated; the V4 Pro official version is still under development.

Anthropic discloses three cybersecurity incidents involving Claude, where the model accessed real systems and obtained unauthorized permissions

Odaily News, July 30 – Anthropic released a report stating that during a review of cybersecurity assessment records, three incidents were discovered in which the Claude model accessed the internet in a third-party evaluation environment and further obtained unauthorized access to three real organizations' systems.Anthropic stated that the review covered 141,000 evaluation runs that could have potentially gained network access, and a total of three related incidents were found. All incidents occurred during Capture The Flag (CTF) cybersecurity tests, where the model was told the environment was a simulation with no internet access; however, due to configuration errors by the evaluation partner, the actual environment had internet connectivity.Among these, Claude Opus 4.7 accessed real company infrastructure during one test and obtained database permissions containing hundreds of production data records; Claude Mythos 5 built a malicious Python package and uploaded it to PyPI, resulting in the package being downloaded and run on 15 real systems; another internal research test model scanned approximately 9,000 targets and accessed a company's internet application through a public vulnerability.Anthropic stated that these incidents were not cases of the model actively seeking to escape or pursue its own goals, but rather the model mistakenly believed the real systems were within the test scope and continued executing the assigned cyberattack tasks. Notably, the newer internal research model stopped attacking after identifying that the targets might be real systems.Anthropic stated that these incidents primarily reflect issues with evaluation environment isolation and operational processes, rather than model alignment failures. The company has suspended related cybersecurity assessments, strengthened security controls in evaluation environments, continuously monitored test records, and will collaborate with third-party organizations to conduct further reviews.

Perplexity CEO: GLM 700B Parameter Model Performance Close to Opus Level

Perplexity CEO Aravind Srinivas tweeted that GLM is a severely underrated model, demonstrating performance close to Opus-level at the 700B parameter scale with extremely high operational efficiency. Previously, Moonshot AI just released the open-source model Kimi K3, and industry attention on the GLM series models under Zhipu AI is rising.

Kimi K3 Released, Gap Between Open-Source and Closed-Source Models Narrows to 4 Points

Moonshot AI launches open-source model Kimi K3, scoring 57 points on the Artificial Analysis Intelligence Index, becoming the third highest-scoring model, second only to Anthropic's Claude Opus 5 (61 points), Claude Fable 5 (60 points), and OpenAI's GPT-5.6 Sol (59 points). The index shows that the gap between leading closed-source models and open-weight models has narrowed to 4 points, the smallest gap since the release of GLM-5 in February. This signifies that open-source models are rapidly catching up to closed-source models in performance.

Claude Opus 5 Officially Available on B.AI API, Dual-Channel Access Balances Performance and Cost

The B.AI API platform officially launches the Claude Opus 5 model. This model provides a context window of up to 1 million tokens and a single output limit of 128,000 tokens, achieving significant performance leaps in scenarios such as complex reasoning, long code engineering development, large-scale document analysis, and knowledge-intensive workflows. To meet diverse deployment needs, this release simultaneously opens a dual-channel API access solution on the B.AI platform: in addition to the official standard interface, a new "Designated Service Provider" cooperation channel is added, through which enterprise users can obtain cost discounts of up to 40%, facilitating flexible balancing of performance and budget expenditure based on actual business loads. Developers can log in to the official B.AI platform starting today to experience the exceptional capabilities of Claude Opus 5 first.

Claude Opus 5 Tops Artificial Analysis Intelligence Index v4.1

Anthropic's Claude Opus 5 (max and xhigh versions) achieved the highest intelligence score on the Artificial Analysis Intelligence Index v4.1. The index integrates 9 evaluation benchmarks including GDPval-AA v2, GPQA Diamond, and Humanity's Last Exam. Previously, Opus 5 has demonstrated leading performance in multiple benchmark tests, including scoring twice that of the previous generation on Frontier-Bench, topping the AA-Briefcase agent benchmark, and outperforming Fable 5 in cost-performance ratio.

Anthropic Launches Opus 5 AI Model, Performance Close to Fable 5, Price Halved

Anthropic has officially released the Opus 5 AI model; the company states that its performance is close to the frontier model Fable 5, but the price is only half that of the latter.

Polymarket: "Next Claude Opus model will be released on July 23, 2026" probability currently at 61%, up 51% in 24 hours

PPP Prediction Market Tool monitoring shows that on Polymarket, the probability of "the next Claude Opus model will be released on July 23, 2026" is currently at 61%, up 51% in 24 hours; additionally, the probability of "release before July 24" is at 9%; the probability of "release before July 25" is at 5%;This event will be settled based on the date (Eastern Time) when Anthropic's next Claude Opus model becomes available to the public. Only models explicitly named "Opus" by Anthropic are valid, and they must be accessible to the public (including public beta testing or open waitlist). Other models such as Sonnet and Haiku are not counted. The final settlement will primarily rely on official Anthropic information.Join the PPP Signal Push Community to stay ahead and seize the opportunity.

月之暗面将发布新模型 Kimi K3,或超越 Opus 4.8

据两名知情人士透露,月之暗面计划在未来几天内发布 Kimi K3,该模型参数规模达 2 万亿至 3 万亿,将成为中国迄今最大的 AI 模型。Anthropic 尚未披露旗下模型的参数规模,但业内人士普遍推测,Opus 4.8 拥有约 1.5 万亿至 2 万亿个参数。

SpaceXAI and Cursor Plan to Launch First Joint AI Model

According to Reuters citing The Information, SpaceXAI and Cursor plan to release their first jointly developed artificial intelligence model as early as Wednesday. The model was originally scheduled for launch early this week but was delayed for efficiency optimization. The report noted that the new model is expected to possess strong rapid information processing capabilities, with some performance aspects potentially competitive with Anthropic's Opus 4.8 and OpenAI's GPT 5.5.

GPT-5.6 rumored to be publicly available as early as July 7, Gemini 3.5 Pro may launch on July 17

Tech blogger Leo has revealed that OpenAI may make GPT-5.6 available to the public between July 7 and July 9, with the earliest possible date being July 7. The new model's plan usage quotas are said to be more generous, and OpenAI is further strengthening its safety strategies ahead of the launch.Additionally, according to sources, Google DeepMind has tentatively scheduled the release of Gemini 3.5 Pro for July 17. Another tech blogger, Astro Polo, stated that Gemini 3.5 Pro will support a 2 million token context window, doubling the 1 million token context window currently supported by Claude Sonnet 5, Claude Opus 4.8, and Claude Fable 5. This makes it more suitable for handling large codebases, lengthy documents, and long conversations.Note: The above information is based on market rumors and has not yet been officially confirmed by OpenAI or Google DeepMind.

Meituan Releases Trillion-Parameter Large Model LongCat-2.0, the First Trillion-Parameter Model to Complete Full-Process Training on a Domestic Computing Cluster

According to Meituan's official release, Meituan has officially launched the new generation large model LongCat-2.0 and open-sourced it simultaneously. The model features a total of 1.6T parameters, making it the industry's first trillion-parameter model to complete full-process training and inference on a domestic computing cluster of 50,000 cards. It natively supports 1M ultra-long context and focuses primarily on code understanding, generation, and execution in Agentic Coding scenarios. Technically, LongCat-2.0 adopts the LongCat Sparse Attention (LSA) sparse attention mechanism, reducing long text computation complexity from quadratic to linear; achieves token-level dynamic activation (33B~56B) via a zero-computation expert mechanism; and introduces the MOPD architecture to fuse three sets of expert capabilities: Agent, Reasoning, and Interaction. In terms of training efficiency, the team spent three years overcoming challenges in adapting to domestic computing power, reducing the monthly average daily failure rate by over 70%, increasing training MFU by 1.5 times, and achieving steady-state daily throughput exceeding 1T tokens/day. In terms of performance evaluation, LongCat-2.0 achieved a score of 59.5 on SWE-bench Pro, surpassing Gemini 3.1 Pro (54.2), GPT-5.5 (58.6), and Claude Opus 4.6 (57.3); and achieved a score of 79.9 on BrowseComp.

Polymarket probability of "Claude Fable 5 restored for US customers before July 1" rises to 73%, up 33% in 24H

Monitoring by Odaily Seer Prophet Channel shows that the probability of "Claude Fable 5 restored for US customers before July 1" on Polymarket has risen to 73%, up 33% in 24H.If Anthropic reopens Claude Fable 5 (or Claude Mythos, or a version confirmed to be the same model) to the US public before the specified date, this event will settle as "Yes"; otherwise, it will settle as "No". Qualifying methods of restoration include public beta or public waitlist; closed testing and private access do not count. Other models such as Haiku, Sonnet, and Opus are not included by default unless they are confirmed to be the same model as Claude Fable 5. Settlement will primarily be based on official announcements from Anthropic, supplemented by consensus reports from mainstream media.Due to export controls imposed by the US government on national security grounds, Anthropic urgently suspended global access to the Claude Fable 5 model just days after its release on June 9. Although the restrictions were primarily aimed at foreign users, due to the difficulty of real-time user identity screening, Anthropic ultimately closed access for all users, including those in the US, and switched requests to less capable models such as Opus 4.8. Anthropic is currently engaged in high-level discussions with the White House on issues including model capabilities and potential jailbreak risks. The company has publicly stated that the ban may have stemmed from a misunderstanding and is actively pushing to restore access.Odaily Seer Prophet Channel continues to monitor the prediction market, seeing changes before they are priced in.

Zhipu Founder Tang Jie: Open-Source Release of GLM-5.2 Will Gradually Narrow the Performance Gap with OpenAI and Anthropic

: Zhipu AI founder Tang Jie posted on the X platform, stating that since the official open-source release of GLM-5.2, it has achieved leading results in multiple international authoritative evaluations and competitive rankings. In the Artificial Analysis Intelligence Index comprehensive evaluation, GLM-5.2 scored 51 points, placing it in the same range as Anthropic's Claude Opus 4.8. In the Code Arena front-end code generation adversarial test, it ranked 2nd globally with an Elo of 1595, and in the DesignArena design and code integration scenario, it scored 1360 points, ranking 1st.Overall, Zhipu GLM-5.2 continues to rank among the top globally in real-world scenario evaluations across multiple areas, including front-end development, design generation, and software engineering. It is steadily narrowing the performance gap with cutting-edge models from OpenAI and Anthropic and will continue to push the upper limits of model capabilities.Previously, in response to Musk's statement that Chinese large models might reach Anthropic's Fable level by the first quarter of next year, Zhipu AI founder Tang Jie replied, "It won't take that long."

Anthropic Launches Phase Two of Project Fetch: Claude Opus 4.7 Achieves 10x Speedup in Robotics Tasks

Anthropic has released the results of Phase Two of “Project Fetch,” evaluating the enhanced capabilities of its latest model in real-world robotic manipulation. Conducted in August 2025, the experiment tasked non-robotics-expert Anthropic employees with completing a series of complex tasks using off-the-shelf quadruped robots. Performance was compared between two conditions: “using Claude models for assistance” versus “relying solely on humans and the internet.” Results show that, under fully autonomous operation by the latest model—Claude Opus 4.7—the robot achieved significantly faster average completion times across all successfully executed tasks than the human team, with execution speed improved by at least 10×.

Zcash Founder Says Claude Mythos Audit Found No Critical Vulnerabilities

Odaily Zcash founder Zooko Wilcox posted on X stating that a security audit conducted by Anthropic's Claude Mythos AI model did not find any "more severe vulnerabilities" in the Zcash protocol. The audit was commissioned by Shielded Labs, a Swiss non-profit organization supporting Zcash development. On June 3, Zcash developers temporarily paused Orchard transactions after discovering a vulnerability in the shielded pool, restoring functionality through an emergency upgrade the same day. The issue stemmed from a four-year-old forging vulnerability in the Orchard shielded pool, identified by security researcher Taylor Hornby with the assistance of Anthropic's Claude Opus 4.8 model. The Zcash Foundation stated there is no evidence that the vulnerability was exploited, nor was any unauthorized value creation detected, and user privacy remained unaffected.Anthropic released the first public version of the Claude Mythos model, Fable 5, on Tuesday, and stated on Friday that it has suspended access to the Fable 5 and Mythos 5 AI models due to export control directives issued by the U.S. government citing national security concerns. (Cointelegraph)

Claude Announces Suspension of Fable 5 Access Due to U.S. Government Ban; Developers Urged to Switch Models Promptly

Claude announced that, due to U.S. government-related restrictions, access to the Fable 5 model will be suspended for all users. New sessions will default to the user’s configured default model or Opus 4.8; existing Fable 5 sessions will terminate immediately with an error. Additionally, API requests to Fable 5 on the Claude Platform will also return errors. Developers must promptly migrate their applications and integrations to other Claude models.

B.AI Officially Launches Claude Opus 4.8, Upgrading Coding and Reasoning Capabilities

The B.AI platform has officially launched the Claude Opus 4.8 model. Building on the same pricing as version 4.7, Opus 4.8 significantly enhances coding and reasoning capabilities, achieving several key breakthroughs: a roughly fourfold reduction in code defect detection failure rate; more objective and accurate task progress feedback; and further improved ability to independently execute complex tasks over extended periods—better supporting automated development and deep-reasoning scenarios. Currently, Opus 4.8 is fully available for API integration and Web Chat access. Users can immediately experience it by logging in at http://chat.b.ai/chat or consult the official documentation for more technical details.

Claude partners with SpaceX to significantly boost computing capacity

Anthropic's Claude has announced a partnership agreement with SpaceX to significantly increase its computing capacity and raise usage limits for Claude Code and Claude API. This includes doubling the 5-hour rate limit for Claude Code in Pro, Max, and Team subscription plans; removing Claude Code rate limits during peak hours for Pro and Max plans; and substantially increasing the Opus model API rate limit. Claude will also utilize the full computing resources of the SpaceX Colossus 1 data center, with plans to add over 300 megawatts of deployment capacity within a month.

WLFI Co-Founder Zach Witkoff: WorldClaw and WorldRouter Ushering in a New Era of AI Infrastructure

Odaily, World Liberty Financial co-founder Zach Witkoff appeared on Fox Business to discuss the company's latest breakthroughs in AI and decentralized finance. Witkoff pointed out that WorldClaw is a key piece of the puzzle in realizing the Agent economy vision. WorldRouter, as an AI model aggregation platform, supports access to over 300 mainstream AI models—including GPT-5.5, Claude Opus 4.7, and Gemini 3.1 Pro—via a single account, with pricing 30% lower than official channels. WorldClaw AgentOS integrates the native stablecoin USD1 for instant and transparent settlement. Users can stake WLFI tokens to obtain AI credit packages and unlock whitelist eligibility. The project has the backing of the Trump family, aiming to establish the United States as the global hub for crypto and AI.