GetChain News
中简 中繁 EN
GetChain News
Toggle sidebar

Regulation/Compliance

News linked to both this project and an event.

Serenity: US AI Companies Lower Inference Costs to Counter the "Model Distillation" Challenge

"White-Haired Stock Guru" Serenity stated that while some observations in the UBS report hold anecdotal truth, the more noteworthy trend is the increasing number of Chinese-language reports regarding the distillation of Anthropic's models. Currently, many US startups and tech companies are opting to use cheaper Chinese models (such as DeepSeek) in their AI applications, as their unit task costs are significantly lower than those of inference models from Gemini, OpenAI, and Anthropic.Serenity believes this trend, driven by capitalism, creates a "typical paradox"—companies naturally gravitate towards lower-cost solutions, thereby eroding the leading advantage of US models. He proposes that the US needs to address this on two fronts:First, build stronger access control and authentication systems, such as "heavy KYC frontier models" for domestic US use and tiered access mechanisms for allies, to reduce the risk of model distillation and misuse. This could also be accompanied by introducing an identity verification system akin to "AI-grade banking authentication" (e.g., biometrics + short-lived permission tokens) to raise the barrier for model calls, and using regulatory measures to restrict account sharing and access resale.Second, enhance the cost efficiency of inference models, allowing them to comprehensively outperform competitors like DeepSeek in both price and performance.Serenity also noted that some high-end models are currently frequently targeted for "distillation exploitation." Ideally, access to models nearing the AGI level should involve increased friction costs. In summary, the core challenge for the US AI industry lies in achieving both "low-cost inference capabilities" and establishing model access security mechanisms comparable to those in the financial system.

Sources: NVIDIA plans to pitch Vera AI CPU to Chinese clients, some cloud providers eyeing test deployment

sources say NVIDIA has begun pitching its first independent central processing unit (CPU) product, Vera, to Chinese clients. Designed specifically for Agentic AI systems, the chip has entered mass production, marking NVIDIA's attempt to further expand its presence in the Chinese market with a CPU offering.According to sources, some Chinese clients have already shown interest in Vera. One major Chinese cloud computing company plans to procure over 300 servers equipped with dual Vera CPUs for testing, and will decide whether to expand procurement after the tests are completed.Built on the Arm Holdings architecture, Vera is NVIDIA's first independent CPU product. NVIDIA has previously stated that Vera's performance in AI agent-related computing tasks is 1.8 times that of comparable competitor products, and expects the product to contribute approximately $20 billion in revenue by the end of this fiscal year (ending January next year).The report notes that as the AI industry's focus gradually shifts from model training to inference computing, CPUs and custom chips are gaining more attention. Vera also positions NVIDIA to directly compete with Intel and Advanced Micro Devices (AMD), which have long dominated the server CPU market.Sources indicate that due to strict U.S. export restrictions on high-end GPUs, CPUs face relatively smaller regulatory hurdles in the Chinese market compared to GPU products. Currently, some Chinese clients plan to first deploy Vera chips for testing in overseas data centers. Meanwhile, software ecosystem compatibility and existing domestic AI chip deployment frameworks may still impact the subsequent large-scale adoption of Vera. (Reuters)