Sibich DFM-2 Large Model Officially Launched to Support Smart Industry Development


Zhipu AI launches GLM-5.3-FlashX, with the API also going live, offering a maximum output of 200 tokens/s. It focuses on intelligence, price, and speed, providing high-throughput, low-latency inference for enterprise developers. The predecessor, GLM-5.3-Flash, was previously introduced overseas under the name Ox Alpha. It gained popularity due to its strong intelligence and cost-effectiveness at the same size, with increasing usage volume. Zhipu is now supporting growing demand.
It has been disclosed that SpaceX's AI department, SpaceXAI, is internally discussing the potential acquisition of customer and operational data from struggling or bankrupt startups to train models such as Grok, in order to obtain high-quality external data to enhance performance. This move indicates that it no longer solely relies on Elon Musk's X platform and internal AI mentor data, but is beginning to expand diverse data sources for Grok.
The WeChat AI assistant 'Weixiao' has sparked controversy over privacy concerns and has topped the hot search list. WeChat clarified that claims about directly reading chat records or spying on privacy are inaccurate. The assistant is a native AI, currently in a limited beta test, supporting text and voice interaction, capable of checking weather, express delivery, and takeout. It can also proactively trigger in scenarios such as Moments, official accounts, chats, and short video sections to assist with content summarization.
The Gates Foundation will invest $1 billion over the next two years to promote the development of artificial intelligence, making cutting-edge technology benefit the poorest regions. Gates also shared his views on the global large model landscape, computing power carbon emissions, and Chinese humanoid robots, pointing out that according to different measures, there are currently four leading large model companies in both the US and China.
Ant Group's AI Security Lab opens the internal security guardrail technology called SingProbe. Unlike external review, it reuses the hidden states of the base model for inference, achieving real-time token-level risk assessment with less than 0.5% additional computing power. It can simultaneously perform intent classification, safety detection, and hallucination recognition, and can block risks at the millisecond level before output in high-risk scenarios such as healthcare. It has been adapted to 29 mainstream large models.