Asia
A new, inexpensive Chinese AI model is catching up with Anthropic, OpenAI on their home turf
Zhipu's AI service on the web, dubbed Z.ai. BEIJING/BENGALURU - Since DeepSeek shocked markets early last year with its cheap but powerful artificial intelligence model, global consumers have been faced with a choice: Chinese offerings with lower prices and less capability or OpenAI or Anthropic, which have poured billions into development. A model called GLM-5.2, launched last month by Beijing-based startup Z.ai, may finally be closing that gap in terms of Western interest. GLM-5.2 has Silicon Valley buzzing with its coding and agent capabilities, or the ability to execute complex tasks with minimal prompting, that almost rival leading U.S. offerings at a fraction of the cost, in what some experts are calling a "mini DeepSeek moment." In a time of both misinformation and too much information, quality journalism is more crucial than ever.
The Dual Nature of LLM Persona: Aggregated Tendencies and Frame-Dependent Geometry
Evaluations of LLM personas via psychometric questionnaires typically rely on aggregate scores, discarding within-instance correlation structure. We test whether this geometric structure is intrinsic or frame-dependent. Constructing within-instance correlation matrices from IPIP-50 responses, we analyze geometry on SPD manifolds under manipulated question orderings in GPT-4o simulating American and Chinese-American personas. We find that persona expression comprises two dissociable components: aggregated features (Big Five scores) degrade under randomization (21% drop) but are frame-robust; geometric features (SPD manifold) collapse under frame misalignment (42% drop) but recover substantially (to 84%) under shared frames, surpassing aggregated features (76%). This collapse-recovery pattern reveals that persona geometry is not intrinsic but a frame-dependent coordination pattern encoding information invisible to aggregation. Our findings establish a dual-nature framework for LLM personas, frame-dependent geometry versus frame-robust aggregates, necessitating frame-aware evaluation and challenging static trait conceptions.
Online Safety Monitoring for LLMs
Schirmer, Mona, Jazbec, Metod, Timans, Alexander, Naesseth, Christian, Waldron, Maja, Nalisnick, Eric
We deploy a simple into our everyday lives as search engines (Jin et al., 2025; statistical framework based on risk control (Angelopoulos Xiong et al., 2024), coding assistants (Zhao et al., 2023), et al., 2022) that converts any safety signal into a binary and companions (Zhang et al., 2025a). As their applicability grows, so does the potential harm caused by malicious decision rule, and offers statistical guarantees on the false LLM outputs. Despite remarkable performance across a alarm or missed detection rate. The framework is universally applicable to different monitoring purposes and can leverage wide range of tasks, LLMs remain prone to generating halarbitrary proxy signals. Through experiments on mathematlucinated, factually incorrect (Ravichander et al., 2025), or ical problem solving and red teaming conversations, we harmful output (Yu et al., 2025) when deployed.
Reflections from ICRA 2026
From the 1st-5th June, the robots descended on Vienna. The 2026 IEEE International Conference on Robotics & Automation (ICRA) brought together the top minds in robotics for one short week to showcase the latest technologies, form new collaborations, and exchange ideas. Held at the Messe Wien, a stone's throw from the bank of the Danube, ICRA proved to be equal parts technological marvel and thought-provoking discussion. The host venue for ICRA 2026: Messe Wien, also known as VIECON. My week at ICRA began with the 2nd ICRA 2026 Workshop on Robot Ethics: Ethical, Legal and User Perspectives in Robotics & Automation (WOROBET) .
The Download: a startup has a solution for AI's groupthink problem
The Download: a startup has a solution for AI's groupthink problem Plus: Scientists say they have built a cell from scratch for the first time. LLMs are stuck in a groupthink groove. This startup is trying to get them out. Open up your chatbot of choice--Claude, ChatGPT, Gemini--and type "Give me a random number between 1 and 10." You're going to get 7. Almost always. That won't work every time--but if it did for you, you may wonder if I have superpowers. The truth is that most large language models are stuck in a rut.
Finance Minister Katayama says G7 will discuss AI defense standards
Finance Minister Satsuki Katayama speaks during an interview on Monday. The Group of Seven nations will discuss standards on artificial intelligence security and defense, Finance Minister Satsuki Katayama has said. Speaking in a recent interview, Katayama said that financial institutions "need to decide the order of priority for fixing their systems," in order to prepare for the possibility of advanced AI models detecting a large number of vulnerabilities in their systems. She added that the G7 nations, which include Japan, will discuss related criteria and work together to tackle cyberattacks. State-of-the-art AI models, such as Claude Mythos, developed by U.S. startup Anthropic, are believed to be highly proficient in identifying system vulnerabilities. Katayama has been negotiating with the United States to ensure that major financial institutions in Japan have access to these technologies.