Zhiyuan WITA Ends 'Naked' Robot Interaction with First Compliance Filing
The embodied intelligence sector has reached a significant milestone. According to the latest announcement from the Shanghai Cyberspace Administration, the WITA large model developed by Zhiyuan has successfully completed the filing process, becoming the first compliantly deployed embodied intelligence interaction large model in the country.
This achievement goes beyond simply obtaining a license. WITA's core purpose is to enable humanoid robots to truly converse, perceive emotions, and develop distinct personalities. Designed specifically for robot interaction scenarios, it uses natural, emotionally expressive communication to transform cold mechanical bodies into silicon companions with continuous memory and individual traits. As the central engine for interactive intelligent deployment, this model is already being applied in commercial settings such as guided tours, shopping assistance, and service retail, addressing the long-standing industry challenge where robots can work but struggle to communicate effectively.

More importantly, Zhiyuan has revealed that WITA Omni 1.0, the first end-to-end multimodal interaction large model for robots, will launch in the third quarter of this year. Its breakthroughs are tangible: interaction latency has been reduced to under 500 milliseconds, closely matching the rhythm of natural human conversation. It supports continuous communication at normal speech speeds, allows interruptions and corrections, and adjusts emotions and tone in real time, making interactions truly feel like talking to a person.
On the technical side, the new model achieves seamless multimodal coordination across language, voice, facial expression, and movement, eliminating the previously disjointed feeling where a robot's mouth moved but its body remained still. More importantly, through a multimodal interaction data flywheel mechanism, the model continuously learns and improves in real-world scenarios, creating a positive cycle of self-optimization.
The strategic implications are equally significant. At the inaugural Hong Kong Embodied Intelligence Industry Summit, Peng Zhihui, co-founder, president, and CTO of Zhiyuan, officially announced the Zhiyuan 358 Vision Plan: targeting revenue of 10 billion yuan by 2027 and 100 billion yuan by 2030. This ambitious roadmap not only demonstrates Zhiyuan's confidence in the commercialization of embodied intelligence but also signals that the entire sector is accelerating from technological validation to large-scale monetization, marking a critical turning point.
Related article
ByteDance’s Seed launches global campus drive, offering virtual shares to win top large model talent
In the competitive landscape of large language models, securing top-tier talent remains the most critical strategic asset.On April 1st, ByteDance announced the launch of its Seed global campus recruitment initiative, part of its large model talent de
Suno to Watermark Songs Amid Legal Battles
Suno, the platform enabling users to generate AI-created music, has unveiled new features to label platform-produced tracks, restrict downloads, and update community standards to curb unauthorized replicas. These updates arrive as Suno confronts mult
Musk Admits Grok Build Leaked User Code, Promises to Erase All Historical Data
Elon Musk directly addressed the privacy controversy surrounding Grok Build, beginning with a simple "True" to confirm the incident's validity. He pledged that all user data previously uploaded to SpaceXAI would be permanently erased, stating, "not a
Related Special Topic Recommendations
Comments (0)
0/500
The embodied intelligence sector has reached a significant milestone. According to the latest announcement from the Shanghai Cyberspace Administration, the WITA large model developed by Zhiyuan has successfully completed the filing process, becoming the first compliantly deployed embodied intelligence interaction large model in the country.
This achievement goes beyond simply obtaining a license. WITA's core purpose is to enable humanoid robots to truly converse, perceive emotions, and develop distinct personalities. Designed specifically for robot interaction scenarios, it uses natural, emotionally expressive communication to transform cold mechanical bodies into silicon companions with continuous memory and individual traits. As the central engine for interactive intelligent deployment, this model is already being applied in commercial settings such as guided tours, shopping assistance, and service retail, addressing the long-standing industry challenge where robots can work but struggle to communicate effectively.

More importantly, Zhiyuan has revealed that WITA Omni 1.0, the first end-to-end multimodal interaction large model for robots, will launch in the third quarter of this year. Its breakthroughs are tangible: interaction latency has been reduced to under 500 milliseconds, closely matching the rhythm of natural human conversation. It supports continuous communication at normal speech speeds, allows interruptions and corrections, and adjusts emotions and tone in real time, making interactions truly feel like talking to a person.
On the technical side, the new model achieves seamless multimodal coordination across language, voice, facial expression, and movement, eliminating the previously disjointed feeling where a robot's mouth moved but its body remained still. More importantly, through a multimodal interaction data flywheel mechanism, the model continuously learns and improves in real-world scenarios, creating a positive cycle of self-optimization.
The strategic implications are equally significant. At the inaugural Hong Kong Embodied Intelligence Industry Summit, Peng Zhihui, co-founder, president, and CTO of Zhiyuan, officially announced the Zhiyuan 358 Vision Plan: targeting revenue of 10 billion yuan by 2027 and 100 billion yuan by 2030. This ambitious roadmap not only demonstrates Zhiyuan's confidence in the commercialization of embodied intelligence but also signals that the entire sector is accelerating from technological validation to large-scale monetization, marking a critical turning point.
ByteDance’s Seed launches global campus drive, offering virtual shares to win top large model talent
In the competitive landscape of large language models, securing top-tier talent remains the most critical strategic asset.On April 1st, ByteDance announced the launch of its Seed global campus recruitment initiative, part of its large model talent de
Suno to Watermark Songs Amid Legal Battles
Suno, the platform enabling users to generate AI-created music, has unveiled new features to label platform-produced tracks, restrict downloads, and update community standards to curb unauthorized replicas. These updates arrive as Suno confronts mult
Musk Admits Grok Build Leaked User Code, Promises to Erase All Historical Data
Elon Musk directly addressed the privacy controversy surrounding Grok Build, beginning with a simple "True" to confirm the incident's validity. He pledged that all user data previously uploaded to SpaceXAI would be permanently erased, stating, "not a





Home






