Sumeru AI Updates Mugen3D to Turn a Single Photo into a Live 3D Teacher
Back to newsroom
Product · EducationJun 29, 2026

Sumeru AI Updates Mugen3D to Turn a Single Photo into a Live 3D Teacher

Deployed across multiple universities — one photo and a voice sample become a talking, gesturing 3D teacher that responds in real time.

Shenzhen, China — Sumeru AI announced a major update to Mugen3D, its real-time interactive 3D content engine, enabling users to turn one photograph and voice sample into a talking, emoting 3D human ready for live conversation. The technology is already deployed across university classrooms in China, marking an early operational use of geometry-based spatial AI in education.

A user uploads one photograph. Within minutes, proprietary geometric algorithms combined with 3D Gaussian Splatting (3DGS) produce a true-to-source model preserving facial structure, hair, fabric texture and surface lighting at 4K resolution. A single unified pipeline handles generation across humans, objects and scenes.

The update shifts Mugen3D from reconstructing the body to enabling real-time interaction. SumeruAI, the company's interaction engine, connects the 3D figure to voice input, multilingual dialogue, role-based knowledge and a speech-to-face animation pipeline with under 150 milliseconds of latency. The result is not a prerendered loop or rigid avatar, but a live persona that speaks, reacts, answers questions and converses.

In its latest public demo, Sumeru AI showcases a Math Teacher persona generated from a real photograph. The avatar explains concepts, answers live questions and switches languages on demand. Institutions already using the platform include Beijing Institute of Technology, Shanghai Jiao Tong University, Shenzhen University, and Hong Kong University of Science and Technology. Nearly 1,000 educators and tens of thousands of students now interact with Sumeru AI's digital teaching personas.

'Most image-to-3D tools stop at the render. We treat that as the starting point,' said Dr. Cheng Feng, CEO of Sumeru AI. 'World models cannot be built on flat video. Reality is 3D. Whether training robots in simulated rooms or teaching students in virtual classrooms, you need geometry with physical bounds, not hallucinated pixels. A 3D teacher can explain a concept, answer a follow-up and remain with a student until the idea sticks. That is presence, not content.'

Mugen3D replaces manual sculpting, rigging and animation with a photo-driven workflow that cuts production time while maintaining precise visual reconstruction. The pipeline supports Unity, Unreal Engine and WebGL for education, robotics simulation, spatial computing, 3D printing and interactive entertainment. Full coverage on Markets Insider.