Technology

How a voice becomes a space

From speech recognition to multi-surface LED output, every stage runs on our own pipeline.

Concept film
Pipeline

What happens in 30 seconds

Scroll to move through each stage.

Concept still
1 / 6
01

Listen

Speech-to-Text (STT) turns what the visitor says into text.

02

Understand

A lightweight language model converts the request into a scene, while a safety filter screens out inappropriate input.

03

Draw

Generative AI creates a scene image that fits the venue's theme.

04

Animate

The still image becomes living, moving video.

05

Accelerate

Real-time inference optimization keeps waiting time short.

06

Surround

The scene plays seamlessly across multiple LED surfaces at once.

Core

Core technology

Generation pipeline orchestration serverConcept still
Core 01

Generation pipeline orchestration server

Ties speech recognition, the language model, image and video generation and output into a single managed flow.

Copyright C-2026-047235

Real-time rendering client engineConcept still
Core 02

Real-time rendering client engine

Plays generated scenes seamlessly and simultaneously across multiple LED surfaces.

Copyright C-2026-047234

Automatic character animationConcept still
Core 03

Automatic character animation

Builds a skeleton for photos or 3D models automatically and generates motions such as walking and running in real time.

Copyright C-2026-047236 · Patent pending 10-2026-0188398

Persona engineConcept still
Core 04

Persona engine

Learns from a figure's records to give evidence-based answers, synchronized with voice and facial expressions.

Persona AI

Before / After

Before and after

Drag the handle to compare.

After "take me under the sea"
Before speaking
After "take me under the sea"Before speaking
Safety & Ops

Safety & operations

On-premise GPU edge server

Generation happens on site with no external cloud. Visitors' voices and photos never leave the venue.

Content safety filter

An allow-list-based filter and output review keep operation within public-space standards.

Remote, unattended operation

Remote monitoring lets us check status and swap content without on-site staff.

Roadmap

R&D roadmap

NOW

Commercial

  • Live MetaCube
  • Lumi Town
  • Persona AI
NEXT

Advancing

  • Multi-person motion interaction
  • More seasonal themes
  • Faster generation
LATER

Research

  • Photoreal space reconstruction with 3DGS
  • Multilingual voice dialogue

3DGS: 3D Gaussian Splatting — reconstructing 3D spaces from real photographs

Looking for joint research or a technology partner?

Contact us