Real-time AI at the Edge: Portable Inference and Interactive Demos
Tokyo AI evening on real-time edge AI: portable inference with ZML, real-time demo engineering and on-device apps.
- When
- Wed, December 16, 2026 · 18:00–21:00 JST
- Where
- Tokyo, Japan · In person
- Region
- Kanto (Tokyo)
- Organizer
- Tokyo AI
- Language
- EN
- Source
- Luma
Summary
Tokyo AI (TAI) hosts a practical technical evening on making AI feel immediate on real hardware. Three proposed 30-minute talks with Q&A cover portable low-latency inference with ZML, the systems engineering behind real-time AI demos such as TORQ and interactive video generation, and integrating on-device inference into responsive applications.
The session is aimed at ML engineers, systems developers and people building edge AI products, with a focus on runtime choices, memory constraints and measured end-to-end latency. Speakers, final talk scopes, the date and the Tokyo venue are all provisional. The evening ends with an hour of networking.
About the community
Tokyo AI is a large international AI community based in Tokyo, bringing together engineers, researchers, investors and product people. It runs frequent English-language technical talks, panels and networking evenings connecting Japan's AI builders with the global ecosystem.
#edge-ai#inference#on-device-ai#zml#latency#ml-systems