WorkshopOnline

Sept 9 - Physical AI Has a Data Problem. It Isn't Collection Workshop

SEP9
21:30 · Asia / Kolkata

About this event

This workshop goes from raw recording to curated corpus. We'll cover what MCAP is and why it's built that way, tour real Physical AI datasets across driving, aquatic, and forest robots, and open an episode with every sensor synced—including channels nothing knows how to decode.

Time, Date and Location

Sep 09, 2026
9:00 AM - 10:00 AM PST
Online. Register for the Zoom!

Physical AI still relies on familiar computer vision tasks—detection, segmentation, depth, tracking. What's changed is the data unit: no longer a single image and label, but an episode—a dozen sensors ticking on independent clocks for minutes, with no frame boundaries.

Most computer vision tooling assumes the old unit and breaks on the new one.

That's why Physical AI teams end up with buckets of .mcap files nobody can characterize. Recording is cheap, so logs pile up faster than anyone curates them. Ask what's actually in there—which tasks, which conditions, how many failures and of what kind—and the honest answer is usually a shrug.
MCAP has been ROS 2's default log format since Iron, and as of FiftyOne 1.19 it opens natively: cameras, LiDAR, GPS, IMU, and logs on one shared timeline, alongside your images and video.

We'll tackle quality, the harder half: what smoothness, sensor-health, and outlier metrics actually measure, where each falls short, and how to turn a score into a defensible decision.

You'll leave knowing how to load your own recordings, query a whole corpus instead of a single file, and which quality signals to trust for which job.

More upcoming events in Germany

OCT
01

Berliner Bitcoin-Stammtisch @ Friedel Richter

Friedel Richter Restaurant, Berlin, Germany · Bitcoin Lab Berlin
In person

Welcome to the Berliner Bitcoin-Stammtisch, where we gather to talk about the digital currency that's worth more than gold and sometimes causes more drama than a Netflix series. Join our tribe of business beings, developers, and activists who believe that Bitcoin is the future of money, and El Salvador seems to agree! Here you can learn the ropes of Bitcoin, from buying your first fraction to safely storing your private keys. Our current hangout spot is Friedel Richter on Torstraße, because ROOM77 is closed until the last block is mined. You can pay in Bitcoin ₿ or Lightning ⚡ and enjoy awesome regional food with an international twist. So come join the fun, the food, and the revolution! We're waiting for you. 🍔🍔🍔

19:00 · WebinarDetails →
OCT
02

Oct 1 - APAC AI, ML and Computer Vision Meetup

Online, Germany · München AI, Machine Learning and Computer Vision Meetup
Online

Join our APAC time-zone friendly virtual meetup to hear talks from experts on cutting-edge topics across AI, ML, and computer vision. Time, Date and Location Oct 1, 2026 6:00 PM - 8:00 PM PDT Online. Register for the Zoom! Beyond Exact Matches: Detecting Modified 3D Assets at Marketplace Scale How can a marketplace identify copied 3D assets when their orientation, geometry, or composition has changed? Drawing on my work in 3D content understanding at Roblox, this talk will explore multi-view and rotation-invariant representations for similarity and duplicate detection, including the challenges posed by deformed and fragmented copies. It will examine how geometric and semantic signals can complement one another, and discuss practical trade-offs in evaluating detection quality and deploying these methods at scale. The presentation will draw on published patent applications and publicly shareable examples to offer practical lessons for engineers building visual search, content-understanding, and marketplace-safety systems. About the Speaker Phani Harish Wajjala is a Principal Machine Learning Engineer at Roblox specializing in 3D computer vision, multimodal AI, and large-scale content understanding. Sign Language: Towards Sign Understanding for Robot Autonomy Navigational signs are common aids for human wayfinding and scene understanding, but are underutilized by robots. We argue that they benefit robot navigation and scene understanding, by directly encoding privileged information on actions, spatial regions, and relations. Interpreting signs in open-world settings remains a challenge owing to the complexity of scenes and signs, but recent advances in vision-language models (VLMs) make this feasible. To advance progress in this area, we introduce the task of visual sign grounding, which parses locations and associated directions from signs, and maps them to region in the sign’s local environment. Additionally, we present a baseline approach using VLMs, and d

06:30 · MeetupDetails →
OCT
02

re:publica Vienna

Germany · re:publica
runs 2–3 October 2026
  • re:publica Vienna is a conference held in Germany that brings together various stakeholders interested in digital culture, technology, and society.
  • The event is aimed at founders, operators, and other professionals looking to explore current trends and topics in the tech ecosystem.
  • Attendees can expect discussions and networking opportunities to connect with like-minded individuals.
Policy & EcosystemFoundersTalks & PanelsOperators
ConferenceDetails →
Share this event