Member of Technical Staff - Data About Us Veeda AI is building the next generation of multimodal foundation world models for Physical AI. We're a small, rapid-moving team of engineers and researchers from leading AI labs, tackling some of the most challenging problems at the intersection of AI, robotics, and embodied intelligence. If you're excited about pushing the boundaries of what's possible with Physical AI, you'll have the opportunity to make an outsized impact from day one. Responsibilities Multimodal Ingest: Build ingest for video, lidar, and robot trajectories on Ray Data and Daft, with GPU decode (NVDEC, DALI) and resharding into WebDataset and Lance layouts that stream sequentially rather than seeking per sample. Curation & Filtering: Decide what earns a slot using blur, exposure,
and camera-trajectory scoring plus embedding deduplication over cuVS indexes, and prove each filter with a downstream ablation, not a dataset-size delta. Annotation & Auto-Labeling: Produce the labels the models need, such as VLM captions, camera pose from feed-forward reconstruction (VGGT, MASt3R), and depth and segmentation pseudo-labels, and hold each to a measured error rate against human review. Real & Synthetic Interop: Normalize episodic data across formats such as LeRobotDataset v3, Open X-Embodiment, and RLDS, reconciling action spaces, control rates, and frame timing, and
📌 Member Of Technical Staff - Data (Toronto)
🏢 Veeda AI
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.