Ayça Takmaz

profile photo

I am currently at Flexion, where I work on long-horizon autonomy for humanoid robots, building the spatial, embodied perception and reasoning abilities that let them carry out full missions on their own. Another core aspect I focus on is turning real-world captures and generated worlds into photorealistic, interactive world representations in which robots can train and be evaluated before acting in the physical world.

My PhD at ETH Zurich focused on how intelligent agents can understand and interact with complex 3D environments, not just recognizing what's in a scene, but reasoning about its geometry, semantics, and spatial relationships well enough to act in it, using language as a way to query and interact with these representations. I was supervised by Prof. Bob Sumner and Prof. Siyu Tang, working closely with Francis Engelmann. During my PhD, I also interned at Google with Johanna Wald and Federico Tombari, and at NVIDIA with Aljoša Ošep and Laura Leal-Taixé.

Prior to that, I did my Master's at ETH Zurich in computer vision and machine learning, working on research projects with Prof. Luc Van Gool and Prof. Marc Pollefeys.

I have also been co-organizing ZurichAI's Vision and Robotics branches since 2024, and the OpenSUN3D workshop series at ICCV, CVPR, and ECCV since 2023.

ETH Zurich Google NVIDIA Flexion

News

Selected Work

Flexion Reflect v1.0 teaser
Flexion Reflect v1.0 - The Path Towards Long-Horizon Autonomous Humanoid Work Flexion Team Autonomous long-horizon mission platform for humanoid robots Technical Blog, 2026 read more
Understanding the 3D World for Intelligent and Interactive Agents teaser
Understanding the 3D World for Intelligent and Interactive Agents Ayça Takmaz Multimodal perception and 3D scene understanding for intelligent, interactive agents PhD Thesis, 2025 thesis
Towards Learning to Complete Anything in Lidar teaser
Towards Learning to Complete Anything in Lidar Ayça Takmaz, Cristiano Saltori, Neehar Peri, Tim Meinhardt, Riccardo de Lutio, Laura Leal-Taixé, Aljoša Ošep International Conference on Machine Learning (ICML) 2025 project page / paper
Search3D teaser
Search3D: Hierarchical Open-Vocabulary 3D Segmentation Ayça Takmaz, Alexandros Delitzas, Robert W. Sumner, Francis Engelmann, Johanna Wald, Federico Tombari IEEE Robotics and Automation Letters (RA-L) 2025 project page / paper
Segment3D teaser
Segment3D: Learning Fine-Grained Class-Agnostic 3D Segmentation without Manual Labels Rui Huang, Songyou Peng, Ayça Takmaz, Federico Tombari, Marc Pollefeys, Shiji Song, Gao Huang, Francis Engelmann European Conference on Computer Vision (ECCV) 2024 project page / paper
SceneFun3D teaser
SceneFun3D: Fine-Grained Functionality and Affordance Understanding in 3D Scenes Alexandros Delitzas, Ayça Takmaz, Federico Tombari, Robert Sumner, Marc Pollefeys, Francis Engelmann Computer Vision and Pattern Recognition (CVPR) 2024 (Oral) project page / paper
OpenMask3D teaser
OpenMask3D: Open-Vocabulary 3D Instance Segmentation Ayça Takmaz*, Elisabetta Fedele*, Robert W. Sumner, Marc Pollefeys, Federico Tombari, Francis Engelmann Conference on Neural Information Processing Systems (NeurIPS) 2023 project page / paper
3D Segmentation of Humans in Point Clouds teaser
3D Segmentation of Humans in Point Clouds with Synthetic Data Ayça Takmaz*, Jonas Schult*, Irem Kaftan, Mertcan Akcay, Bastian Leibe, Robert Sumner, Francis Engelmann, Siyu Tang International Conference on Computer Vision (ICCV) 2023 project page / paper
Unsupervised Monocular Reconstruction of Non-Rigid Scenes teaser
Unsupervised Monocular Reconstruction of Non-Rigid Scenes Ayça Takmaz, Danda Pani Paudel, Thomas Probst, Ajad Chhatkuli, Martin R. Oswald, Luc Van Gool International Conference on 3D Vision (3DV) 2021 project page / paper