Referenz: 

Researcher (m/f/d/x) in Visual Perception for Interactive 3D World

Researcher (m/f/d/x) in Visual Perception for Interactive 3D World

Stellenausschreibung Titel lang - DE
Researcher (m/f/d/x) in Visual Perception for Interactive 3D Worlds
Stellenausschreibung Titel lang - EN
Researcher (m/f/d/x) in Visual Perception for Interactive 3D Worlds
Tätigkeitsbereich
Wissenschaft
Bereich
Erweiterte Realität
Standort
Kaiserslautern
Ansprechpartner
Dr. Alain Pagani
Ansprechpartner Tel.Nr.
+49 631 20575 3530
Adresse
Deutsches Forschungszentrum für Künstliche Intelligenz GmbH (DFKI);Trippstadter Str 122;67663 Kaiserslautern
E-Mail
alain.pagani@dfki.de
Anstellungsart
Vollzeit / Teilzeit
Vertragsart
Befristet
Department Description - DE
3994
Department Description - EN
3965
Laufzeit (Monate)
30
Datum bis
31.10.2026
Ref No
11390-2026

The core activities of the Department Augmented Vision at the German Research Center for Artificial Intelligence (DFKI) in Kaiserslautern lie in the fields of Image Processing and Computer Vision, Image Understanding, Augmented Reality, Virtual Reality, and 3D Reconstruction.

In the context of new national and European research projects, we are looking for highly motivated researchers to work on visual perception for interactive 3D worlds. Our research focuses on learning rich representations of humans, activities, and complex 3D environments using modern vision architectures, multimodal foundation models, world models, and end-to-end transformer-based approaches. Particular emphasis is placed on dynamic scene understanding, multimodal perception, vision-language reasoning, and embodied visual intelligence.

The position is intended for candidates who wish to pursue a PhD and contribute to high-quality scientific publications and research demonstrators.

Ihre Aufgaben

Research and develop novel methods for visual perception and representation learning in interactive 3D environments.

Investigate multimodal perception, dynamic scene understanding, human activity understanding, world models, and vision-language reasoning.

Design and conduct rigorous experimental evaluations and contribute to scientific publications.

Implement research prototypes using modern vision architectures, transformer-based models, and GPU-based computing.

Contribute research components to demonstrators in Extended Reality, digital twins, human-AI interaction, and embodied AI.

Unsere Anforderungen

Master’s degree or equivalent in Computer Science, Artificial Intelligence, Electrical Engineering, Robotics, Mathematics, or a related field.

Strong background in Computer Vision and machine learning.

Very good programming skills in Python and experience with modern research frameworks such as PyTorch.

Experience in some of the following areas: vision transformers, multimodal foundation models, world models, vision-language models, 3D Computer Vision, video understanding, human activity understanding, representation learning, or generative models.

Good understanding of modern neural architectures, large-scale representation learning, and experimental evaluation.

Strong motivation to conduct scientific research and pursue a PhD.

Excellent communication skills, ability to work independently, and willingness to contribute to a collaborative research environment.

Additional desirable qualifications

Experience with one or more of the following topics would be an advantage: end-to-end transformers, multimodal reasoning, video foundation models, Gaussian Splatting, neural scene representations, scene graphs, open-vocabulary perception, visual memory, continual learning, embodied AI, uncertainty-aware perception, event-based vision, or real-time visual systems.

Previous experience with scientific publications, open-source research code, benchmark datasets, or applied research projects is also welcome.

Was Sie erwarten können

We offer excellent working conditions with challenging research topics in an interdisciplinary and international team at an internationally renowned research institute.

The position provides the opportunity to work on current research at the intersection of Computer Vision, multimodal perception, world models, 3D scene understanding, Extended Reality, and artificial intelligence. The successful candidate will contribute to visible research results, international publications, and demonstrators in collaboration with academic and industrial partners.

We expect the researcher to start or continue a PhD at RPTU Kaiserslautern-Landau during the project.

Das Deutsche Forschungszentrum für Künstliche Intelligenz GmbH (DFKI) wurde 1988 als gemeinnützige Public-Private-Partnership (PPP) gegründet. Das DFKI verbindet wissenschaftliche Spitzenleistung und wirtschaftsnahe Wertschöpfung mit gesellschaftlicher Wertschätzung. Das DFKI forscht seit über 35 Jahren an KI für den Menschen und orientiert sich an gesellschaftlicher Relevanz und wissenschaftlicher Exzellenz in den entscheidenden zukunftsorientierten Forschungs- und Anwendungsgebieten der Künstlichen Intelligenz. In der internationalen Wissenschaftswelt zählt das DFKI zu den wichtigsten „Centers of Excellence“.

Schwerbehinderte Bewerberinnen und Bewerber und Gleichgestellte werden bei gleicher Eignung besonders berücksichtigt. Das DFKI beabsichtigt, den Anteil von Frauen im Wissenschaftsbereich zu erhöhen und fordert deshalb Frauen ausdrücklich auf, sich zu bewerben.

Zurück zur Übersicht
Bewerbungsfrist
Fachliche Fragen zu dieser Position beantwortet Ihnen gerne:
Stelle teilen: