Share this job
AI Engineer
Tokyo, Tokyo, Japan
Apply for this job

Introduction

We are currently supporting a search for an AI & Multimodal Machine Learning Engineer within the robotics industry. This role will lead the development and fine-tuning of cutting-edge Vision-Language and action-oriented multimodal models. It offers an incredible opportunity to shape an open, scalable data ecosystem using real-world operational data.


About the Role

Operating as a specialized AI Machine Learning Engineer within the robotics industry, you will drive the architecture, training pipelines, and deployment of advanced Vision-Language and multimodal models. You will translate state-of-the-art research into scalable production systems while working closely with cross-functional software and hardware engineering teams.


Responsibilities

Model Development & Research

  • Design, implement, and fine-tune Vision-Language models and related multimodal architectures for real-world robotics applications.
  • Continuously research leading academic trends, select appropriate technologies, and integrate new findings into core model performance.
  • Perform structured error analysis and iterative performance tuning to maximize model quality.

Pipeline & Data Engineering

  • Build scalable, reliable training pipelines designed for continuous learning cycles with real-world operational data.
  • Preprocess and manage diverse datasets encompassing image, video, natural language, and robot action parameters.
  • Establish robust evaluation metrics to measure operational performance effectively.

Infrastructure & Cross-Functional Collaboration

  • Set up scalable training, inference, and experimental environments to ensure seamless deployment to production setups.
  • Partner with software and robotics engineers to define technical requirements, structure validation plans, and drive development forward.

Requirements

  • Experience: Minimum 3+ years of experience in developing, deploying, and operating machine learning models in production environments using Python and PyTorch.
  • Industry Background: Solid hands-on experience within the robotics industry, artificial intelligence research, or related advanced hardware/software sectors.
  • Technical Expertise: Direct hands-on experience fine-tuning Vision-Language or multimodal models, along with building real-world data pipelines and MLOps loops.
  • Language: Fluent English communication skills (Japanese language ability is not required).
  • Soft Skills: Collaborative mindset, strong analytical problem-solving skills, and the capacity to convert academic research into practical, commercial solutions.


Why Apply

  • High autonomy in driving research and operational ML implementations.
  • Opportunity to work with cutting-edge, open-source AI and large-scale data ecosystems.
  • Collaborative and multidisciplinary engineering environment combining software, AI, and physical hardware.
  • Convenient central location in Tokyo.


Apply for this job
Powered by