AI Engineer
Makana Partners ·www.makanapartners.com
Apply directIntroduction
We are currently supporting a confidential search for an AI & Multimodal Machine Learning Engineer within the robotics industry. This role will lead the development and fine-tuning of cutting-edge Vision-Language and action-oriented multimodal models. It offers an incredible opportunity to shape an open, scalable data ecosystem using real-world operational data.
About the Role
Operating as a specialized AI Machine Learning Engineer within the robotics industry, you will drive the architecture, training pipelines, and deployment of advanced Vision-Language and multimodal models. You will translate state-of-the-art research into scalable production systems while working closely with cross-functional software and hardware engineering teams.
Key Responsibilities
Model Development & Research
- Design, implement, and fine-tune Vision-Language models and related multimodal architectures for real-world robotics applications.
- Continuously research leading academic trends, select appropriate technologies, and integrate new findings into core model performance.
- Perform structured error analysis and iterative performance tuning to maximize model quality.
Pipeline & Data Engineering
- Build scalable, reliable training pipelines designed for continuous learning cycles with real-world operational data.
- Preprocess and manage diverse datasets encompassing image, video, natural language, and robot action parameters.
- Establish robust evaluation metrics to measure operational performance effectively.
Infrastructure & Cross-Functional Collaboration
- Set up scalable training, inference, and experimental environments to ensure seamless deployment to production setups.
- Partner with software and robotics engineers to define technical requirements, structure validation plans, and drive development forward.
Requirements
- Experience: Minimum 3+ years of experience in developing, deploying, and operating machine learning models in production environments using Python and PyTorch.
- Industry Background: Solid hands-on experience within the robotics industry, artificial intelligence research, or related advanced hardware/software sectors.
- Technical Expertise: Direct hands-on experience fine-tuning Vision-Language or multimodal models, along with building real-world data pipelines and MLOps loops.
- Language: Fluent English communication skills (Japanese language ability is not required).
- Soft Skills: Collaborative mindset, strong analytical problem-solving skills, and the capacity to convert academic research into practical, commercial solutions.
Why Apply
- High autonomy in driving research and operational ML implementations.
- Opportunity to work with cutting-edge, open-source AI and large-scale data ecosystems.
- Collaborative and multidisciplinary engineering environment combining software, AI, and physical hardware.
- Convenient central location in Tokyo.