AI Engineer

  • Tokyo
  • Partial Remote
  • Full-time
  • August 18, 2026
Conditions
location-icon
Apply from Anywhere 👍
visa-icon
Relocation to Japan 👍
(Overseas visa sponsorship supported)
Requirements
language-icon
Language Requirements
Japanese: Business Level
English: Business Level
career-icon
Minimum Experience
Mid-level or above

About Kanaria Tech

Kanaria Tech is a Tokyo-based frontier physical AI lab building the core AI capabilities for physically embodied intelligence. Our flagship technology is the Kanaria Robotic Model (KRM), an embodiment-agnostic, multimodal foundation model that gives robots socially aware and anticipatory navigation. We are starting with AMRs and will extend to other embodiments over time.

 

About the Physical AI Team

The Physical AI team owns KRM (the multimodal navigation foundation model) at the core of everything we ship. This is our frontier research and model-development group: it defines KRM's architecture, trains it at scale, and drives the roadmap toward social navigation.

 

The Role

We are looking for an AI engineer to help design, train, and advance KRM. You will work at the intersection of foundation models, computer vision, and robot learning. You'll build a model that perceives multimodal scenes, predicts how they will evolve, and produces both robot trajectories and human-readable interpretability outputs.

 

What You'll Do

  • Design, train, and iterate on KRM that ingests vision (plus LiDAR, radar, audio, and depth when available) and outputs navigation signal together with interpretability signals.
  • Develop and improve the observation encoders and the cross-attention world-state representation.
  • Improve the existing decoders that predict semantic segmentation, optical flow, depth, and camera pose for the current and future frames (the model's short-horizon world model).
  • Advance the camera-pose-prediction learning objective that helps to gather training data at scale.
  • Work on per-robot RL policies that adapt the shared model to specific hardware.
  • Drive roadmap items: better world representations, observation/action decoupling, language integration, and feedback loops.
  • Build and scale the large-scale video data pipeline that feeds training.
  • Build a benchmarking system for social robot navigation.

 

What We're Looking For

  • Strong foundation in deep learning and modern neural network architectures (transformers, attention).
  • Hands-on experience training large models in PyTorch (or an equivalent framework).
  • Solid Python engineering skills and comfort working with large-scale data.
  • Background in one or more of: computer vision, multimodal learning, robot learning, or world models.

 

Nice to Have

  • Experience with Vision-Language-Action (VLA) models, world models, video understanding, and 3D geometry / SLAM.
  • Reinforcement learning, especially sim-to-real or robot control.
  • Simulation experience (e.g., Isaac Sim) and synthetic data generation.
  • A research track record (publications) in relevant areas.

Kanaria Tech is a leading physical AI company in Japan building its own AI navigation model — one that works in any wheeled robot.

The company is making it effortless for robots to live among people, moving through human spaces with minimal disruption. Their customized robot, Wakko, has already been deployed at Kusatsu City Hall.

View Kanaria Tech's company page
AI Engineer at Kanaria Tech
APPLY NOW  ➜Japanese Required ⚠️