Skip to main content
AI Engineer
Back to Jobs

AI Engineer

Kanaria TechTokyoPosted 1 days ago

求人概要

Job Type
正社員
Japanese Level
ビジネス (N2)
Category
Tech & Engineering
ビザサポート

職務内容

About Kanaria Tech Kanaria Tech is a Tokyo-based frontier physical AI lab building the core AI capabilities for physically embodied intelligence. Our flagship technology is the Kanaria Robotic Model (KRM), an embodiment-agnostic, multimodal foundation model that gives robots socially aware and anticipatory navigation. We are starting with AMRs and will extend to other embodiments over time. About the Physical AI Team The Physical AI team owns KRM (the multimodal navigation foundation model) at the core of everything we ship. This is our frontier research and model-development group: it defines KRM's architecture, trains it at scale, and drives the roadmap toward social navigation. The Role We are looking for an AI engineer to help design, train, and advance KRM. You will work at the intersection of foundation models, computer vision, and robot learning. You'll build a model that perceives multimodal scenes, predicts how they will evolve, and produces both robot trajectories and human-readable interpretability outputs. What You'll Do Design, train, and iterate on KRM that ingests vision (plus LiDAR, radar, audio, and depth when available) and outputs navigation signal together with interpretability signals. Develop and improve the observation encoders and the cross-attention world-state representation. Improve the existing decoders that predict semantic segmentation, optical flow, depth, and camera pose for the current and future frames (the model's short-horizon world model). Advance the camera-pose-prediction learning objective that helps to gather training data at scale. Work on per-robot RL policies that adapt the shared model to specific hardware. Drive roadmap items: better world representations, observation/action decoupling, language integration, and feedback loops. Build and scale the large-scale video data pipeline that feeds training. Build a benchmarking system for social robot navigation. What We're Looking For Strong foundation in deep learning and modern neural network architectures (transformers, attention). Hands-on experience training large models in PyTorch (or an equivalent framework). Solid Python engineering skills and comfort working with large-scale data. Background in one or more of: computer vision, multimodal learning, robot learning, or world models. Nice to Have Experience with Vision-Language-Action (VLA) models, world models, video understanding, and 3D geometry / SLAM. Reinforcement learning, especially sim-to-real or robot control. Simulation experience (e.g., Isaac Sim) and synthetic data generation. A research track record (publications) in relevant areas.

類似求人

東京都の関連情報

よくある質問

Does this role offer visa sponsorship?
Yes, this position is listed as offering visa sponsorship.
Is this job remote?
This role is listed as on-site. Location: Tokyo. Employment type: full_time.
What level of Japanese is required?
The listing specifies business.