
Compensation
Salary undisclosedDescription
We’re looking for a Forward Deployed Engineer to own the technical outcome of the AI systems we build with our customers. You will work alongside customer engineering and operations teams to make complete systems work in production or near-production environments. This engineering role spans hardware and software integration, deployment, validation, automation, debugging, and ongoing operations. You will contribute production code, establish reliable operating procedures, and drive customer systems through resolution until they work as intended.
This role is Hybrid, based out of Tokyo, Japan.
We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting.
Who You Are
- You take ownership of customer outcomes across system integration, validation, production readiness, and ongoing operations.
- You understand how accelerator compute, memory, and networking topology constrain AI workloads, and do not treat hardware as a black box.
- You work directly with customer engineering and operations teams to understand their environments, constraints, and desired outcomes.
- You are comfortable debugging across the inference stack and communicating technical trade-offs clearly with customer leadership and core engineering teams.
What We Need
- Strong software engineering skills with 5+ years of relevant experience in applied engineering, machine learning engineering, MLOps, platform engineering, infrastructure engineering, or site reliability engineering.
- Proficiency in Python; C++ experience is a plus.
- Experience turning ambiguous requirements into production-ready implementations, acceptance criteria, validation plans, and operating procedures.
- Experience with Kubernetes and Linux systems administration at multi-node, HPC, or AI-cluster scale, including Helm-based deployments, observability, infrastructure automation, CI/CD, release engineering, and LLM inference frameworks such as vLLM, SGLang, or Mooncake.
- Bilingual who is native or business level Japanese with Business level English speaker.
What You Will Learn
- How co-design across AI hardware and software translates silicon capabilities into measurable latency and throughput gains.
- How to operate disaggregated inference systems on Kubernetes while balancing performance, reliability, maintainability, and cost.
- How to take enterprise AI deployments from requirements through integration, software delivery, validation, production readiness, and ongoing ownership.
- How customer deployment and operational insights shape product improvements and platform direction.
Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer.
Stack
- Posted
- Sep 30, 2026
- Last seen
- Sep 30, 2026
- First seen
- Sep 30, 2026


