Staff Forward Deployed Engineer

Tenstorrent
Remote, North America / Santa Clara, CA / Austin, TX2026-08-06

About the job

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities.

We’re looking for a Forward Deployed Engineer who’s excited to build with the engineers using the AI computers Tenstorrent makes. You will create continuity between customers, engineering, and AI inference service products. This is an engineering role first: you contribute production code, operate deployments, and you can explain a trade-off to customer leadership as clearly as to core engineering teams. This is a high-autonomy role with direct customer impact.

Responsibilities

- Contribute production code and operate deployments

- Explain trade-offs to customer leadership and core engineering teams

- Work directly with customers to understand their challenges and provide effective solutions

- Debug across the full inference stack: from failing requests, through the serving layer, down to OOMs or kernel dispatch

- Bring feedback in the form of pull requests, reproducible code, benchmarks, and telemetry data

- Turn ambiguous customer requirements or issues into verifiable acceptance criteria

Qualifications

Minimum

- Strong software engineering skills with 5+ years of relevant technical experience (e.g. Applied Engineer, Machine Learning Engineer, MLOps Engineer, Platform Engineer, Infrastructure Engineer, Site Reliability Engineer, Field Application Engineer)

- Experience turning ambiguous customer requirements or issues into verifiable acceptance criteria

- Kubernetes and Helm experience at multi-node, HPC, or AI cluster scale

- Experience with observability and infrastructure automation, e.g. Prometheus, Grafana, OpenTelemetry

- Experience with LLM inference serving engines and technologies, e.g. vLLM, SGLang, Mooncake, NIM, Dynamo, LMCache

Preferred

No preferred qualifications listed.