T
Not Specified Permanent

Austin, Texas · USA job

Staff Forward Deployed Engineer

Tenstorrent

Austin, Texas

Job description

Overview

In this role you will bridge customers, engineering, and AI inference services to drive production outcomes on Tenstorrent AI systems. You'll deploy and operate production code with high autonomy, explaining trade-offs to both customer leadership and engineering teams. You'll work directly with customers to diagnose challenges across the inference stack and contribute measurable improvements. This is a hands-on, impact-driven engineering position at a remote-first company with hubs in North America.

Compensation / Benefits
  • competitive compensation package
  • remote work flexibility (North America)
  • equal opportunity employer
  • global hub locations
Responsibilities
  • Contribute production code and operate deployments for AI inference workloads
  • Serve as the continuity link between customers, engineering, and AI inference products
  • Explain trade-offs and present recommendations to both customer leaders and engineering teams
  • Debug across the full inference stack from requests to serving layer and kernel dispatch when needed
  • Provide feedback via pull requests, reproducible code, benchmarks, and telemetry data
  • Collaborate with customers to understand challenges and craft effective solutions
  • Scale disaggregated inference services on Kubernetes while balancing performance and reliability
Key requirements
  • 5+ years of relevant software engineering experience (e.g., Applied Engineer, ML/AI/Platform/Infrastructure/SRE roles)
  • Kubernetes and Helm experience at multi-node, HPC, or AI cluster scale
  • Experience with observability and automation (Prometheus, Grafana, OpenTelemetry)
  • Experience with LLM inference serving engines and technologies (e.g., vLLM, SGLang, Mooncake, NIM, Dynamo, LMCache)
  • Customer-focused communication
  • Collaborative problem-solving
  • Ability to translate ambiguous requirements into acceptance criteria
  • Kubernetes and Helm at scale
  • Observability tooling (Prometheus, Grafana, OpenTelemetry)
  • LLM inference serving technologies (vLLM, SGLang, Mooncake, NIM, Dynamo, LMCache)

Explore related USA jobs

Similar jobs you may like

Related roles with a similar title and location.

Same category Same area Same location Same country

ASIC Design Engineer II, Annapurna Labs - Cloud-Scale Machine Learning Acceleration

Annapurna Labs (U.S.) Inc.

Austin, Texas, United States · 78716

Same category Same area Same location Same country

DFT Design Engineer, Machine Learning Acceleration

Annapurna Labs (U.S.) Inc.

Austin, Texas, United States · 78716

Same category Same location Same country

Electrical Superintendent

Faith Technologies

Fort Stockton, Texas, United States · 79735

Same category Same location Same country

Electrical Superintendent

Faith Technologies

El Paso, Texas, United States · 88568

Same category Same location Same country

Engineering Operation Technician Night Role

Amazon Data Services, Inc.

Lubbock, Texas, United States · 79430

Same category Same location Same country

Product Engineer

A.O. Smith

Haltom City, Texas, United States · 76117

Same category Same location Same country

Critical Infrastructure Mechanical Engineer, Field Engineering

Amazon Data Services, Inc.

Wink, Texas, United States · 79789

Same category Same location Same country

Plant Engineer II

Calpine Operating Services Company

Midland, Texas, United States · 79707

Need help finding a job?

Chat with our AI assistant.