Tenstorrent Inc. Logo

Tenstorrent Inc.

Staff Forward Deployed Engineer

Posted 2 Months Ago
In-Office
Austin, TX, USA
100K-500K Annually
Senior level
In-Office
Austin, TX, USA
100K-500K Annually
Senior level
Build and operate AI inference deployments, write production code, debug full inference stack, collaborate with customers and engineering, and drive performance/observability improvements using Kubernetes, Helm, and LLM serving tools.
The summary above was generated by AI

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities.

We’re looking for a Staff Forward Deployed Engineer who’s excited to build with the engineers using the AI computers Tenstorrent makes. You will create continuity between customers, engineering, and AI inference service products. This is an engineering role first: you contribute production code, operate deployments, and you can explain a trade-off to customer leadership as clearly as to core engineering teams. This is a high-autonomy role with direct customer impact.

This role is remote, based out of North America, with preference near one of our main hubs: Santa Clara, CA; Austin, TX; or Toronto, ON.

We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting.


Who You Are

  • You understand how accelerator compute, memory, and networking topology constrain AI workloads, and don't treat hardware as a black box.
  • You're an early adopter of AI for your work from coding to building agentic workflows that multiply your impact.
  • You work directly with customers to understand their challenges and provide effective solutions.
  • You are comfortable debugging across the full inference stack: from failing requests, through the serving layer, down to OOMs or kernel dispatch if need be.
  • You bring feedback in the form of pull requests, reproducible code, benchmarks, and telemetry data.

What We Need

  • Strong software engineering skills with 5+ years of relevant technical experience (e.g. Applied Engineer, Machine Learning Engineer, MLOps Engineer, Platform Engineer, Infrastructure Engineer, Site Reliability Engineer, Field Application Engineer).
  • Experience turning ambiguous customer requirements or issues into verifiable acceptance criteria.
  • Kubernetes and Helm experience at multi-node, HPC, or AI cluster scale.
  • Experience with observability and infrastructure automation, e.g. Prometheus, Grafana, OpenTelemetry.
  • Experience with LLM inference serving engines and technologies, e.g. vLLM, SGLang, Mooncake, NIM, Dynamo, LMCache.

What You Will Learn

  • Where co-design of AI hardware and software translates into unique latency and throughput performance.
  • How to scale disaggregated inference services on Kubernetes while balancing performance, reliability, and tactical tradeoffs.
  • What makes enterprise AI deployments successful: from technical requirements through software delivery, cluster-scale validation, and production ownership.
  • Why customer insights from the field shape the best products.
  • How to build agentic workflows for asymmetric impact.

Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made.

Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer.

This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology.  Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2).   These requirements apply to persons located in the U.S. and all countries outside the U.S.  As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency.  If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.

Similar Jobs

2 Months Ago
In-Office or Remote
USA
165K-350K Annually
Mid level
165K-350K Annually
Mid level
Artificial Intelligence • Legal Tech
Embed with customers to design, build, and deploy GC AI integrations into legal workflows. Develop production-grade API/webhook integrations, troubleshoot deployments, create reference implementations and documentation, and feed product insights back to Engineering to improve the platform. Travel to customer sites up to 25% as needed.
Top Skills: APIsLlmsPythonSdksTypescriptWebhooksWorkflow Automation
19 Days Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
169K-273K Annually
Senior level
169K-273K Annually
Senior level
Artificial Intelligence • Machine Learning • Retail • Social Impact • Software
Deploy and build Afresh's AI platform: embed with customers to integrate data and ship production LLM/agent systems, then harden field learnings into reusable platform components, tooling, and observability to improve quality and scale.
Top Skills: Agent FrameworksAgentsBigQueryDatabricksKnowledge GraphLanggraphLlmsMlopsModel ServingObservabilityOntologyPgvectorPineconeRetrieval/RagSemantic LayerSnowflakeVector StoresWeaviate
30 Minutes Ago
Hybrid
Senior level
Senior level
Machine Learning • Payments • Security • Software • Financial Services
Lead the target architecture and hands-on delivery of an AI-driven commercial lending modernization platform. Design cloud-native systems using agentic AI, LLMs, microservices, APIs, event-driven architecture, and Kubernetes. Guide legacy code analysis and rewriting, make real-time technical decisions, lead design sessions, establish scalable engineering patterns, and communicate architecture tradeoffs to technical and executive stakeholders. The role also addresses security, resiliency, auditability, model risk, and regulatory controls.
Top Skills: Agentic AiAi Agent OrchestrationAi Observability ToolsAPIsAutogenAWSAzureContainersEvent-Driven ArchitectureJavaKubernetesLanggraphLlmsMcpMicroservicesPythonRagSemantic KernelTogaf

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account