VAST Data Logo

VAST Data

Senior Solutions Engineer, AI Infrastructure

Reposted 19 Days Ago
Remote or Hybrid
Hiring Remotely in United States
Senior level
Remote or Hybrid
Hiring Remotely in United States
Senior level
The Senior Solutions Engineer will design and implement infrastructure for AI and HPC workloads, engage with customers, and lead technical discovery and architecture design.
The summary above was generated by AI
Description

We're looking for a deeply technical Solutions Architect to help customers design, evaluate, and deploy infrastructure for large-scale AI, HPC, analytics, and data-intensive workloads.

This is a customer-facing technical role for someone who has lived inside production infrastructure. You may have been a platform engineer, infrastructure engineer, SRE, MLOps engineer, AI infrastructure engineer, storage engineer, cloud engineer, or HPC systems engineer. What matters most is that you have built, operated, or architected real systems, and can bring that credibility into customer conversations.

Our customers are building infrastructure at serious scale: GPU clusters, high-performance storage systems, Kubernetes platforms, distributed training environments, inference platforms, data pipelines, lakehouses, and large enterprise systems. You'll help them reason about architectures involving 10,000+ GPUs, 100PB+ of storage, high-performance networking, distributed filesystems, orchestration layers, and demanding production workloads.

You'll own technical discovery, architecture design, PoC planning, competitive positioning, and customer technical strategy. You'll work from the first whiteboard session through evaluation, deployment planning, and production success. You'll also partner closely with product and engineering teams to bring field feedback into the roadmap.

We're looking for someone who can go deep technically, communicate clearly, operate without a rigid playbook, and translate complex infrastructure into customer outcomes.

Responsibilities

  • Lead technical discovery with customers across infrastructure, platform, ML, data, and executive stakeholders.
  • Design architectures for large-scale AI, HPC, analytics, and enterprise data workloads.
  • Help customers evaluate infrastructure involving GPUs, storage, networking, orchestration, and data movement.
  • Translate complex technical requirements into clear solution designs, reference architectures, and deployment guidance.
  • Debug customer issues across Linux, storage, networking, Kubernetes, schedulers, GPUs, and application workloads.
  • Build technical assets, runbooks, and field guidance for repeatable customer engagements.
  • Partner with product and engineering to communicate customer requirements, gaps, and roadmap opportunities.
  • Help customers move from architecture design to production deployment.
Requirements
  • 8 to 12+ years of technical experience, with significant hands-on infrastructure experience.
  • Experience building, operating, or architecting production platform infrastructure.
  • Strong understanding of Linux kernel implementation details, distributed systems including PAXOS and raft, storage implementations details like NAND or write amplification, networking store/forward, load balancing designs, and production operations.
  • Experience with one or more of: GPU infrastructure, large scale HPC systems, Kubernetes platforms from scratch, MLOps, storage systems, cloud infrastructure, data platforms, or large-scale enterprise infrastructure.
  • Ability to communicate credibly with engineers, architects, technical executives, and business stakeholders.
  • Strong discovery, problem-solving, and systems debugging skills.
  • Comfort operating in ambiguous, fast-moving environments.
  • Interest in customer-facing technical work, solution design, and business outcomes.

Preferred Experience

  • Experience with large-scale GPU clusters, distributed training, inference infrastructure, or AI platforms.
  • Experience with petabyte-scale storage or high-performance data systems.
  • Experience with Kubernetes, Slurm, Ray, Spark, or other orchestration / scheduling systems.
  • Domain Expertise with one or more of these - Lustre, Ceph, Weka, BeeGFS, GPFS, VAST, object storage, or distributed filesystems.
  • Experience with large-scale InfiniBand, RoCE, RDMA, high-performance Ethernet, or NVIDIA/Mellanox networking.
  • Direct Experience with CUDA, NCCL, DCGM, GPUDirect, checkpointing, dataset staging, or model-serving infrastructure.
  • Experience across multiple industries or customer environments.

Similar Jobs

Senior level
Artificial Intelligence • Big Data • Healthtech • Software • Biotech
Leads end-to-end IVD, companion diagnostic, clinical trial assay, tumor profiling, and SaMD development programs. Aligns internal cross-functional teams and biopharma partners across strategy, validation, clinical operations, regulatory submission, manufacturing, and commercialization. Owns governance, milestones, budgets, resources, risks, and partner relationships while delivering programs through regulatory approval or pivotal clinical trials.
Top Skills: Analytical ValidationAssay DevelopmentBioinformaticsBlaClinical ValidationCompanion Diagnostics (Cdx)Design ControlsEu IvdrFda Regulatory RequirementsIndIvdMaaNdaPhase I/Ii/IiiSoftware As A Medical Device (Samd)
2 Hours Ago
Remote or Hybrid
USA
170K-260K Annually
Expert/Leader
170K-260K Annually
Expert/Leader
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Lead product marketing strategy and go-to-market execution for Falcon Exposure Management. Develop positioning, messaging, launches, thought leadership, web strategy, and field enablement while partnering with product, sales, marketing, and executive teams. Translate cybersecurity capabilities into customer value and business outcomes, guide market education, analyze buyers and competitors, and strengthen CrowdStrike’s category leadership. The role requires deep exposure management expertise, strong cybersecurity knowledge, exceptional communication skills, and 12–15+ years of relevant experience.
Top Skills: AIB2B SaasCloud SecurityCybersecurityEnterprise SoftwareExternal Attack Surface ManagementFalcon Exposure ManagementIdentity SecuritySecurity OperationsVulnerability Management
2 Hours Ago
Remote or Hybrid
USA
120K-180K Annually
Senior level
120K-180K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Lead program management for CrowdStrike’s Office of Adversary Strategy, coordinating complex cross-functional projects from scoping through delivery. Responsibilities include managing schedules, dependencies, risks, metrics, communications, change management, process improvements, and escalations. The role partners with technical teams and leadership, applies Agile project management practices, coaches teams through obstacles, and uses AI technologies to improve decision-making, workflows, and business outcomes.
Top Skills: AgileArtificial IntelligenceChatgptClaudeExcelGeminiGenerative AiJIRAKanbanScaled Agile FrameworkScrumSystems Development Lifecycle

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account