Sentient Foundation Logo

Sentient Foundation

Applied ML Engineer

Posted An Hour Ago
In-Office or Remote
Hiring Remotely in Austin, TX, USA
Entry level
In-Office or Remote
Hiring Remotely in Austin, TX, USA
Entry level
Build end-to-end machine learning systems spanning research reproduction, model evaluation, inference infrastructure, backend services, and product interfaces. Design rigorous experiments, evaluation datasets, and verification workflows; analyze model internals; investigate provenance and evasion; and convert findings into reliable production systems with APIs, background jobs, observability, testing, documentation, and user-facing React/TypeScript experiences.
The summary above was generated by AI
Applied ML EngineerThe Role

We’re looking for an Applied ML Engineer to build systems at the intersection of machine learning research and production software.

This is an end-to-end engineering role. You should be comfortable reading a research paper, identifying what is actually testable, building the smallest useful experiment, evaluating it rigorously, and turning the result into a production system that users can interact with.

You’ll work across model evaluation, model internals, inference infrastructure, backend systems, and product interfaces. The goal is not simply to reproduce research. It is to turn promising methods into reliable, measurable, and usable products.

What You’ll Do
  • Reproduce and evaluate research methods using open-weight and API-accessible models.

  • Design evaluation datasets, probes, scoring methods, baselines, calibration tests, and experiment harnesses.

  • Work directly with model weights, logits, hidden states, activations, model APIs, and inference infrastructure when required.

  • Build and extend our evaluation infrastructure, including runners, judges, persistence, experiment orchestration, and reporting.

  • Turn research workflows into product experiences, including experiment configuration, runs, traces, comparisons, reports, and review workflows.

  • Investigate how verification methods behave under model modification, including fine-tuning, merging, quantization, distillation, safety removal, and deliberate evasion.

  • Design controlled experiments that separate meaningful signals from artifacts or confounders.

  • Write clear technical reports that distinguish measured evidence, interpretation, and hypotheses.

  • Ship production-quality systems with APIs, background jobs, observability, testing, and documentation.

What We’re Looking For
  • Strong Python engineering skills and hands-on experience with PyTorch and Hugging Face Transformers.

  • A strong understanding of ML evaluation, including dataset design, baselines, metrics, calibration, false positives, false negatives, statistical uncertainty, and reproducibility.

  • Ability to read ML research papers and implement methods from first principles rather than relying entirely on existing packages.

  • Experience building production software beyond notebooks, including APIs, asynchronous jobs, databases, logging, testing, and deployment.

  • Comfort working with open-weight models and understanding how modern LLM inference systems operate.

  • Ability to work across backend and frontend boundaries. Our product surface is primarily React/TypeScript, and you should be able to make complex experiments and results understandable to users.

  • Strong technical judgment about what experimental evidence does and does not support. For example, evidence that one model was derived from another is not necessarily evidence that it was directly trained on that model's outputs.

  • High agency and a strong sense of ownership. You are comfortable identifying problems, proposing solutions, and driving work forward without waiting for detailed instructions.

  • Comfortable working in a fast-moving startup environment where priorities can evolve quickly and individuals are expected to operate across functions.

Useful Experience

Experience in any of the following is a plus:

  • Model provenance, fingerprinting, watermarking, distillation detection, red-teaming, safety evaluations, or interpretability.

  • Activation and representation analysis, probing, model hooks, logits, hidden states, or other model-internals work.

  • Evaluation and inference infrastructure such as DSPy, LiteLLM, Temporal, Ray, vLLM, PostgreSQL/pgvector, or similar systems.

  • Next.js, React, TypeScript, data visualization, or experiment dashboards.

  • Running and serving open-weight models on GPUs and reasoning about latency, throughput, memory, precision, and cost tradeoffs.

  • Designing adversarial evaluations or testing systems against deliberate attempts to evade detection.

What Success Looks Like in the First Six Months

You will:

  • Reproduce at least one published model-provenance or verification method and clearly document its capabilities, assumptions, and limitations.

  • Build a repeatable model-verification runner with versioned inputs, artifacts, metrics, and reports.

  • Add at least one verification workflow to Construct and make it accessible through the Eldros UI.

  • Run controlled experiments across base models, fine-tuned models, merged models, quantized models, and known distilled models.

  • Improve our ability to understand when verification methods succeed, when they fail, and why.

  • Leave behind production-quality code, tests, tooling, and documentation that another engineer can confidently operate and extend.

This Role Is Not
  • A pure research role where work ends with a paper or notebook.

  • A generic model-training or fine-tuning position.

  • A frontend-only or backend-only engineering role.

  • A role where benchmark scores are accepted at face value without understanding how they were produced.

  • A role for someone who wants to stay within a single layer of the stack.

We are looking for someone who enjoys moving between research, experimentation, engineering, and product, and who cares about building systems that produce evidence people can actually trust.

Similar Jobs

One Month Ago
In-Office or Remote
Senior level
Senior level
Artificial Intelligence • Information Technology • Consulting
Design, train, and deploy retrieval, reranking, and relevance ML models for a production agent-native search platform. Build embedding-based indexing and large-scale retrieval systems, define evaluation metrics and pipelines, optimize latency/quality/cost trade-offs, and collaborate with engineers to integrate models into 24x7 production services.
Top Skills: C++GoPython
Yesterday
Remote or Hybrid
Site of Old Bullion, NV, USA
112K-207K Annually
Senior level
112K-207K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads Pfizer’s environmental sustainability strategy across climate, Scope 3 emissions, biodiversity, nature, product sustainability, and sustainable sourcing. Manages supplier decarbonization programs, sustainability metrics, regulatory horizon scanning, risk assessments, governance, training, and communications. Facilitates cross-functional teams, advises stakeholders, represents Pfizer in external initiatives, and serves as a sustainability subject matter expert across the enterprise.
Top Skills: Microsoft TeamsWebex
5 Days Ago
In-Office or Remote
USA
215K-358K Annually
Senior level
215K-358K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads measurement and product marketing for eight enterprise AI platforms. Builds shared KPI, taxonomy, scorecard, data-quality, and value-realization frameworks; translates usage and outcome data into executive insights and investment guidance. Oversees positioning, internal launches, campaigns, enablement, adoption, and audience segmentation. Partners across product, engineering, data, communications, and business teams while building and managing teams responsible for analytics, marketing, and communications.
Top Skills: Ai PlatformsBusiness IntelligenceDashboardsData ContractsData FabricData VisualizationEvent TaxonomyExperimentationKnowledge GraphsKpi FrameworksOkr PlatformsProduct AnalyticsTelemetry

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account