Raydar Logo

Raydar

Forward Deployed Machine Learning Engineer

Posted Yesterday
Remote
Hiring Remotely in United States
170K-270K Annually
Mid level
Remote
Hiring Remotely in United States
170K-270K Annually
Mid level
Build production machine learning benchmarks and evaluation systems for foundation models. Own backend infrastructure including data pipelines, execution environments, storage, and orchestration, as well as sandboxed environments for agentic evaluations. Partner with researchers, enterprise customers, and company leadership to scope and deliver technical solutions. Identify scalable evaluation patterns, communicate clearly, and operate effectively in an ambiguous, high-ownership startup environment.
The summary above was generated by AI

About the company

Our client is a fast-growing AI data infrastructure company building a secure marketplace and data lab for high-quality model-training data. The company helps data holders license sensitive, real-world datasets to vetted AI teams while protecting governance, privacy, intellectual property, and security. It has raised $65 million, including a $55 million Series A backed by leading venture firms, employs approximately 80 people, and is scaling quickly after reaching its full-year growth target early.

The role and why it matters

This is the first Machine Learning Engineer dedicated to the company's Benchmarks and Evaluations vertical. You will partner directly with the general manager, researchers, and early enterprise customers to establish the technical foundation for evaluating foundation models across domains and modalities. The role combines hands-on ML evaluation, backend infrastructure, and customer-facing delivery in a high-ownership environment.

What you'll do

• Define, design, and build production benchmarks and evaluations with customers and internal researchers.

• Build and own backend infrastructure, including data pipelines, execution environments, storage, and orchestration.

• Create sandboxed environments for agentic evaluations involving tools, code execution, and multi-step tasks.

• Own the engineering portion of customer engagements from technical scoping through production delivery.

• Identify repeatable evaluation patterns and infrastructure gaps that can become scalable products.

• Move quickly through ambiguity while maintaining strong technical judgment and clear written communication.


Requirements

What we're looking for

• 4 or more years of engineering experience, including hands-on machine learning model evaluation work.

• Experience deploying end-to-end ML evaluation or benchmark systems to production against demanding customer timelines.

• Strong proficiency with ML evaluation frameworks and benchmark design, including approaches such as LLM-as-judge.

• Ownership of backend and infrastructure systems such as large-scale data pipelines, execution environments, storage, and orchestration.

• Customer-facing engineering experience managing enterprise stakeholders.

• A degree in computer science, physics, or a related technical field.

• Current unrestricted U.S. work authorization without visa sponsorship.

Bonus points

• Experience building evaluations or human-data pipelines for large language models.

• Experience working directly with AI researchers or foundation-model labs.

• Published work or meaningful open-source contributions in ML evaluations or benchmarks.

• Early-stage B2B startup, forward-deployed engineering, or high-ownership generalist experience.


Benefits

Compensation and benefits

• $170K-$270K base salary.

• Competitive equity.

• Comprehensive benefits provided by a well-capitalized, high-growth company.

• High autonomy and the opportunity to build a new technical vertical from the ground up.

Location and work model

• Full-time and remote within the United States.

• Strong independent ownership, customer responsiveness, and cross-functional collaboration are expected.

Similar Jobs

Yesterday
Remote
3 Locations
Senior level
Senior level
Information Technology • Business Intelligence • Consulting
Lead enterprise delivery of Databricks-based data, machine learning, and Generative AI solutions. Design production data pipelines, analytics engines, RAG applications, and ML platforms; migrate legacy data estates to the Lakehouse; and advise executive client stakeholders. Drive platform adoption, identify expansion opportunities, mentor technical teams, and develop reusable engineering accelerators. The role combines hands-on architecture, engineering, consulting, customer engagement, and technical practice building.
Top Skills: Amazon RedshiftSparkAWSAws GlueAws LambdaAzureDatabricksDatabricks JobsDatabricks WorkflowsDelta LakeGCPGenerative AiHadoopLlmsMlflowMosaic AiPysparkPythonRagS3SnowflakeSQL ServerTeradataUnity CatalogVector Databases
9 Days Ago
Remote
United States
175K-200K Annually
Senior level
175K-200K Annually
Senior level
Other
Build and deploy production generative AI and machine learning systems for utility clients. Own engagements end to end, including discovery, architecture, retrieval, model selection, evaluation, deployment, monitoring, operator-facing applications, and post-go-live support. Integrate with client data and business systems, establish AI safety and governance controls, manage scope and risk, and communicate with technical teams and executives. Contribute reusable accelerators, evaluation frameworks, and reference architectures across client engagements.
Top Skills: SparkAWSAzureCi/CdDashDatabricksDockerGenerative AiGitGoogle Cloud PlatformGradioGraphql ApisHybrid RetrievalJavaKnowledge GraphsLarge Language ModelsMachine LearningPythonReactRest ApisRetrieval-Augmented GenerationScalaStreamlitTypescriptVector Databases
28 Days Ago
Remote
United States of America
Entry level
Entry level
Information Technology • Software • Consulting
Build and deploy AI-powered applications for customers, combining Python and TypeScript application engineering with LLMs, generative AI, agent orchestration, APIs, distributed systems, and cloud platforms. Partner with stakeholders to understand requirements, design scalable architectures, rapidly prototype solutions, integrate AI capabilities, and evolve proofs of concept into production systems. The role also involves customer-facing consulting, technical decision-making, and collaboration throughout the solution lifecycle.
Top Skills: Agent OrchestrationAi ObservabilityAi/MlAPIsAWSAzureDistributed SystemsGCPGenerative AiLarge Language Models (Llms)MlopsModel DeploymentPythonRagTypescriptVector Databases

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account