BJAK Logo

BJAK

Staff Machine Learning Engineer

Reposted 4 Days Ago
Remote
Hiring Remotely in United States
Senior level
Remote
Hiring Remotely in United States
Senior level
Lead production-grade ML systems: build data pipelines, training and evaluation workflows, and scalable GPU inference. Fine-tune and adapt large models (LoRA, QLoRA, SFT, DPO, distillation), optimize deployments for latency/cost/reliability, and collaborate with product and engineering to ship robust, observable model-driven features.
The summary above was generated by AI
About the Role

There are over 5 billion users using basic applications today such email, notes, tasks, calendar and they're not AI-native. Our mission is to build proactive applications for anyone in the world, who are not used to complex prompting. We aim to bring intelligence to conversations, errands, organising and workflows, with minimal to no prompting.

Our product focuses on achieving high reliability for long-running workflows, persistent context, and real-world task completion. We believe products will greatly reduce hallucinations

Our objective is to organise anyone's life, allowing us all to spend time on valuable and meaningful things

As Staff Machine Learning Engineer, you own the execution layer of our intelligence, turning research and model capabilities into reliable, scalable production systems.

You will work across the model lifecycle: data, training, evaluation, inference, and deployment. This is a hands-on leadership role for someone who wants to operate at the intersection of research, systems, and product.

 
What You'll Own
  • Own the end-to-end ML systems powering our company, from data and training to evaluation, inference, and deployment.

  • Build and evolve training and fine-tuning pipelines for large models.

  • Design evaluation systems that measure capability, robustness, safety, and real-world product performance.

  • Architect high-performance inference systems, optimizing latency, GPU utilization, memory, cost, and reliability.

  • Build data pipelines and systems for high-quality real-world and synthetic training data.

  • Establish reliable production infrastructure for deploying, monitoring, and continuously improving models.

  • Partner closely with research and application engineering to turn model capabilities into product improvements.

  • Make pragmatic technical trade-offs and rapidly iterate based on real-world performance.

 
What We're Looking For
  • Experience building and shipping ML systems used in production, not just research prototypes.

  • Strong understanding of modern large-model training, fine-tuning, evaluation, and inference.

  • Strong software engineering and systems fundamentals.

  • Experience operating ML workloads at meaningful scale, particularly GPU-based systems.

  • Strong technical judgment and the ability to navigate ambiguous problems independently.

  • A bias toward experimentation, measurement, and shipping.

  • High standards for correctness, reliability, and production quality.

 
Outcomes
  • Research and models reliably translate into production-ready solutions with clear performance and quality targets.

  • ML pipelines, training loops, and inference systems are stable, efficient, and maintainable.

  • Production issues are detected, debugged, and resolved quickly, minimizing user impact.

  • Team members are supported, aligned, and able to deliver high-impact ML work with minimal friction.

  • Iterations on models and systems are measurable, safe, and improve user experience over time.

 
Tech Stack
  • Python

  • PyTorch / JAX

  • GPU-based training and inference system

 
Ideal Experience
  • You have built or shipped real ML systems used by people, not just demos.

  • You are comfortable working with large models and understanding their failure modes.

  • You write strong, production-grade code and care about system correctness.

 
How We Work

We are a small, high-talent-density, hands-on team. Engineers have broad ownership and are expected to exercise strong judgment and execute independently.

We make decisions quickly, work closely together, and balance speed with engineering fundamentals. We care less about process and more about building something exceptional.

 
Interview process

If there appears to be a fit, we'll reach to schedule 3, but no more than 4 interviews.

Applications are evaluated by our technical team members. Interviews will be conducted via virtual meetings and/or onsite.

We value transparency and efficiency, so expect a prompt decision. If you've demonstrated the exceptional skills and mindset we're looking for, we'll extend an offer to join us. This isn't just a job offer; it's an invitation to be part of a team that's bringing AI to have practical benefits to billions globally.

Similar Jobs

7 Hours Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
234K-349K Annually
Expert/Leader
234K-349K Annually
Expert/Leader
eCommerce • Healthtech • Kids + Family • Retail • Social Media
Own the strategy, architecture, development, deployment, monitoring, and continuous improvement of production personalization and recommendation systems. Build custom embeddings, ranking models, and shared representations across feeds, search, and recommendations. Lead ambiguous problems from concept to measurable customer impact, establish AI evaluation standards, partner with product and data teams, make cross-team technical decisions, and coach senior engineers.
Top Skills: AirflowAWSAws SagemakerBedrock AgentcoreClaude CodeCoderabbitDatadogDbtDeep LearningDevinGithub ActionsHexIncident.IoIterableKotlinKubernetesLangchainLaunchdarklyLinearMatrix FactorizationMaximMlflowMySQLPackwerkPandasPythonPyTorchRuby on RailsRankingReactRetrievalScikit-LearnShopifySidekiqSigmaSnowflakeSwiftTerraformTypescriptWardenWarpstreamWeaviateXgboost
2 Days Ago
Easy Apply
Remote or Hybrid
Easy Apply
158K-225K Annually
Senior level
158K-225K Annually
Senior level
Cloud • Information Technology • Security • Software • Cybersecurity
Build and operate production LLM-powered agents and machine learning systems for customer risk identification and threat detection. Translate security heuristics into agent logic, collaborate with threat researchers, develop data requirements and pipelines, evaluate precision and recall, and deploy solutions through CI/CD. The role also requires production reliability ownership, observability, incident debugging, cloud infrastructure expertise, and senior staff-level technical leadership.
Top Skills: AthenaAWSCi/CdDockerElasticsearchLangchainLanggraphLlm AgentsNumpyOpensearchPandasPolarsPrestoPythonSQL
9 Days Ago
In-Office or Remote
277K-415K Annually
Expert/Leader
277K-415K Annually
Expert/Leader
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Lead end-to-end machine learning initiatives for Block’s conversational support systems. Responsibilities include developing and maintaining chatbot and recommendation models, researching LLM, RAG, fine-tuning, and real-time inference architectures, shaping long-term ML roadmaps, and guiding cross-functional teams. The role also requires communicating technical strategy and outcomes to leadership and external stakeholders while delivering scalable, production-ready AI solutions across Cash App, Square, and other Block products.
Top Skills: Deep LearningJaxLarge Language Models (Llms)Natural Language Processing (Nlp)PythonPyTorchReal-Time InferenceRetrieval-Augmented Generation (Rag)TensorFlow

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account