Metasys Logo

Metasys

MLOps Engineer Internship

Reposted Yesterday
Remote
Hiring Remotely in United States
Internship
Remote
Hiring Remotely in United States
Internship
Build and maintain MLOps infrastructure: CI/CD for models, monitoring, deployment into a NestJS monolith, feature store/versioning, retraining workflows, A/B experimentation, and LLM operationalization.
The summary above was generated by AI
Overview: Operationalizing AI and Personalization

The MLOps Engineer is crucial for bridging the gap between data science and production, responsible for the reliable, scalable, and secure deployment of machine learning models. You will operationalize the models powering our AI agents and the e-commerce personalization systems, ensuring continuous integration, delivery, and monitoring of our predictive analytics and recommendation engines.

Internship Details

Duration: 3 months
Start Date: Immediate
Location: Remote
Stipend: None initially. Based on your first-quarter performance, you may be offered a paid full-time opportunity, or even be absorbed directly by the client as an FTE.

Key Responsibilities & Core Projects

You will build the automation infrastructure that turns static models into continuously improving production systems.

  • Model CI/CD Pipelines: Design and build robust CI/CD pipelines dedicated to the machine learning lifecycle: automated model training, validation, and deployment using tools integrated with our main Makefile CI/CD setup.

  • Model Monitoring & Tracking: Implement comprehensive monitoring and alerting for model performance (e.g., drift detection, prediction accuracy, latency) and track experiments and model artifacts using version control tools.

  • Production Deployment: Operationalize the deployment of ML models powering AI agents and e-commerce services, ensuring they integrate seamlessly into the NestJS modular monolith architecture.

  • Versioning & Feature Stores: Manage model versioning and lineage. Collaborate on the design and maintenance of a centralized Feature Store to ensure consistent data for training and serving.

  • Experimentation Infrastructure: Implement and manage the infrastructure necessary for A/B testing different model versions or personalization strategies in a production environment (e-commerce storefront).

  • Retraining Workflows: Define and automate the model retraining workflows based on data drift or performance degradation triggers, ensuring models remain relevant to the dynamic supply chain and customer behavior.

Required Technologies & Tools

Candidates must possess hands-on expertise in the tools and methodologies used for production ML and MLOps:

  • MLOps Tools: Experience with model registries, experiment tracking, and serving platforms (e.g., MLflow, Kubeflow, Sagemaker).

  • CI/CD & Automation: Proficiency in building pipelines (using Python/Bash scripting) and experience with Docker and Terraform.

  • Data & Compute: Experience managing data pipelines for ML (ETL/ELT) and optimizing compute resources for training and inference.

  • Programming: Strong proficiency in Python and familiarity with TypeScript/Node.js for deployment integration.

  • Methodology: Deep understanding of MLOps best practices, responsible AI principles, and monitoring concepts.

AI Agent Focus

You will ensure the reliability and continuous improvement of the core AI layer.

  • LLM Operationalization: Implement specific pipelines for the fine-tuning, validation, and deployment of Large Language Models (LLMs) used in our AI agents.

  • Agent Performance Tracking: Develop metrics and tracking systems to measure the business impact and operational efficiency of multi-agent systems and recommendation engines.

  • Framework Integration: Operationalize models built using frameworks like LangChain or LlamaIndex, ensuring they are secure, versioned, and scalable in a production environment.

Success Metrics & Career Path

Performance will be measured by:

  • Deployment Velocity: Speed and reliability of deploying new or retrained models to production.

  • Model Performance: Maintaining model accuracy and minimizing performance drift in production.

  • Pipeline Automation: Percentage of the ML lifecycle (training, validation, deployment) that is fully automated.

Mentorship Structure: Reports to the Solution Architect or Head of Technology, collaborating closely with Data Architects, Data Scientists, and SREs to maintain a reliable AI ecosystem.

Similar Jobs

25 Seconds Ago
In-Office or Remote
129K-233K Annually
Mid level
129K-233K Annually
Mid level
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Own the Pittsburgh territory through field-based, full-cycle sales. Generate pipeline through prospecting, networking, events, partnerships, and door-to-door outreach; conduct demos; close deals; manage Salesforce activity and forecasts; and exceed quotas. Build trusted relationships with local businesses across restaurants, retail, and services while coordinating onboarding and referrals. The role requires approximately 80% field work and 50–60 targeted business visits weekly.
Top Skills: Payment ProcessingSalesforceSquare Software And Hardware
29 Seconds Ago
In-Office or Remote
130K-234K Annually
Mid level
130K-234K Annually
Mid level
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Own the full outbound sales cycle for mid-market merchants, from prospecting and pipeline development through discovery, demonstrations, negotiation, and close. Build multi-product solutions, acquire new logos, manage complex multi-stakeholder deals, maintain Salesforce forecasting accuracy, collaborate cross-functionally, and consistently achieve revenue targets.
Top Skills: Salesforce
8 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
204K-290K Annually
Senior level
204K-290K Annually
Senior level
Big Data • Fintech • Mobile • Payments • Financial Services
Lead technical strategy and execution for foundational Trust and Safety systems supporting consumer risk and credit reporting. Design and operate highly available backend systems, improve reliability and monitoring, establish engineering standards, guide complex cross-functional initiatives, and mentor engineers. The role involves ownership of system architecture, technical planning, operational readiness, code quality, and team development.
Top Skills: SparkAWSKotlinKubernetesMySQLPython

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account