Metasys Logo

Metasys

MLOps Engineer Internship

Posted 6 Hours Ago
Be an Early Applicant
Remote
Hiring Remotely in United States
Internship
Remote
Hiring Remotely in United States
Internship
Build and maintain MLOps infrastructure: CI/CD for models, monitoring, deployment into a NestJS monolith, feature store/versioning, retraining workflows, A/B experimentation, and LLM operationalization.
The summary above was generated by AI
Overview: Operationalizing AI and Personalization

The MLOps Engineer is crucial for bridging the gap between data science and production, responsible for the reliable, scalable, and secure deployment of machine learning models. You will operationalize the models powering our AI agents and the e-commerce personalization systems, ensuring continuous integration, delivery, and monitoring of our predictive analytics and recommendation engines.

Internship Details

Duration: 3 months
Start Date: Immediate
Location: Remote
Stipend: None initially. Based on your first-quarter performance, you may be offered a paid full-time opportunity, or even be absorbed directly by the client as an FTE.

Key Responsibilities & Core Projects

You will build the automation infrastructure that turns static models into continuously improving production systems.

  • Model CI/CD Pipelines: Design and build robust CI/CD pipelines dedicated to the machine learning lifecycle: automated model training, validation, and deployment using tools integrated with our main Makefile CI/CD setup.

  • Model Monitoring & Tracking: Implement comprehensive monitoring and alerting for model performance (e.g., drift detection, prediction accuracy, latency) and track experiments and model artifacts using version control tools.

  • Production Deployment: Operationalize the deployment of ML models powering AI agents and e-commerce services, ensuring they integrate seamlessly into the NestJS modular monolith architecture.

  • Versioning & Feature Stores: Manage model versioning and lineage. Collaborate on the design and maintenance of a centralized Feature Store to ensure consistent data for training and serving.

  • Experimentation Infrastructure: Implement and manage the infrastructure necessary for A/B testing different model versions or personalization strategies in a production environment (e-commerce storefront).

  • Retraining Workflows: Define and automate the model retraining workflows based on data drift or performance degradation triggers, ensuring models remain relevant to the dynamic supply chain and customer behavior.

Required Technologies & Tools

Candidates must possess hands-on expertise in the tools and methodologies used for production ML and MLOps:

  • MLOps Tools: Experience with model registries, experiment tracking, and serving platforms (e.g., MLflow, Kubeflow, Sagemaker).

  • CI/CD & Automation: Proficiency in building pipelines (using Python/Bash scripting) and experience with Docker and Terraform.

  • Data & Compute: Experience managing data pipelines for ML (ETL/ELT) and optimizing compute resources for training and inference.

  • Programming: Strong proficiency in Python and familiarity with TypeScript/Node.js for deployment integration.

  • Methodology: Deep understanding of MLOps best practices, responsible AI principles, and monitoring concepts.

AI Agent Focus

You will ensure the reliability and continuous improvement of the core AI layer.

  • LLM Operationalization: Implement specific pipelines for the fine-tuning, validation, and deployment of Large Language Models (LLMs) used in our AI agents.

  • Agent Performance Tracking: Develop metrics and tracking systems to measure the business impact and operational efficiency of multi-agent systems and recommendation engines.

  • Framework Integration: Operationalize models built using frameworks like LangChain or LlamaIndex, ensuring they are secure, versioned, and scalable in a production environment.

Success Metrics & Career Path

Performance will be measured by:

  • Deployment Velocity: Speed and reliability of deploying new or retrained models to production.

  • Model Performance: Maintaining model accuracy and minimizing performance drift in production.

  • Pipeline Automation: Percentage of the ML lifecycle (training, validation, deployment) that is fully automated.

Mentorship Structure: Reports to the Solution Architect or Head of Technology, collaborating closely with Data Architects, Data Scientists, and SREs to maintain a reliable AI ecosystem.

Similar Jobs

24 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
173K-255K Annually
Senior level
173K-255K Annually
Senior level
Big Data • Fintech • Mobile • Payments • Financial Services
Lead product for Affirm's 1:Many Distribution Partnerships, owning partner portfolio, building repeatable partner acceleration capabilities, defining roadmaps, metrics, and go-to-market strategies, and coordinating cross-functional teams to launch and scale platform integrations that drive merchant growth.
Top Skills: Figma
2 Hours Ago
Remote
United States
110K-130K Annually
Mid level
110K-130K Annually
Mid level
AdTech • Artificial Intelligence • Big Data • Digital Media • eCommerce • Machine Learning • Marketing Tech
Plan and execute regional and large-scale industry events, design executive-level activations, manage partner marketing and budgets, track lead flow and ROI, collaborate with Sales/Product/Global Marketing, and standardize event and partner marketing processes.
Top Skills: AsanaCRMRoi DashboardsSpreadsheets
5 Hours Ago
Easy Apply
Remote or Hybrid
Easy Apply
166K-196K Annually
Senior level
166K-196K Annually
Senior level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Own and execute the multi-year telematics hardware platform roadmap end-to-end. Drive commercial outcomes for gateways, sensors, cabling and accessories through field validation, launch, lifecycle management, and cross-functional alignment. Engage customers and sales on strategic deals, use market and fleet data to prioritize, and mentor junior PMs while ensuring hardware strategy supports revenue, margin, and company growth.
Top Skills: Can BusCloudEdge AiFccIotJ1939Obd-IiPtcrbSensorsTelematicsVehicle GatewaysWireless

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account