Design, build, and maintain scalable batch and streaming data pipelines using Python, SQL, Kafka, and GCP. Develop REST and GraphQL data services, CI/CD automation, observability, and data-quality frameworks. Integrate ML and NLP models into production workflows, including continuous retraining and real-time inference. Support structured and unstructured healthcare datasets, distributed systems, microservices, and reliable data infrastructure while collaborating with data science and MLOps teams.
Responsibilities
- Design, build, and maintain scalable data pipelines to support analytics, ML, and operational reporting.
- Develop robust data ingestion, transformation, and integration workflows using Python, SQL, and modern data engineering frameworks.
- Build and maintain batch and streaming data pipelines leveraging technologies such as Kafka (or similar pub/sub tools).
- Work with Google Cloud Platform (GCP) services, including Cloud Storage, Dataflow, Pub/Sub, BigQuery, Cloud Spanner and Cloud Functions
- Develop and manage data APIs and interfaces (REST and GraphQL) to enable high-performance data access across microservices.
- Implement CI/CD automation for data pipelines using GitHub Actions, Argo CD, or equivalent tools.
- Collaborate with Data Scientists and MLOps teams to integrate ML/NLP models into data pipelines and production workflows.
- Build and operationalize NLP data pipelines for structured and unstructured data sources (e.g., Rx claims, clinical documents).
- Enable continuous learning and model‑retraining workflows using Vertex AI, Kubeflow, or similar GCP‑native tooling.
- Implement frameworks for observability and data quality, ensuring ML predictions, confidence scores, and fallback events are logged into data lakes or monitoring systems.
- Support distributed data systems and ensure reliability, performance, and scalability of data infrastructure.
Required Qualifications
- 5+ years of experience building data pipelines or backend data workflows using Python, Java, or similar languages.
- 2+ years of experience designing REST/GraphQL data services or integrating data APIs.
- Hands‑on experience working with ML/AI model integration in production (e.g., Vertex AI Endpoints, TensorFlow Serving, ML REST APIs).
- Experience handling structured and unstructured datasets, including healthcare data (Rx claims, clinical documents, NLP text).
- Familiarity with the end-to-end ML lifecycle: data ingestion, feature engineering, training, deployment, and real‑time inference.
- 2+ years of experience with cloud platforms (GCP preferred; AWS or Azure acceptable).
- 2+ years working with streaming platforms like Kafka or equivalent.
- 2+ years of experience with databases (Postgres or similar relational systems).
- 2+ years of experience with CI/CD tools (GitHub Actions, Jenkins, Argo CD, etc.).
Preferred Qualifications
- Direct, hands-on experience with Google Cloud Platform, especially BigQuery, Dataflow, GKE, Composer and Vertex AI.
- Knowledge of Kubernetes concepts and experience running data services or pipelines on GKE.
- Strong understanding of distributed systems, microservice patterns, and data‑centric system design.
- Experience using Vertex AI, Kubeflow, or other ML orchestration platforms for model training and serving.
- Knowledge of GenAI pipelines, LLM prompt workflows, and agent orchestration frameworks (e.g., LangChain, transformers).
- Experience deploying Python-based ML/NLP services into microservice ecosystems using REST, gRPC, or sidecar architectures.
- Domain experience in healthcare, claim adjudication, or Rx data processing.
Education
- Bachelor’s degree in Computer Science, Data Engineering, Information Systems, or equivalent experience
(High School Diploma + 4 years of relevant experience acceptable).
Compensation, Benefits and Duration
Minimum Compensation: USD 42,000
Maximum Compensation: USD 147,000
Compensation is based on actual experience and qualifications of the candidate. The above is a reasonable and a good faith estimate for the role.
Medical, vision, and dental benefits, 401k retirement plan, variable pay/incentives, paid time off, and paid holidays are available for full-time employees.
This position is available for independent contractors
No applications will be considered if received more than 120 days after the date of this post
Similar Jobs
Automotive • Big Data • Insurance • Software • Transportation
Handle high-volume inbound roadside assistance calls from stranded motorists, gather location and vehicle details, dispatch tow and service providers, de-escalate stressful situations, document interactions, and coordinate service using multiple digital systems. The role requires reliable full-time schedule adherence, weekend and holiday flexibility, strong communication, multitasking, accurate typing, and a dedicated home office with employer-specified equipment and hardwired internet.
Top Skills:
Ai ToolsCrm SystemsDispatch SoftwareEthernetGmailGoogle ChatGoogle ChromeGoogle DocsGoogle MapsGoogle SheetsGoogle WorkspaceHarver System CheckerMozilla FirefoxSwoopWindows 11Zoom
Automotive • Big Data • Insurance • Software • Transportation
Provide real-time inbound customer support for roadside emergencies, gather location and vehicle details, dispatch tow and service providers, de-escalate stressful situations, document calls, and track service progress using multiple digital tools. The role requires reliable full-time schedule adherence, weekend and holiday availability, remote collaboration, and a dedicated home workspace meeting strict technology and internet requirements.
Top Skills:
CRMDispatch SoftwareGmailGoogle ChatGoogle ChromeGoogle DocsGoogle MapsGoogle SheetsGoogle WorkspaceHarverMozilla FirefoxSwoopWindows 11Zoom
Automotive • Big Data • Insurance • Software • Transportation
Handle high-volume inbound roadside assistance calls from stranded motorists, gather location and vehicle details, dispatch tow and service providers, de-escalate stressful situations, document cases, and provide clear customer guidance. Associates use multiple web-based systems and dispatch tools while collaborating remotely. The role requires reliable schedule adherence, weekend and holiday availability, strong communication, empathy, multitasking, accurate typing, and the ability to make sound decisions under pressure.
Top Skills:
Ai ToolsCrm SystemsDispatch SoftwareGoogle ChatGoogle ChromeGoogle MapsGoogle WorkspaceHarverMozilla FirefoxWindows 11Zoom
What you need to know about the Austin Tech Scene
Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.
Key Facts About Austin Tech
- Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
- Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
- Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
- Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
- Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

