Urban SDK Logo

Urban SDK

Data Engineer

Reposted 10 Hours Ago
Remote
Hiring Remotely in United States
Mid level
Remote
Hiring Remotely in United States
Mid level
Design, build, and maintain scalable ETL/ELT data pipelines on Databricks and AWS S3; manage large geospatial and temporal datasets; productionize ML models; implement data validation, testing, and monitoring; optimize storage and processing; document workflows; collaborate cross-functionally.
The summary above was generated by AI

About Urban SDK

Urban SDK is shaping the Future of Smart Cities. We are pioneers in geospatial AI technology, providing public leaders with insights and automation for mission-critical decisions. We equip critical public services with geospatial AI, enabling precise, data-driven decisions with efficiency and confidence.

Our Commitment to People

We are committed to aligning business growth with professional outcomes for every employee. Our commitment has been recognized by Jacksonville Business Journal and Will Reed as a noted Best Places to Work.


About the role

We are looking for a skilled Data Engineer to design, build, and maintain scalable data pipelines and platforms that support our geospatial traffic analytics applications. The ideal candidate will have experience with Python, Databricks, S3, and modern data engineering practices, including automated testing, CI/CD, and data quality monitoring.

Responsibilities

  • Design, implement, and maintain scalable data pipelines and ETL/ELT workflows on Databricks and cloud platforms.
  • Manage large-scale geospatial and temporal datasets stored in AWS S3.
  • Collaborate with data scientists to productionize machine learning models and ensure smooth data availability.
  • Implement data validation, testing, and monitoring frameworks to ensure data accuracy, consistency, and reliability.
  • Optimize data storage and processing strategies to handle high volumes of traffic and mobility data efficiently.
  • Develop and maintain documentation for data workflows, architecture, and processes.
  • Work closely with cross-functional teams to understand data requirements and ensure timely delivery.
  • Stay up-to-date with the latest trends and best practices in data engineering, cloud technologies, and big data processing.

Qualifications

  • Bachelor's or Master's degree in Computer Science, Data Engineering, or a related field.
  • 3+ years of experience as a data engineer or in a similar role.
  • Strong proficiency in Python and associated libraries for data engineering (pandas, PySpark, etc.).
  • Hands-on experience with Databricks and Spark for large-scale data processing.
  • Experience with AWS services, especially S3, and knowledge of cloud-based data architectures.
  • Solid understanding of data pipeline testing, version control, and CI/CD practices.
  • Experience with SQL and NoSQL databases.
  • Strong problem-solving skills and attention to detail.


Preferred Skills

  • Familiarity with geospatial data formats and processing (GeoJSON, Shapefiles, PostGIS).
  • Experience with workflow orchestration tools (Databricks, Prefect, or similar).
  • Knowledge of containerization (Docker/Kubernetes) and cloud-native data solutions.
  • Experience supporting machine learning pipelines in production.


Compensation

  • Location: Jacksonville, FL (Town Center Area) or Remote
  • Type:  Full-time
  • Reports to: Director of Engineering
  • Salary Based on Experience 
  • Annual Bonus
  • Medical, Vision, Dental, 401(k)  
  • 21 Days Vacation
  • Office Lunch provided Daily

Similar Jobs

2 Days Ago
Remote or Hybrid
113K-193K Annually
Senior level
113K-193K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design and operate scalable batch and streaming data platforms supporting machine learning and generative AI. Build pipelines for structured, unstructured, OCR, document, and image data; develop RAG, semantic search, and LLM-powered solutions; and establish data quality, observability, governance, orchestration, and deployment practices. Partner with stakeholders, lead platform scalability and cost optimization, mentor engineers, and translate ambiguous needs into production-ready technical roadmaps while securely handling sensitive data.
Top Skills: AirflowAmazon KinesisSparkAWSAzureAzure Event HubsChart.JsDatabricksDeequDelta LakeDockerGithub ActionsGCPGreat ExpectationsJavaKafkaKubernetesLlmsMlopsPlotlyPysparkPythonRagScalaSeabornSnowflakeSQLTerraform
3 Days Ago
Remote
United States
Mid level
Mid level
Artificial Intelligence • Information Technology • Professional Services • Software • Analytics • Generative AI • Big Data Analytics
Configure and maintain Adobe Experience Platform and Real-Time CDP data pipelines, XDM schemas, datasets, identity resolution, Profile enablement, segmentation readiness, and destination activation. Validate data quality, troubleshoot ingestion and identity issues, manage sandboxes, and implement privacy, consent, governance, and access controls. Partner with architects, data engineers, consultants, analysts, data scientists, and marketing teams to support reporting, personalization, and machine learning readiness.
Top Skills: Adobe AnalyticsAdobe Experience PlatformAdobe I/O RuntimeAdobe Journey OptimizerAdobe Real-Time CdpAdobe Source ConnectorsAdobe TargetAPIsAWSAzureBatch IngestionBigQueryCcpaData PrepEltETLGCPGdprJavaScriptPythonQuery ServiceRedshiftSalesforce CdpSegmentSnowflakeSQLStreaming IngestionWeb SdkXdm
Yesterday
In-Office or Remote
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Lead cloud data modernization by designing and implementing Azure/Snowflake/Databricks data platforms, building scalable ETL/ELT pipelines, ensuring data quality, security and governance, implementing CI/CD, mentoring engineers, and supporting healthcare data solutions and compliance.
Top Skills: AirflowAzureAzure Data Factory (Adf)Azure Data Lake Storage Gen2 (Adls Gen2)Change Data Capture (Cdc)CptDatabricksDatabricks GenieEtl/EltFacetsFhirGitGithub ActionsGithub CopilotHcpcsHl7Icd-10LlmsLoincPrompt EngineeringPysparkPythonRag PipelinesSnowflakeSnowflake CortexSparkSQL ServerSsisVector StoresVisioX12 Edi (837/835/834)

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account