Urban SDK Logo

Urban SDK

Data Engineer

Reposted Yesterday
Remote
Hiring Remotely in United States
Mid level
Remote
Hiring Remotely in United States
Mid level
Design, build, and maintain scalable ETL/ELT data pipelines on Databricks and AWS S3; manage large geospatial and temporal datasets; productionize ML models; implement data validation, testing, and monitoring; optimize storage and processing; document workflows; collaborate cross-functionally.
The summary above was generated by AI

About Urban SDK

Urban SDK is shaping the Future of Smart Cities. We are pioneers in geospatial AI technology, providing public leaders with insights and automation for mission-critical decisions. We equip critical public services with geospatial AI, enabling precise, data-driven decisions with efficiency and confidence.

Our Commitment to People

We are committed to aligning business growth with professional outcomes for every employee. Our commitment has been recognized by Jacksonville Business Journal and Will Reed as a noted Best Places to Work.


About the role

We are looking for a skilled Data Engineer to design, build, and maintain scalable data pipelines and platforms that support our geospatial traffic analytics applications. The ideal candidate will have experience with Python, Databricks, S3, and modern data engineering practices, including automated testing, CI/CD, and data quality monitoring.

Responsibilities

  • Design, implement, and maintain scalable data pipelines and ETL/ELT workflows on Databricks and cloud platforms.
  • Manage large-scale geospatial and temporal datasets stored in AWS S3.
  • Collaborate with data scientists to productionize machine learning models and ensure smooth data availability.
  • Implement data validation, testing, and monitoring frameworks to ensure data accuracy, consistency, and reliability.
  • Optimize data storage and processing strategies to handle high volumes of traffic and mobility data efficiently.
  • Develop and maintain documentation for data workflows, architecture, and processes.
  • Work closely with cross-functional teams to understand data requirements and ensure timely delivery.
  • Stay up-to-date with the latest trends and best practices in data engineering, cloud technologies, and big data processing.

Qualifications

  • Bachelor's or Master's degree in Computer Science, Data Engineering, or a related field.
  • 3+ years of experience as a data engineer or in a similar role.
  • Strong proficiency in Python and associated libraries for data engineering (pandas, PySpark, etc.).
  • Hands-on experience with Databricks and Spark for large-scale data processing.
  • Experience with AWS services, especially S3, and knowledge of cloud-based data architectures.
  • Solid understanding of data pipeline testing, version control, and CI/CD practices.
  • Experience with SQL and NoSQL databases.
  • Strong problem-solving skills and attention to detail.


Preferred Skills

  • Familiarity with geospatial data formats and processing (GeoJSON, Shapefiles, PostGIS).
  • Experience with workflow orchestration tools (Databricks, Prefect, or similar).
  • Knowledge of containerization (Docker/Kubernetes) and cloud-native data solutions.
  • Experience supporting machine learning pipelines in production.


Compensation

  • Location: Jacksonville, FL (Town Center Area) or Remote
  • Type:  Full-time
  • Reports to: Director of Engineering
  • Salary Based on Experience 
  • Annual Bonus
  • Medical, Vision, Dental, 401(k)  
  • 21 Days Vacation
  • Office Lunch provided Daily

Similar Jobs

3 Days Ago
Remote or Hybrid
113K-193K Annually
Senior level
113K-193K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design and operate scalable batch and streaming data platforms supporting machine learning and generative AI. Build pipelines for structured, unstructured, OCR, document, and image data; develop RAG, semantic search, and LLM-powered solutions; and establish data quality, observability, governance, orchestration, and deployment practices. Partner with stakeholders, lead platform scalability and cost optimization, mentor engineers, and translate ambiguous needs into production-ready technical roadmaps while securely handling sensitive data.
Top Skills: AirflowAmazon KinesisSparkAWSAzureAzure Event HubsChart.JsDatabricksDeequDelta LakeDockerGithub ActionsGCPGreat ExpectationsJavaKafkaKubernetesLlmsMlopsPlotlyPysparkPythonRagScalaSeabornSnowflakeSQLTerraform
4 Days Ago
Remote
United States
Mid level
Mid level
Artificial Intelligence • Information Technology • Professional Services • Software • Analytics • Generative AI • Big Data Analytics
Configure and maintain Adobe Experience Platform and Real-Time CDP data pipelines, XDM schemas, datasets, identity resolution, Profile enablement, segmentation readiness, and destination activation. Validate data quality, troubleshoot ingestion and identity issues, manage sandboxes, and implement privacy, consent, governance, and access controls. Partner with architects, data engineers, consultants, analysts, data scientists, and marketing teams to support reporting, personalization, and machine learning readiness.
Top Skills: Adobe AnalyticsAdobe Experience PlatformAdobe I/O RuntimeAdobe Journey OptimizerAdobe Real-Time CdpAdobe Source ConnectorsAdobe TargetAPIsAWSAzureBatch IngestionBigQueryCcpaData PrepEltETLGCPGdprJavaScriptPythonQuery ServiceRedshiftSalesforce CdpSegmentSnowflakeSQLStreaming IngestionWeb SdkXdm
2 Days Ago
In-Office or Remote
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Lead cloud data modernization by designing and building cloud-native ETL/ELT pipelines, data warehouses, and lakehouse architectures using Azure, Snowflake, and Databricks. Ensure data quality, security, governance, CI/CD, and healthcare interoperability while mentoring engineers and collaborating with stakeholders to deliver scalable, HIPAA-compliant analytics platforms.
Top Skills: AdlsApache AirflowAzureAzure Data Factory (Adf)Azure Data Lake Storage Gen2 (Adls Gen2)Azure OpenaiCi/CdCptDatabricksDatabricks GenieDatabricks Mosaic AiDockerEltETLFhirGitGithub ActionsHcpcsHl7Icd-10KafkaKubernetesLlmsLoincPysparkPythonRag PipelinesSnowflakeSnowflake CortexSparkSQL ServerSsisTerraformVector StoresVisioX12 Edi (837/835/834)

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account