Responsibilities: Design, develop, and maintain scalable data ingestion pipelines to onboard structured, semi-structured, and unstructured data from batch and streaming sources (e.g., APIs, databases, flat files, message queues) into the Azure/Databricks environment. Work with unstructured or structured data and converting those data sets using a variety of analyses such as optimization, simulation, classical and spatial statistics, and/or programming languages.