D
Posted Yesterday•Remote
Data Engineer (Python, PySpark & Databricks)
MiddleRemoteSalary undisclosed
Required Skills
PythonAzureSQLCI/CDDatabricksApache Spark
Job Description
We are looking for a Data Engineer to design, build, modernize, and operate scalable data pipelines supporting Tax technology platforms. This role focuses on developing and enhancing data processing workloads using Python, Apache Spark, PySpark, and Databricks, while ensuring reliable, secure, and maintainable data solutions. The ideal candidate has strong hands-on engineering skills, experience with production ETL/ELT pipelines, data ingestion, APIs, data quality, and cloud-based data platforms, along with strong problem-solving and collaboration skills.
Responsibilities
- Design, build, and support scalable data pipelines using Python, Apache Spark, PySpark, and Databricks.
- Modernize and migrate legacy data processing workloads to secure, cloud-native platforms.
- Build and maintain batch data ingestion pipelines from structured and unstructured sources.
- Integrate data from REST APIs, SharePoint, document repositories, enterprise applications, and cloud platforms.
- Implement data quality, monitoring, observability, and operational controls.
- Optimize data workloads for performance, scalability, reliability, and cost efficiency.
- Develop document extraction, classification, metadata enrichment, and automation pipelines.
- Apply software engineering practices including Git, CI/CD, automated testing, and code reviews.
- Build, deploy, troubleshoot, and maintain production ETL/ELT p
Ready to apply? Optimize your CV for this specific jobAI customizes your experience bullets and increases chances to get hired.