We are working with a leading provider of information technology and workforce solutions, helping companies modernize and optimize their analytics data assets.
The Role
- Refactor and optimize legacy SQL/PySpark logic into modular, scalable pipelines.
- Architect reliable, production-grade analytics data products with strong governance standards.
- Process large-scale data workloads using PySpark and Databricks.
- Collaborate with the core analytics team to create assets for handoff to broader engineering and offshore teams.
- Standardize data product outputs, including schema enforcement and metadata documentation.
What You'll Need
- Expert-level experience with PySpark and Databricks for large-scale data processing.
- Advanced SQL expertise, particularly in AWS Redshift environments with high-volume queries.
- Proven ability to transition experimental sandbox work into production-ready assets.
- Experience in refactoring legacy SQL/PySpark logic into optimized pipelines.
- Strong communication and partnership skills to work effectively with analytics and engineering teams.
What's On Offer
- Opportunity to work on modernizing and optimizing critical analytics data assets.
- Flexible work environment with remote options.
- Competitive hourly pay rate.
- Access to benefits including medical, dental, vision, and 401K contributions for full-time consultants.
Apply via Haystack today!