Job Title: Senior Data Engineer
Location: Remote (U.S. — EST preferred)
Employment Type: Contract (W2 through ZipStaff)
About the Opportunity:
ZipStaff is seeking a Senior Data Engineer to build complex data pipelines on a modern data-ingestion platform for a leading healthcare analytics organization.
You will deliver accurate, timely data across products that support healthcare data management, validation, and analytics. This role requires strong big-data pipeline skills (Hadoop / Spark), solid RDBMS/SQL depth, and the ability to troubleshoot production data issues in a fast-paced Agile environment.
What You'll Do:
- Build data pipelines from transformation specifications to load source data into the data lake using a proprietary big-data processing platform.
- Support and improve current data-ingestion processes for proprietary healthcare data applications.
- Develop and maintain data-engineering processes using T-SQL, Spark, Scala, and shell scripting.
- Review and test data for accuracy and validity before loading to the data lake.
- Perform data analysis, data mining, and root-cause investigation using modern analysis tools.
- Work with Technical Operations to troubleshoot complex database / environment issues (OS, storage, servers).
- Provide off-hours support to resolve production issues when necessary.
- Develop tools and techniques that improve process efficiency and data performance.
- Mentor junior team members in data engineering and quality best practices.
- Support delivery of business priorities in a Scaled Agile Framework (SAFe) environment.
Required Qualifications:
- Bachelor's degree in Computer Science, Engineering, or a related field, with 10+ years of industry experience.
- 8+ years of experience with data aggregation, standardization, linking, quality-check mechanisms, and reporting.
- 5+ years of experience with big-data technologies such as Hadoop and Spark.
- 5+ years of experience with RDBMS (Oracle or MS SQL Server) and SQL or other data-integration / ETL tools.
- Solid Linux experience, including shell scripting and file systems.
- Ability to debug data issues, identify root cause, and fix problems in a fast-paced environment.
- EST timezone preferred.
- Must be legally authorized to work in the United States without sponsorship now or in the future.
Preferred Qualifications:
- Healthcare data experience, preferably in a data-operations role.
- Experience with Spark and Scala on large-scale ingestion workloads.
- Experience operating in an Agile / SAFe environment.
- Advanced SQL and end-to-end Linux / Hadoop operations experience.
About ZipStaff:
ZipStaff partners with leading organizations to connect skilled technology professionals with high-impact contract opportunities. We focus on quality matches and long-term success.
To apply, please submit your resume highlighting your Hadoop/Spark pipeline work, SQL Server or Oracle ETL experience, and any healthcare data background.