- Information Systems
- Remote within the U.S. (100%)
- Full-Time (Exempt)
- 110,000—130,000 **
- Medical
- Dental
- Vision
- PTO
- OPEN: Wednesday, August 12, 2026
- CLOSES: Open until filled
Role Summary
The Data Integration Engineer designs, develops, deploys, administers, and continuously improves the company’s data-integration platform and end-to-end data pipelines. This role manages Apache NiFi-based ETL workflows and supporting SQL, Python, Amazon Aurora/RDS, SFTP, and SMB integrations across development, test, and production environments and helps evaluate, implement, and operate Amazon Redshift as the company’s analytical data platform.
Working closely with Data Analytics and other technical and operational teams, the engineer ensures that client data is securely ingested, mapped, transformed, validated, reconciled, loaded, and reliably delivered to internally developed application tools, business-intelligence solutions, and other downstream consumers. This position combines hands-on data integration engineering with NiFi platform administration, database and data-warehouse development, systems analysis, production support, and data-quality management.
Key Responsibilities
• Design, develop, test, deploy, monitor, and maintain Apache NiFi dataflows supporting inbound client data, internal processing, and outbound data delivery.
• Create source-to-target mappings and implement data extraction, transformation, validation, routing, loading, reconciliation, and exception-handling logic.
• Develop and maintain SQL scripts and integration-related Amazon RDS objects, including schemas, tables, views, materialized views, stored procedures, and supporting database structures.
• Configure and maintain NiFi processors, controller services, parameter contexts, queues, retry paths, back-pressure controls, external services and controllers, logging, and related platform components.
• Support SFTP- and SMB-based ingress and egress pipelines, including connectivity, authentication, permissions, directory structures, and file-handling conventions.
• Develop and maintain Python scripts used for file validation, pre-load cleansing, normalization, reconciliation, and exception identification.
• Investigate rejected, incomplete, or inaccurate data loads; identify root causes; correct or normalize affected data under established controls; and execute documented replay procedures.
• Ensure that integrated data is accurately structured and presented for internally developed applications, BI solutions, reporting processes, and analytical datasets.
• Work closely with the Data Analytics team, a primary internal consumer of the integrated data, to understand requirements, resolve discrepancies, and validate downstream usability.
• Participate in the evaluation, proof of concept, architecture, implementation, and ongoing operation of Amazon Redshift or other cloud data warehouse solutions, including analytical data modeling, schema and database-object development, ETL/ELT design, workload monitoring, performance optimization, access controls, data quality, lineage, and integration with internal applications and BI platforms.
• Support the onboarding of new clients and data sources while maintaining the reliability and accuracy of existing production integrations.
• Create and maintain version-controlled technical documentation, operational procedures, data mappings, scripts, validation rules, system dependencies, and recovery processes.
• Help establish and maintain data governance practices for data quality, lineage, definitions, access, retention, documentation, and change control as the company’s data infrastructure and operations mature.
Required Qualifications
- Bachelor’s degree in computer science, information systems, information technology, engineering, or a related technical discipline; equivalent relevant professional experience may be considered.
- 3+ years of hands-on experience developing, administering, and troubleshooting Apache NiFi dataflows or a closely comparable enterprise ETL or data-integration platform.
- 3+ years of direct experience working with PostgreSQL, production database processes, and data-warehouse environments.
- Advanced SQL skills in PostgreSQL environments, including experience creating and maintaining database objects and developing logic for data mapping, transformation, validation, reconciliation, and loading.
- Experience developing and maintaining Python scripts for data validation, cleansing, automation, or file processing.
- Ability to troubleshoot issues across source files, ETL workflows, databases, integration services, internal applications, and downstream reporting or analytics.
- Experience with secure file-transfer technologies and network file-sharing protocols, including SFTP and SMB.
- Exceptional attention to detail and a disciplined approach to data accuracy, reconciliation, documentation, versioning, and controlled production changes.
- Excellent written and verbal communication skills, with the ability to explain technical issues, data defects, processing status, and implementation requirements to both technical and nontechnical stakeholders.
- Ability to balance concurrent production support, client onboarding, development, maintenance, and infrastructure-improvement priorities in a fast-moving and evolving environment.
Preferred Qualifications
- Master’s degree in computer science, information systems, information technology, engineering, or a related technical discipline.
- 5+ years building and supporting separate NiFi development, test, and production environments, including multi-node or clustered deployments.
- 3+ years of experience designing, implementing, or operating Amazon Redshift—Serverless or provisioned—including analytical data modeling, warehouse loading patterns, materialized views, workload monitoring, performance tuning, and secure access to sensitive data; experience with Amazon Aurora/RDS zero-ETL integration is a plus.
- Experience implementing automated data-quality controls, exception reporting, controlled data replay, and root-cause corrective actions.
- Experience integrating operational data with internally developed applications, BI platforms, analytics workflows, or regulated healthcare-data environments.
- Experience working with protected health information (PHI) and healthcare data in HIPAA-regulated environments, including secure data handling, minimum-necessary access, privacy safeguards, and sound health information management practices.
- Experience working with health insurance claims, eligibility, and entitlement data, including coordination of benefits (COB), payment integrity, member coverage, and claims-processing workflows.
Technology Stack
Data Integration & Development: Apache NiFi, PostgreSQL, SQL, Python
AWS: Amazon Aurora/RDS, Amazon Redshift, Amazon S3, Amazon EC2
File Transfer & Connectivity: SFTP, SMB
Environment & Tooling: Linux, VS Code, GitHub, Codex
** The base salary range for this full-time, exempt position is $110,000 to $130,000 USD. Salary ranges are determined by role and level. Compensation is determined by additional factors, including job-related skills, experience, and relevant education or training. This compensation reflects the annual salary only, and does not include equity or benefits.
Questions?