At Lumenalta, we partner with forward-thinking organizations to build technology solutions that scale, delight users, and accelerate business growth. Our global teams bring curiosity, commitment, and technical excellence to every project. We value transparency, autonomy, and impact, empowering every team member to do their best work.
We're seeking a Senior Data Engineer with deep expertise in distributed streaming architectures to join a high-impact enterprise engagement. This role is an opportunity to architect and deliver complex pipelines across multiple sources, from raw landing zones through to analytics-ready curated layers.
What You’ll Be Doing
- Architect and implement data landing zones for high-volume streaming event data (Kafka Design and deploy end-to-end batch and streaming pipelines on AWS using Spark and Flink, with reliable state management, exactly-once semantics, and incremental processing.
- Develop transformation logic to calculate descriptive statistics (min, max, avg, counts) at ingestion time and build advanced curated-layer pipelines for complex multi-stream time alignment.
- Tune and manage the Spark and Flink compute environment
- Lead production cutover activities, including environment hardening, catch-up batch processing for historical backfill, and phased loading of historical data. *Define and enforce data quality standards and pipeline observability practices across all layers of the platform.
- Collaborate with cross-functional stakeholders to translate domain requirements into robust, scalable data architecture decisions.
What We’re Looking For
- 5+ years in data engineering with a proven track record of delivering production-grade pipelines in complex, high-volume enterprise environments.
- Advanced hands-on experience with Apache Kafka at scale, including topic and partition design, consumer group management, schema registry, and operational tuning.
- Deep experience with Apache Spark (batch and Structured Streaming) and Apache Flink for stateful stream processing, windowing, and event-time handling.
- Strong AWS background across the data stack: MSK or Kinesis, EMR, Glue, S3, Lambda, and IAM. *Strong proficiency in Python/PySpark and advanced SQL for both pipeline development and complex transformation logic.
- Experience managing high-frequency time-series datasets and performing synchronization across misaligned temporal streams. *Familiarity with cloud cost management practices, including compute sizing, spend monitoring, and resource optimization. *Comfortable owning go-live activities: environment hardening, cutover planning, and post-deployment stabilization in critical production settings.
Nice to Have
Experience working with HIPAA-regulated data and PHI, including secure handling, anonymization, access controls, and audit requirements.
Why Lumenalta is an amazing place to work at
At Lumenalta, you can expect that you will:
- Be 100% dedicated to one project at a time so that you can innovate and grow.
- Be a part of a team of talented and friendly senior-level developers.
- Work on projects that allow you to use leading tech.
Location
This is a fully remote position; however, candidates must be based in regions that align with the Pacific, Central, or Eastern U.S. time zones to ensure effective collaboration with client and team schedules.
Application Deadline
Applications will be accepted until October 11, 2026. Candidates can expect feedback by October 19, 2026.

