Job description
The Senior Data Engineer will design, build, and optimize large-scale data infrastructure, ensuring efficient data collection, processing, and availability to support advanced analytics and data-driven decision-making.
Role description:
Design ETL/ELT pipelines for network telemetry data
Interface with on-premises data collection systems
Implement Iceberg table structures, also for time-series data
• Optimize data processing for analytics
• Manage data quality and lineage
• Implement pipeline observability
Implement modern data protection strategies
Qualifications:
5+ years experience with AWS Glue, S3, Athena, IAM, Lakeformation, Lambda, S3 Tables
Apache Iceberg and lakehouse architectures
Java (Python, Scala) for data processing
Familiarity with data formats (JSON, Avro, Parquet)
• Time-series database queires
Data processing (Spark, DuckDB, Flink, AirFlow, Prefect)
• Graph Theory
• Excellent SQL Knowledge
• English: B1
Nice to have:
• AWS Certifications
Originally posted on Himalayas
Who can apply
Eligible countries: United States. Accepted UTC offsets: UTC-10, UTC-9, UTC-8, UTC-7, UTC-6, UTC-5, UTC+14. Review the full description for employer-specific work authorization, residency and schedule requirements.