About the company
Webflow is the agentic web marketing platform for modern marketing teams, helping organizations build, manage, and optimize high-performing web experiences.
Responsibilities
- Design data solutions across batch, streaming, and real-time workloads
- Design and build reliable data pipelines using Spark, Kafka, Iceberg, and Airflow/MWAA
- Contribute to the evolution of data lake and data platform, including ingestion, processing, storage, and serving patterns
- Build, maintain and scale cloud data infrastructure that includes Kafka, Airflow, Druid, EMR on EKS with reliability and observability in mind
- Implement and improve data quality, observability, reliability, and governance across data pipelines and systems
- Define and improve standards for event instrumentation and governance, including event schemas, validation, and schema evolution
- Build platform solutions for centralized PII handling, data retention, deletion, and residency to support GDPR and CCPA requirements
- Build agent harness capabilities that encode data engineering standards, pipeline patterns, and quality gates into AI tools
- Contribute hands-on to complex technical initiatives, troubleshoot production issues, mentor engineers
Requirements
- BA/BS degree or equivalent experience
- 5+ years of experience in data engineering, building and operating production data pipelines and platforms
- Strong experience with Spark and distributed data processing
- Experience designing batch, real-time, and streaming data solutions using technologies such as Kafka, Spark Structured Streaming, and CDC
- Experience implementing data quality, observability, and reliability across production data systems
- Experience with cloud data infrastructure, CI/CD, and infrastructure-as-code
- Demonstrated experience leading complex data engineering initiatives end-to-end
- Strong Programming and SQL skills