About the company
At Preply, we’re all about creating life-changing learning experiences. We help people discover the magic of the perfect tutor, craft a personalised learning journey, and stay motivated to keep growing. Our approach is human-led, tech-enabled - and it’s creating real impact.
We’ve just reached unicorn status with a $150M Series D, accelerating our vision to transform education through human-led, AI-enhanced learning. Today, 100,000+ tutors teach 90+ languages to learners in 180 countries - and we’re only getting started.
Responsibilities
- Build trusted ingestion & enrichment foundations (Data Lake and Data as a Product)
- Own end-to-end ingestion pipelines (batch & streaming)
- Data quality, contracts & early validation
- Enrichment, modeling & lifecycle management
- Observability, reliability & operational excellence
- Governance & compliance by design
- Enable self-service & standardization
- Cross-team collaboration & ownership
Requirements
- Exposure to and experience building architectural patterns of a large, high-scale application
- Solid experience working in platform or data engineering teams
- Familiarity with cloud platforms (AWS/GCP or equivalent) and modern DevOps practices
- Hands-on experience with real-time and batch data processing using Spark, Flink, Spark streaming, Kafka, Debezium, etc.
- Expertise with orchestration tools such as Airflow, dbt, or similar
- Exceptional problem-solving skills
- Strong communication and cross-functional collaboration skills (English level B2+)