Preply

Data Engineer

Гибрид · Барселона, Испания · Английский B2

Не указано: грейд

Навыки

  • Airflow
  • AWS
  • Data Lake
  • Data Quality
  • dbt
  • Flink
  • Форензика и реагирование
Ещё 11
  • Google Cloud
  • Kafka
  • Обучаемость
  • Machine Learning
  • Мониторинг и observability
  • Ответственность за результат
  • Решение задач
  • Retention и отток
  • SOLID и паттерны
  • Spark
  • Потоковая обработка

О компании и продукте

  • Care to change the world - We are passionate about our work and care deeply about its impact to be life changing.
  • We do it for learners - For both Preply and tutors, learners are why we do what we do. Every day we focus on empowering tutors to deliver an exceptional learning experience.
  • Keep perfecting - To create an outstanding customer experience, we focus on simplicity, smoothness, and enjoyment, continually perfecting it as every detail matters.
  • Now is the time - In a fast-paced world, it matters how quickly we act. Now is the time to make great things happen.

Задачи

  • Contribute to trusted ingestion & enrichment foundations (Data Lake and Data as a Product)
  • Build and maintain components of Preply's data lake
  • Ensure every dataset has clear ownership, purpose, schemas, and quality expectations from first ingestion through downstream consumption by analytics, product, and ML teams
  • Treat trust, correctness, and predictability as first-class features of the platform
  • Develop end-to-end ingestion pipelines (batch & streaming)
  • Build and operate reliable batch and streaming ingestion pipelines that support both real-time and analytical use cases. Contribute to defining clear raw
  • standardized
  • consumption layers with explicit responsibilities, lineage, and retention strategies. Balance performance, cost, and reliability as the platform scales
  • Data quality, contracts & early validation
  • Implement data contracts between producers and consumers, covering schema, freshness, volume, and quality guarantees
  • Embed validation, anomaly detection, and quality checks early in the ingestion lifecycle to catch issues before they propagate
  • Enrichment, modeling & lifecycle management
  • Build enrichment logic that joins, standardizes, and contextualizes data across domains using shared definitions and reusable patterns
  • Support historical tracking, point-in-time correctness, and dataset versioning so downstream users can confidently analyze changes and impacts over time
  • Observability, reliability & operational excellence
  • Instrument ingestion pipelines with strong observability: freshness, latency, data quality, and cost metrics
  • Contribute to SLOs, alerting, and incident response playbooks so data failures are visible, diagnosable, and recoverable
  • Help move the platform from reactive firefighting to proactive reliability management
  • Governance & compliance by design
  • Ensure sensitive data is properly masked, minimized, or anonymized by default, and that all data flows you own are auditable and traceable
  • Enable self-service & standardization
  • Contribute to standardized ingestion templates, shared libraries, and platform tooling that enable teams to onboard new data sources independently
  • Improve discoverability, documentation, and metadata so datasets you own are easy to find and trust without relying on tribal knowledge
  • Cross-team collaboration & ownership
  • Work closely with Product, Backend, Analytics, and ML partners to align on ingestion requirements and trade-offs

Требования

  • Hands-on experience building components of large, high-scale applications (e.g., data pipelines, well-structured APIs, efficient algorithms)
  • Solid experience working in platform or data engineering teams (or equivalent) with the ability to deliver within a multi-stakeholder environment
  • Familiarity with cloud platforms (AWS/GCP or equivalent) and modern DevOps practices
  • Hands-on experience designing and implementing real-time and batch data processing pipelines using modern frameworks like Spark, Flink, Spark Streaming, Kafka, Debezium, etc
  • Experience with orchestration tools such as Airflow, dbt, or similar
  • Exceptional problem-solving skills paired with a proactive, innovative mindset focused on continuous improvement
  • Strong communication and cross-functional collaboration skills (English level B2+)

Условия

  • An open, collaborative, dynamic, and diverse culture
  • A generous monthly allowance for lessons on http://preply.comPreply.com http://Preply.com, Learning & Development budget, and time off for your self-development
  • A competitive financial package with equity, leave allowance, and health insurance
  • Access to free mental health support platforms
  • The opportunity to unlock the potential of learners and tutors through language learning and teaching in 175 countries (and counting!)
  • We’ve just reached unicorn status with a $150M Series D, accelerating our vision to transform education through human-led, AI-enhanced learning
  • Today, 100,000+ tutors teach 90+ languages to learners in 180 countries - and we’re only getting started
  • As a category-defining company, we’re shaping what the future of learning looks like at global scale
  • Every Preply lesson sparks change, fuels ambition, and drives progress that matters
  • Joining Preply means helping define the future of education at global scale, and building something that truly matters for millions of people, every day
  • MEET THE TEAM!
  • At Preply, the Data Ingestion and Enrichment team provides a single, trusted, and scalable data foundation
  • The team ensures that all analytics, machine learning, and product features are built on unified, governed, and production-grade data assets in Preply's Lake House, including the extraction, normalization, and generation of structured data from Preply's unstructured assets, forming a durable data moat for AI-driven products

Паспорт вакансии

История публикации

Появилась в Вакандии30 дней
Перепубликации5 разпубликаций всего: 6
Проверяли на источникеВидели 26 дней назад
Среди похожихНет данныху карточки не хватает полей, чтобы найти похожие
Долго открыта
Вакансия поднималась в источнике 5 раз. Это не значит, что она закрыта, но откликаться стоит с этой оговоркой.

Откуда что взялось

Отмечено то, что вывели мы. Без пометки — значение назвал работодатель.

Грейдне указан
Формат работыГибрид
ГеографияБарселона, Испаниявычитано из текста вакансии
Зарплата≈ 15 833 USD в месяцнаша оценка, в вакансии не названа

Почему на этом месте в выдаче

Порядок выдачи объявлен контрактом: свежесть решает между днями, полнота и зарплата — внутри дня.

Полнота карточки753 из 4 полей: грейд, формат, география, зарплата
Зарплата названа0вилки работодателя нет, показана наша оценка

Проверка Вакандии

Источники и свежесть

Тип источника
Карьерный сайт работодателя
Найдено публикаций
6
Посмотреть публикации и даты
  • ashbyОсновная публикация · 2026-07-29
  • ashbyПовторная публикация · 2026-07-29
  • ashbyПовторная публикация · 2026-07-29
  • ashbyПовторная публикация · 2026-07-29
  • ashbyПовторная публикация · 2026-07-29
  • ashbyПовторная публикация · 2026-07-29

Работодатель

Preply

39 активных вакансий · вилка работодателя указана в 0%

Открыть профиль компании

Безопасность

Отклик уходит на сайт источника

Вакандия показывает вакансию, но не отправляет отклик и не проверяет работодателя. Сам отклик вы оставляете на внешнем сайтеjobs.ashbyhq.com.

Признаки мошенничества
  • Просят предоплату, «залог» или деньги за обучение и оборудование.
  • Требуют код из SMS, данные банковской карты или доступ к «Госуслугам».
  • Быстро уводят в мессенджер и торопят с решением.
  • Обещают большой доход без опыта и без деталей задач.

Настоящий работодатель не просит денег и платёжных данных до трудоустройства.

Продолжить поиск

Похожие вакансии

Причина сходства указана на каждой карточке