Resumen del rol

Data Engineer

Requisitos y responsabilidades

Contenido del rol extraído en secciones para revisar más rápido.

The Data Engineer Role.

  • Design, build, and operate large-scale data processing pipelines handling multi-terabyte and streaming datasets, including audio/video transcoding, feature extraction, and preprocessing workflows.
  • Deploy, scale, and troubleshoot containerized workloads on Kubernetes and AWS in production environments.
  • Build and maintain distributed data processing jobs using frameworks such as Spark and Ray.
  • Design and operate workflow orchestration systems (e.g., Airflow) with dependency management, retries, monitoring, and alerting for production pipelines.
  • Administer and tune enterprise databases, including performance tuning, backup/recovery, access control, and scaling strategies.
  • Partner with ML engineers and researchers to support training pipelines, model retraining triggers, feature stores, and other MLOps workflows.

Who you are.

  • Hands-on experience with Kubernetes and AWS, including deploying, scaling, and troubleshooting containerized workloads in production environments.
  • Proficiency with high-performance/distributed computing frameworks such as Spark and Ray for processing large-scale datasets.
  • Experience with workflow orchestration tools such as Airflow (or comparable systems like Dagster, Prefect, or Luigi) to schedule and manage complex data pipelines.
  • Strong programming skills in Python and SQL; experience with Golang is a plus.
  • Demonstrated track record building and operating large-scale data processing pipelines, ideally handling multi-terabyte or streaming datasets.
  • Experience working with audio or video data at scale is a strong plus (e.g., transcoding, feature extraction, or preprocessing pipelines).
  • Familiarity with common data transformation patterns applied to large datasets (ETL/ELT, batch and stream processing, data validation and quality checks).
  • Experience designing and maintaining job orchestration systems, including dependency management, retries, monitoring, and alerting for production pipelines.
  • Bonus: experience orchestrating machine learning workflows (training pipelines, model retraining triggers, feature stores, or MLOps tooling).

What we offer.

  • Healthcare plans with 100% premium coverage for employees and partial coverage available for dependents
  • Dental and Vision plans with 100% premium coverage for employees and their dependents
  • Short/Long-term disability and life insurance plans with 100% premium coverage for employees
  • FSA/HSA and 401k programs
  • Equity compensation
  • 20 days of PTO per year
  • 12 weeks of Parental Leave
  • Learning and Development budget
  • Monthly wellness benefits
  • Annual company-sponsored offsite

What we offer.

  • Daily in-office lunch through UberEats
  • Commuter benefits
  • Remote Fridays
  • Happy Hours and other local events
Roles similares

Mantén una lista de respaldo.

Ver stack
FocoData EngineerÁrea del rol
Señal de seniorityMiddleNivel del candidato
StackAWS, Golang, KubernetesSkills principales
Ubicación1 país aceptadoElegibilidad

Stack

Usa estas tags para comparar roles remotos similares.

Elegibilidad de ubicación

Candidatos deberían aplicar solo cuando el país del perfil aparece aquí.

Tu perfilPaís no definidoInicia sesión para comparar tu país con este rol.

Flujo de contratación

WithMira muestra el rol y luego envía candidatos a la aplicación de la empresa.

1Revisa fit del rol, stack y elegibilidad de ubicación en WithMira.
2Abre la página de aplicación de la empresa desde el link rastreado.
3Guarda el rol o suscríbete a oportunidades similares antes de salir.