Datos y analítica
Senior Data Engineer - Azure, Databricks & ML Pipelines | Remote - vacante remota
Toptal
Senior Data Engineer - Azure, Databricks & ML Pipelines | Remote - Datos y analítica, remoto
La empresa Toptal busca cubrir el puesto de Senior Data Engineer - Azure, Databricks & ML Pipelines | Remote: es un puesto senior, así que la empresa espera que las decisiones las tome usted. La vacante pertenece al área de Datos y analítica y es totalmente remota. La empresa no pone ninguna restricción sobre el lugar de residencia, así que usted puede postular desde donde vive.
Las herramientas principales que nombra el anuncio son: Python, SQL, Machine Learning. El CV debería mostrar ejemplos concretos justo de esas capacidades.
La empresa no publicó ninguna cifra, eso se aclara en la entrevista. Sobre el horario, el anuncio no pone ninguna condición.
Esta vacante pasó una revisión automática: los anuncios que exigen un permiso de trabajo extranjero, patrocinio de visa, una nacionalidad concreta o residencia en un país determinado no entran en la lista.
Herramientas
En resumen
- Empresa
- Toptal
- Área
- Datos y analítica
- Quién puede postular
- Desde cualquier país del mundo
- Modalidad
- Totalmente remoto
- Nivel
- Senior
- Herramientas
- Python, SQL, Machine Learning
- Publicada
- 29 de julio de 2026 (hace 26 días)
- Activa
- hasta el 7 de septiembre de 2026
- Fuente
- We Work Remotely
Vacantes nuevas por correo: Datos y analítica
El tablón se actualiza varias veces al día. Escribimos solo cuando aparece una vacante nueva ya verificada. Nada de spam y no hace falta cuenta.
¿Ya está suscrito? Gestionar sus preferencias
Descripción de la empresa
Headquarters: Remote
URL: https://www.toptal.com/
About the Role
We're looking for a Senior Data Engineer to design, build, and maintain scalable data pipelines and ML-ready infrastructure on Azure and Databricks. This is a hands-on engineering role: you'll own the full data pipeline lifecycle - ingestion, transformation, orchestration, and deployment - while supporting machine learning workflows with clean, reliable data. If you're comfortable owning infrastructure decisions and writing production-quality Python at scale, this role is built for that.
What You'll Do
* Design, build, and maintain data pipelines using Databricks and Azure-native data services
* Develop and optimize ETL/ELT processes to support analytics and machine learning workloads
* Build and maintain CI/CD pipelines for data engineering and ML deployment workflows
* Write clean, efficient, production-quality Python for data processing and pipeline automation
* Support machine learning teams with well-structured, high-quality datasets and feature pipelines
* Design and manage data architecture across Azure services (e.g., Azure Data Factory, Azure Data Lake, Azure Synapse)
* Monitor pipeline performance, troubleshoot data quality issues, and implement reliability improvements
* Implement data governance, security, and access control best practices
* Collaborate with data scientists, analysts, and software engineers to align data infrastructure with business needs
* Participate in code reviews, architecture discussions, and technical planning
What You Bring
* Strong hands-on experience with Azure cloud data services
* Proven experience building and maintaining pipelines on Databricks
* Solid experience designing and managing CI/CD pipelines for data or ML workflows
* Strong Python skills for data engineering and pipeline development
* Working knowledge of machine learning workflows and how data engineering supports them
* Experience with SQL and relational/distributed data systems
* Understanding of data pipeline orchestration, monitoring, and reliability practices
* Strong problem-solving skills and ability to work independently on complex data infrastructure challenges
* Solid communication skills for collaborating with data science and engineering teams
Nice to Have
* Experience with MLOps practices and tools (MLflow, Azure ML)
* Familiarity with Spark internals and performance tuning within Databricks
* Experience with infrastructure-as-code (Terraform, Bicep, ARM templates)
* Exposure to real-time/streaming data pipelines (Kafka, Event Hubs, Structured Streaming)
* Relevant Azure or Databricks certifications
Why This Role
* Full pipeline ownership: Own data infrastructure end to end, from ingestion through ML-ready delivery
* Modern data stack: Work with Azure and Databricks, leading platforms in enterprise data engineering
* Cross-functional impact: Directly enable machine learning and analytics outcomes, not just move data
* Flexibility: Remote-friendly engagement structure
How to Apply
Ready to bring your data engineering expertise to Azure and Databricks-powered ML infrastructure? Apply through Toptal here: https://www.toptal.com/talent/apply
To apply: https://weworkremotely.com/remote-jobs/toptal-senior-data-engineer-azure-databricks-ml-pipelines-remote
El texto se conserva en el idioma original de la empresa, porque en ese mismo idioma se postulará usted.
Preguntas frecuentes sobre esta vacante
¿Puedo postularme al puesto de Senior Data Engineer - Azure, Databricks & ML Pipelines | Remote desde donde vivo?
Sí. Para esta vacante, Toptal acepta candidatos de cualquier país del mundo, así que usted no necesita un permiso de trabajo de otro país. El anuncio pasó una revisión automática: si la empresa hubiera exigido un permiso de trabajo, patrocinio de visa o residencia en un país determinado, no estaría en este tablón.
¿Qué remuneración se indica?
Para esta vacante, Toptal no publicó ninguna cifra. La mayoría de los anuncios remotos no publica un monto: se acuerda en la entrevista.
¿Cómo me postulo?
Usted se postula directamente ante la empresa, a través del anuncio original publicado en We Work Remotely. Donator no recibe candidaturas, no cobra comisión y no guarda ningún CV.
¿Qué tipo de vacante es esta?
Un puesto totalmente remoto en el área de Datos y analítica. Los anuncios híbridos y todo lo que exija presencia en la oficina no se publican en este tablón.
Certificados gratuitos para esta vacante
Este anuncio pide Python, SQL, Machine Learning. Abajo están las credenciales gratuitas que cubren justo esas herramientas.
Kaggle Learn: 17 micro-courses
certificateKaggle · ~4 h
Intro to ML, Pandas, SQL, Deep Learning, Computer Vision and more. Only a Google account is needed, and the certificate has a public link.
moderate weightCS50x: Introduction to Computer Science
certificateHarvard CS50 · ~100 h
Harvard's own branded certificate is free once you pass every problem set and the final project. The edX verified certificate is a separate paid product and is not needed.
strong brandDeveloper certificates (10+ tracks)
certificatefreeCodeCamp · ~300 h
A non-profit; a card is never requested. The certificate is issued after five projects are submitted and pass.
moderate weightDeep Reinforcement Learning course
certificateHugging Face · ~30 h
Entirely free with no deadline. 80% earns a Certificate of Completion, 100% a Certificate of Honors.
moderate weight
Vacantes similares
Senior ML Engineer (Europe-based/Remote)
SWORD Health · hace 3 días
Desde EuropaMachine LearningData Analytics Engineer
Ruby Labs · hace 3 días
Desde EuropaSQLData Scientist (AI Data & LLM Specialist)
Eclipse Laboratories · hace 4 días
Desde cualquier lugarPythonMachine LearningSenior Data Analyst
Ruby Labs · hace 4 días
Desde Europa