Toptal

Senior Data Engineer — Azure, Databricks & ML Pipelines | Remote

Remote, United States remote Entry Salary not listed
remote Technology & IT Curated
Sign in to apply Free account — we bring you straight back to this role.

About the role

Headquarters: Remote
URL: https://www.toptal.com/

About the Role

We're looking for a Senior Data Engineer to design, build, and maintain scalable data pipelines and ML-ready infrastructure on Azure and Databricks. This is a hands-on engineering role: you'll own the full data pipeline lifecycle — ingestion, transformation, orchestration, and deployment — while supporting machine learning workflows with clean, reliable data. If you're comfortable owning infrastructure decisions and writing production-quality Python at scale, this role is built for that.

What You'll Do

Design, build, and maintain data pipelines using Databricks and Azure-native data services

Develop and optimize ETL/ELT processes to support analytics and machine learning workloads

Build and maintain CI/CD pipelines for data engineering and ML deployment workflows

Write clean, efficient, production-quality Python for data processing and pipeline automation

Support machine learning teams with well-structured, high-quality datasets and feature pipelines

Design and manage data architecture across Azure services (e.g., Azure Data Factory, Azure Data Lake, Azure Synapse)

Monitor pipeline performance, troubleshoot data quality issues, and implement reliability improvements

Implement data governance, security, and access control best practices

Collaborate with data scientists, analysts, and software engineers to align data infrastructure with business needs

Participate in code reviews, architecture discussions, and technical planning

What You Bring

Strong hands-on experience with Azure cloud data services

Proven experience building and maintaining pipelines on Databricks

Solid experience designing and managing CI/CD pipelines for data or ML workflows

Strong Python skills for data engineering and pipeline development

Working knowledge of machine learning workflows and how data engineering supports them

Experience with SQL and relational/distributed data systems

Understanding of data pipeline orchestration, monitoring, and reliability practices

Strong problem-solving skills and ability to work independently on complex data infrastructure challenges

Solid communication skills for collaborating with data science and engineering teams

Nice to Have

Experience with MLOps practices and tools (MLflow, Azure ML)

Familiarity with Spark internals and performance tuning within Databricks

Experience with infrastructure-as-code (Terraform, Bicep, ARM templates)

Exposure to real-time/streaming data pipelines (Kafka, Event Hubs, Structured Streaming)

Relevant Azure or Databricks certifications

Why This Role

Full pipeline ownership: Own data infrastructure end to end, from ingestion through ML-ready delivery

Modern data stack: Work with Azure and Databricks, leading platforms in enterprise data engineering

Cross-functional impact: Directly enable machine learning and analytics outcomes, not just move data

Flexibility: Remote-friendly engagement structure

How to Apply

Ready to bring your data engineering expertise to Azure and Databricks-powered ML infrastructure? Apply through Toptal here: https://www.toptal.com/talent/apply

 

To apply: https://weworkremotely.com/remote-jobs/toptal-senior-data-engineer-azure-databricks-ml-pipelines-remote

Interview prep

Walk in with sharper answers.

Use this as a quick practice sheet before you speak with the employer.

Role
Technology & IT Data Analysis MySQL Python Remote Collaboration Sales remote

Likely questions

  1. Tell us about work you have done that is close to the Senior Data Engineer — Azure, Databricks & ML Pipelines | Remote role.
  2. How would you approach your first 30 days at Toptal?
  3. Which of Data Analysis, MySQL and Python have you used recently, and what did it help you achieve?
  4. Describe a time you solved a problem without waiting to be told exactly what to do.
  5. How do you stay organised and communicate clearly when working remotely?

Prepare before the call

  • A recent example that proves your experience with Data Analysis, MySQL and Python.
  • One short story with a problem, your action, and the result.
  • Two examples that show the strengths listed on your CV.
  • A clear reason why this role and company interest you.
  • Your availability, preferred work style, and salary expectations.

Ask them

  • What would success look like in the first 90 days?
  • What are the main problems this hire should help solve?
  • How does the team give feedback and measure good work?
  • What does a normal working week look like for this role?
Practice line

I am interested in the Senior Data Engineer — Azure, Databricks & ML Pipelines | Remote role because I can bring practical experience in Data Analysis, MySQL and Python, learn the team quickly, and contribute to the outcomes Toptal needs from this hire.

Related jobs.

More roles from this company or category.