Archived listing. This role was posted over 30 days ago and is no longer accepting applications. We keep it for reference, but the employer may have already filled it. See today's verified listings.

AWS Data Lake & Pipeline Architecture

Toptal Anywhere in the World

See live roles like this Save · sign in Alert me to jobs like this

This one is closed

Live roles like this one

See every "AWS Data Lake" role →

Posted 15 Sep 2026
Last seen 15 Sep 2026
Location Anywhere in the World
Lifecycle mature
Grade D
This listing scored 52/100, which is a D. It lost the most ground on pay transparency. See the breakdown

Every figure above is arithmetic over the posting itself — its salary field, its text, its age, its tags and how many sources carry it. How the grades work →

This listing does not state a salary

$126k – $163k

That is the middle half of what comparable roles paid on this board over the last 90 days — 39 listings that did publish a figure, median $140k. It is not this employer's offer, and we have no idea what they pay. It is only what the rest of the market advertised.

What this role involves →

Headquarters:

Summary
We are seeking a Data Engineer to support the development of a Data Intelligence Platform. This role focuses on data modeling, data services, data pipelines, and cloud-based data infrastructure for reporting, analytics, and data science.

General information
This Data Engineer will support data-related operations across a broader Data Intelligence Platform team. The work includes building and consuming web services, integrating search technologies, supporting personalization and recommendation engines, and applying data engineering best practices across the platform.

The role spans cloud-based data architecture, database performance, pipeline automation, and support for data scientists, researchers, and internal business units. The environment includes AWS-based infrastructure and a mix of structured and unstructured data sources.

Tasks and Deliverables
- Participate in the architecture design and implementation of high-performance, scalable, and optimized data solutions.
- Create data models from scratch using strong SQL fundamentals.
- Write and optimize in-application SQL statements.
- Ensure the performance, security, and availability of databases.
- Prepare documentation and technical specifications.
- Handle database procedures such as upgrades, backups, recovery, and migration.
- Profile server resource usage and optimize configurations as necessary.
- Design, build, and automate the deployment of data pipelines and applications.
- Integrate data from on-premise databases and external data sources using REST APIs and harvesting tools.
- Collaborate with business units and data science teams on data access, transformation, processing, and reporting needs.
- Support implementation, technical issues, and training related to the data lake ecosystem.
- Work with the team to manage AWS resources, including EMR and ECS clusters.
- Support provisioning, monitoring, configuration, and maintenance of AWS tools.
- Evaluate and promote new cloud technologies that improve capabilities and lower operating costs.
- Support automation efforts using Infrastructure as Code with Terraform and CI/CD tools such as Jenkins.
- Work with the team to implement data governance, access control, and security risk reduction.

Required experience
- 7-9 years of experience designing and developing cloud-based data models, ETL pipelines, and infrastructure.
- Experience working with both structured and unstructured data.
- Strong proficiency with SQL across popular databases.
- Experience optimizing large, complex SQL statements.
- Knowledge of best practices for relational databases.
- Experience configuring database engines and orchestrating clusters.
- Ability to plan resource requirements from high-level specifications.
- Ability to troubleshoot common database issues.
- Experience with Spark, Glue, EMR, and Apache Kafka or AWS Kinesis.
- Experience with version control tools such as Git or Subversion.
- Experience using automated build systems and CI/CD workflows.
- Experience programming in Java, Python, and Scala.
- Knowledge of data structures and algorithms.
- Knowledge of relational, NoSQL, graph, document, key-value, and time-series databases.
- Knowledge of scalable data model design and management.
- Knowledge of ML model deployment.
- Knowledge of AWS cloud platforms.
- Knowledge of TDD and BDD.
- Strong interest in improving software development skills, frameworks, and technologies.

Engagement highlights
- Opportunity to work across data modeling, data services, and data science within a broader Data Intelligence Platform.
- Exposure to a varied technical environment spanning AWS, ETL pipelines, databases, search technologies, and recommendation systems.
- Direct collaboration with data science teams and internal stakeholders on reporting, transformation, and platform capabilities.

To apply: https://weworkremotely.com/remote-jobs/toptal-senior-data-engineer-aws-data-lake-pipeline-architecture

Apply for this role Opens weworkremotely.com — the link as listed; we have not yet verified it is the employer's own page

Quick question · anonymous · one tap

Would you apply to this job?

Answer to see what other job seekers said.

Description review

48/100 HR standards 89/100 Title ↔ description not measured Fit to the official role
Read the marked-up listing →

Your turn · no account needed

Help the next applicant

You may know something about this listing that we cannot see from here. One tap. No account needed. Signed-in reports earn points once the evidence agrees with you.

I know what it pays

What you were offered, quoted in an interview, or paid in this role. A range is fine.

Sign in with Google to earn points for reports — 100 confirmed points buy a week of Early Access.

Where this listing came from

  1. 15 Sep 2026 We Work Remotely first sighting

Seen on 1 board over 7 days.