IT 与编程
Senior Data Engineer - AWS Data Lake & Pipeline Architecture(远程职位)
Toptal
链接会跳转到原始招聘信息。Donator 不代收申请。
Telegram 上的最新职位
每一条新的远程职位,发布即推送。
@donator_jobs_zh
Senior Data Engineer - AWS Data Lake & Pipeline Architecture:「IT 与编程」类别,完全远程
这条招聘信息来自 Toptal,岗位是 Senior Data Engineer - AWS Data Lake & Pipeline Architecture:这是高级职位,公司期待由你自己拿主意。职位属于「IT 与编程」类别,完全远程。公司不限制候选人的居住地,在任何地方都可以申请。
招聘信息里列出的主要工具是:AWS、SQL。简历上最好能给出正是这几项能力的具体例子。
公司没有公布具体数字,这一项会在面试时谈定。对工作时间,招聘信息没有提出任何条件。
这个职位通过了自动核查:凡是要求外国工作许可、签证担保、特定国籍,或者必须居住在指定国家的招聘信息,都不会进入列表。
要点
- 公司
- Toptal
- 类别
- IT 与编程
- 谁可以申请
- 来自世界任何国家
- 工作方式
- 完全远程
- 级别
- 高级
- 工具
- AWS, SQL
- 发布时间
- 2026年9月15日 (24 天前)
- 有效期
- 截至 2026年10月25日
- 来源
- We Work Remotely
新职位邮件提醒:IT 与编程
职位板每天更新数次。只有出现新的、已核查的职位时才会写信。不发垃圾邮件,也不需要注册账号。
已经订阅过了? 管理订阅设置
公司发布的职位描述
Headquarters:
Summary We are seeking a Data Engineer to support the development of a Data Intelligence Platform. This role focuses on data modeling, data services, data pipelines, and cloud-based data infrastructure for reporting, analytics, and data science. General information This Data Engineer will support data-related operations across a broader Data Intelligence Platform team. The work includes building and consuming web services, integrating search technologies, supporting personalization and recommendation engines, and applying data engineering best practices across the platform. The role spans cloud-based data architecture, database performance, pipeline automation, and support for data scientists, researchers, and internal business units. The environment includes AWS-based infrastructure and a mix of structured and unstructured data sources. Tasks and Deliverables - Participate in the architecture design and implementation of high-performance, scalable, and optimized data solutions. - Create data models from scratch using strong SQL fundamentals. - Write and optimize in-application SQL statements. - Ensure the performance, security, and availability of databases. - Prepare documentation and technical specifications. - Handle database procedures such as upgrades, backups, recovery, and migration. - Profile server resource usage and optimize configurations as necessary. - Design, build, and automate the deployment of data pipelines and applications. - Integrate data from on-premise databases and external data sources using REST APIs and harvesting tools. - Collaborate with business units and data science teams on data access, transformation, processing, and reporting needs. - Support implementation, technical issues, and training related to the data lake ecosystem. - Work with the team to manage AWS resources, including EMR and ECS clusters. - Support provisioning, monitoring, configuration, and maintenance of AWS tools. - Evaluate and promote new cloud technologies that improve capabilities and lower operating costs. - Support automation efforts using Infrastructure as Code with Terraform and CI/CD tools such as Jenkins. - Work with the team to implement data governance, access control, and security risk reduction. Required experience - 7-9 years of experience designing and developing cloud-based data models, ETL pipelines, and infrastructure. - Experience working with both structured and unstructured data. - Strong proficiency with SQL across popular databases. - Experience optimizing large, complex SQL statements. - Knowledge of best practices for relational databases. - Experience configuring database engines and orchestrating clusters. - Ability to plan resource requirements from high-level specifications. - Ability to troubleshoot common database issues. - Experience with Spark, Glue, EMR, and Apache Kafka or AWS Kinesis. - Experience with version control tools such as Git or Subversion. - Experience using automated build systems and CI/CD workflows. - Experience programming in Java, Python, and Scala. - Knowledge of data structures and algorithms. - Knowledge of relational, NoSQL, graph, document, key-value, and time-series databases. - Knowledge of scalable data model design and management. - Knowledge of ML model deployment. - Knowledge of AWS cloud platforms. - Knowledge of TDD and BDD. - Strong interest in improving software development skills, frameworks, and technologies. Engagement highlights - Opportunity to work across data modeling, data services, and data science within a broader Data Intelligence Platform. - Exposure to a varied technical environment spanning AWS, ETL pipelines, databases, search technologies, and recommendation systems. - Direct collaboration with data science teams and internal stakeholders on reporting, transformation, and platform capabilities.
To apply: https://weworkremotely.com/remote-jobs/toptal-senior-data-engineer-aws-data-lake-pipeline-architecture
职位正文保留公司发布时的原文,因为投递的时候用的也是同一种语言。
关于这个职位的常见问题
我可以在自己居住的地方申请「Senior Data Engineer - AWS Data Lake & Pipeline Architecture」这个职位吗?
可以。对于这个职位,Toptal 接受来自世界任何国家的候选人,所以你不需要其他国家的工作许可。这条招聘信息通过了自动核查:如果雇主要求工作许可、签证担保,或者必须居住在某个特定国家,它就不会出现在本站。
标明的报酬是多少?
对于这个职位,Toptal 没有公布报酬。多数远程招聘信息不给出具体数字,这件事会在面试时谈定。
怎么申请?
你通过发布在 We Work Remotely 上的原始招聘信息,直接向雇主提交申请。Donator 不代收申请,不收取佣金,也不保存简历。
这是什么类型的职位?
这是「IT 与编程」类别中的一个完全远程职位。混合办公的招聘信息,以及任何需要到办公室的职位,本站都不会发布。
适合这个职位的免费证书
这条招聘信息要求 AWS、SQL。下面是正好覆盖这几项工具的免费证书。
CS50x: Introduction to Computer Science
certificateHarvard CS50 · ~100 h
Harvard's own branded certificate is free once you pass every problem set and the final project. The edX verified certificate is a separate paid product and is not needed.
strong brandDeveloper certificates (10+ tracks)
certificatefreeCodeCamp · ~300 h
A non-profit; a card is never requested. The certificate is issued after five projects are submitted and pass.
moderate weightKaggle Learn: 17 micro-courses
certificateKaggle · ~4 h
Intro to ML, Pandas, SQL, Deep Learning, Computer Vision and more. Only a Google account is needed, and the certificate has a public link.
moderate weightAWS Educate: Introduction to Generative AI
digital badgeAWS · ~4 h
Email only: no card and no AWS account. The labs run in the browser and the minimum age is 13.
moderate weight
