IT 与编程
Data Engineer (Spark)(远程职位)
Addepto
链接会跳转到原始招聘信息。Donator 不代收申请。
Telegram 上的最新职位
每一条新的远程职位,发布即推送。
@donator_jobs_zh
Data Engineer (Spark):「IT 与编程」类别,完全远程
这条招聘信息来自 Addepto,岗位是 Data Engineer (Spark)。职位属于「IT 与编程」类别,完全远程。公司不限制候选人的居住地,在任何地方都可以申请。
招聘信息单独点名了 Python,也就是说,正是这一项工具的使用经验会成为决定因素。
公司没有公布具体数字,这一项会在面试时谈定。对工作时间,招聘信息没有提出任何条件。
这个职位通过了自动核查:凡是要求外国工作许可、签证担保、特定国籍,或者必须居住在指定国家的招聘信息,都不会进入列表。
工具
要点
- 公司
- Addepto
- 类别
- IT 与编程
- 谁可以申请
- 来自世界任何国家
- 工作方式
- 完全远程
- 工具
- Python
- 用工形式
- 全职
- 发布时间
- 2026年9月22日 (今天)
- 有效期
- 截至 2026年11月21日
- 来源
- Himalayas
新职位邮件提醒:IT 与编程
职位板每天更新数次。只有出现新的、已核查的职位时才会写信。不发垃圾邮件,也不需要注册账号。
已经订阅过了? 管理订阅设置
公司发布的职位描述
Addepto is a leading AI consulting () and data engineering () company that builds scalable, ROI-focused AI solutions for some of the world's largest enterprises and pioneering startups, including Rolls Royce, Continental, Porsche, ABB, and WGU. With an exclusive focus on Artificial Intelligence and Big Data, Addepto helps organizations unlock the full potential of their data through systems designed for measurable business impact and long-term growth. The company's work extends beyond client engagements. Drawing from real-world challenges and insights, Addepto has developed its own product - ContextClue - and actively contributes open-source solutions to the AI community. This commitment to transforming practical experience into scalable innovation has earned Addepto recognition by Forbes as one of the top 10 AI consulting companies worldwide. As part of KMS Technology, a US-based global technology group, Addepto combines deep AI specialization with enterprise-scale delivery capabilities-enabling the partnership to move clients from AI experimentation to production impact, securely and at scale. As a Data Engineer , you will have the exciting opportunity to work with a team of technology experts on challenging projects across various industries, leveraging cutting-edge technologies. Here are some of the projects we are seeking talented individuals to join:
* Development and maintenance of a large platform for processing automotive data. A significant amount of data is processed in both streaming and batch modes. The technology stack includes Spark, Cloudera, Airflow, Iceberg, Python, and AWS.
* Design and development of a universal data platform for global aerospace companies. This Azure and Databricks powered initiative combines diverse enterprise and public data sources. The data platform is at the early stages of the development, covering design of architecture and processes as well as giving freedom for technology selection.
* Centralized reporting platform for a growing US telecommunications company. This project involves implementing BigQuery and Looker as the central platform for data reporting. It focuses on centralizing data, integrating various CRMs, and building executive reporting solutions to support decision-making and business growth.
🚀 Your main responsibilities:
* Develop and maintain a high-performance data processing platform for automotive data, ensuring scalability and reliability.
* Design and implement data pipelines that process large volumes of data in both streaming and batch modes.
* Optimize data workflows to ensure efficient data ingestion, processing, and storage using technologies such as Spark, Cloudera, and Airflow.
* Work with data lake technologies (e.g., Iceberg) to manage structured and unstructured data efficiently.
* Collaborate with cross-functional teams to understand data requirements and ensure seamless integration of data sources.
* Monitor and troubleshoot the platform, ensuring high availability, performance, and accuracy of data processing.
* Leverage cloud services (AWS) for infrastructure management and scaling of processing workloads.
* Write and maintain high-quality Python (or Java/Scala) code for data processing tasks and automation.
Requirements 🎯 What you'll need to succeed in this role:
* At least 4 years of commercial experience implementing, developing, or maintaining Big Data systems, data governance and data management processes.
* Strong programming skills in Python (or Java/Scala): writing a clean code, OOP design.
* Hands-on with Big Data technologies like Spark , Cloudera, Kafka, Data Platform, Airflow, NiFi, Docker, and Iceberg.
* Excellent understanding of dimensional data and data modeling techniques.
* Experience implementing and deploying solutions in cloud environments.
* Consulting experience with excellent communication and client management skills, including prior experience directly interacting with clients as a consultant.
* Ability to work independently and take ownership of project deliverables.
* Fluent English (at least C1 level).
* Bachelor’s degree in technical or mathematical studies.
➕ Nice to have:
* Experience with an MLOps framework such as Kubeflow or MLFlow.
* Familiarity with Databricks and/or dbt.
🎁 Discover our perks & benefits:
* Work in a supportive team of passionate enthusiasts of AI & Big Data.
* Engage with top-tier global enterprises and cutting-edge startups on international projects.
* Enjoy flexible work arrangements, allowing you to work remotely or from modern offices and coworking spaces.
* Accelerate your professional growth through career development paths, knowledge-sharing initiatives, language classes, and sponsored training and conferences . Benefit from partnerships with Databricks and Anthropic , which provide access to industry-leading training materials and certification programs.
* Participate in team-building events and utilize the integration budget .
* Celebrate work anniversaries, birthdays, and milestones.
* Access medical and sports packages , eye care, and well-being support services, including psychotherapy and coaching.
* Get full work equipment for optimal productivity, including a laptop and other necessary devices.
* With our backing, you can boost your personal brand by speaking at conferences, writing for our blog, or participating in meetups.
* Experience a smooth onboarding with a dedicated buddy, and start your journey in our friendly, supportive, and autonomous culture.
Originally posted on Himalayas
职位正文保留公司发布时的原文,因为投递的时候用的也是同一种语言。
关于这个职位的常见问题
我可以在自己居住的地方申请「Data Engineer (Spark)」这个职位吗?
可以。对于这个职位,Addepto 接受来自世界任何国家的候选人,所以你不需要其他国家的工作许可。这条招聘信息通过了自动核查:如果雇主要求工作许可、签证担保,或者必须居住在某个特定国家,它就不会出现在本站。
标明的报酬是多少?
对于这个职位,Addepto 没有公布报酬。多数远程招聘信息不给出具体数字,这件事会在面试时谈定。
怎么申请?
你通过发布在 Himalayas 上的原始招聘信息,直接向雇主提交申请。Donator 不代收申请,不收取佣金,也不保存简历。
这是什么类型的职位?
这是「IT 与编程」类别中的一个完全远程职位。混合办公的招聘信息,以及任何需要到办公室的职位,本站都不会发布。
适合这个职位的免费证书
这条招聘信息要求 Python。下面是正好覆盖这几项工具的免费证书。
CS50x: Introduction to Computer Science
certificateHarvard CS50 · ~100 h
Harvard's own branded certificate is free once you pass every problem set and the final project. The edX verified certificate is a separate paid product and is not needed.
strong brandCS50P: Programming with Python
certificateHarvard CS50 · ~60 h
Same mechanism as CS50x: at least 70% on every assignment plus a final project. The certificate does not expire.
strong brandPython Essentials 1 and 2
digital badgeCisco Networking Academy · ~60 h
A two-part course with a separate badge for each part. It also doubles as preparation for OpenEDG's paid PCEP exam.
strong brandDeveloper certificates (10+ tracks)
certificatefreeCodeCamp · ~300 h
A non-profit; a card is never requested. The certificate is issued after five projects are submitted and pass.
moderate weight
