Data Engineer (m/w/d) at LinkedIn
Berlin, Berlin, Germany -
Full Time


Start Date

Immediate

Expiry Date

21 Dec, 26

Salary

60000.0

Posted On

22 Sep, 26

Experience

2 year(s) or above

Remote Job

Yes

Telecommute

Yes

Sponsor Visa

No

Skills

Industry

Information Technology & Services

Description

About this job

This position is suitable if...

Do you see data as more than just zeros and ones, but rather as the foundation for sound decisions?

Do you enjoy transforming chaotic legacy systems into clean data structures?

Do you think in pipelines, not scripts?

Are error handling and retry logic second nature to you, not optional extras?

Do you want to learn from experienced engineers while simultaneously implementing your own ideas?

Then you've come to the right place!

The tasks

As a Data Engineer, you will build the data pipelines that enable our digital transformation. You will be responsible for the clean, reliable, and scalable flow of data between hundreds of ERP systems and our modern data architecture.


Data Pipeline Engineering: You will design and implement robust ETL/ELT pipelines that reliably extract, transform, and load data from hundreds of ERP systems. You will utilize modern orchestration tools and ensure that everything runs smoothly, even with complex data flows. You will build pipelines that are not only functional but also maintainable and extensible.


Data Lake & Data Warehouse Architecture: You will help build our data lake infrastructure on AWS from the ground up. This includes structuring data effectively, implementing data governance concepts, and ensuring that data is optimally prepared for analytics and business intelligence. You understand the difference between raw, staging, and curated data and implement the corresponding architectures.


Greenfield Project – Legacy System Integration: You'll be right in the middle of an exciting greenfield project: the integration of several hundred legacy ERP systems. This is no standard integration – it involves complex data structures, different formats, and the challenge of creating a consistent data model from heterogeneous sources.


Workflow Orchestration & Reliability: You'll implement resilient workflows for time-critical data processing. You'll utilize modern orchestration patterns and ensure that pipelines continue to run or restart smoothly, even in the event of errors. You'll build monitoring and alerting systems so that problems are detected before they lead to actual outages.


Collaboration & Development: You'll work closely with our Staff Engineer, the Product Owners, and the approximately 10-person development team. You'll learn from experienced colleagues while contributing your own ideas. You'll grow with the challenges and continuously develop your skills.

Data Quality & Documentation: You implement data quality checks and ensure that poor-quality data never enters our systems. You document data flows and transformation logic so that others understand what happens to the data.

The profile

Must-haves:

  • At least 3 years of experience as a Data Engineer or in comparable roles
  • Practical experience with ETL/ELT pipelines and data lakes
  • Good Python skills for data processing and pipeline development
  • Experience with cloud platforms, ideally AWS (S3, Glue, Athena, RDS, etc.)
  • Basic knowledge of Infrastructure as Code (Terraform or CDK)
  • Understanding of data modeling and solid SQL knowledge
  • Experience with Git and CI/CD

Nice-to-haves:

  • Experience with Temporal or similar workflow orchestration frameworks (Airflow, Prefect, Dagster)
  • Knowledge of integrating legacy ERP systems
  • Experience with streaming technologies (Kafka, Kinesis)
  • Container knowledge (Docker, ECS)
  • Monitoring mit Grafana
  • Interest in healthcare or e-commerce
  • AWS Certifications

Why us?

What we offer:

  • Hybrid model: 2 days per week in the office in Cologne, 3 days flexible working from wherever you want
  • Greenfield project: You'll help build our modern data architecture from the ground up – no legacy code you have to understand first.
  • A real challenge: Integrating several hundred ERP systems – complex, but also extremely instructive.
  • Moderner Tech-Stack: AWS, Python, Terraform, moderne Data Engineering Tools
  • Mentoring & Growth: You work alongside experienced engineers and develop your skills.
  • Scope for creativity: You contribute your ideas and help shape our data infrastructure.
  • Growing team: You will become part of a motivated development team of approximately 10 people.
  • Impact: Your pipelines ensure that people have easier access to essential resources.

How To Apply:

Incase you would like to apply to this job directly from the source, please click here

Responsibilities

About this job

This position is suitable if...

Do you see data as more than just zeros and ones, but rather as the foundation for sound decisions?

Do you enjoy transforming chaotic legacy systems into clean data structures?

Do you think in pipelines, not scripts?

Are error handling and retry logic second nature to you, not optional extras?

Do you want to learn from experienced engineers while simultaneously implementing your own ideas?

Then you've come to the right place!

The tasks

As a Data Engineer, you will build the data pipelines that enable our digital transformation. You will be responsible for the clean, reliable, and scalable flow of data between hundreds of ERP systems and our modern data architecture.


Data Pipeline Engineering: You will design and implement robust ETL/ELT pipelines that reliably extract, transform, and load data from hundreds of ERP systems. You will utilize modern orchestration tools and ensure that everything runs smoothly, even with complex data flows. You will build pipelines that are not only functional but also maintainable and extensible.


Data Lake & Data Warehouse Architecture: You will help build our data lake infrastructure on AWS from the ground up. This includes structuring data effectively, implementing data governance concepts, and ensuring that data is optimally prepared for analytics and business intelligence. You understand the difference between raw, staging, and curated data and implement the corresponding architectures.


Greenfield Project – Legacy System Integration: You'll be right in the middle of an exciting greenfield project: the integration of several hundred legacy ERP systems. This is no standard integration – it involves complex data structures, different formats, and the challenge of creating a consistent data model from heterogeneous sources.


Workflow Orchestration & Reliability: You'll implement resilient workflows for time-critical data processing. You'll utilize modern orchestration patterns and ensure that pipelines continue to run or restart smoothly, even in the event of errors. You'll build monitoring and alerting systems so that problems are detected before they lead to actual outages.


Collaboration & Development: You'll work closely with our Staff Engineer, the Product Owners, and the approximately 10-person development team. You'll learn from experienced colleagues while contributing your own ideas. You'll grow with the challenges and continuously develop your skills.

Data Quality & Documentation: You implement data quality checks and ensure that poor-quality data never enters our systems. You document data flows and transformation logic so that others understand what happens to the data.

The profile

Must-haves:

  • At least 3 years of experience as a Data Engineer or in comparable roles
  • Practical experience with ETL/ELT pipelines and data lakes
  • Good Python skills for data processing and pipeline development
  • Experience with cloud platforms, ideally AWS (S3, Glue, Athena, RDS, etc.)
  • Basic knowledge of Infrastructure as Code (Terraform or CDK)
  • Understanding of data modeling and solid SQL knowledge
  • Experience with Git and CI/CD

Nice-to-haves:

  • Experience with Temporal or similar workflow orchestration frameworks (Airflow, Prefect, Dagster)
  • Knowledge of integrating legacy ERP systems
  • Experience with streaming technologies (Kafka, Kinesis)
  • Container knowledge (Docker, ECS)
  • Monitoring mit Grafana
  • Interest in healthcare or e-commerce
  • AWS Certifications

Why us?

What we offer:

  • Hybrid model: 2 days per week in the office in Cologne, 3 days flexible working from wherever you want
  • Greenfield project: You'll help build our modern data architecture from the ground up – no legacy code you have to understand first.
  • A real challenge: Integrating several hundred ERP systems – complex, but also extremely instructive.
  • Moderner Tech-Stack: AWS, Python, Terraform, moderne Data Engineering Tools
  • Mentoring & Growth: You work alongside experienced engineers and develop your skills.
  • Scope for creativity: You contribute your ideas and help shape our data infrastructure.
  • Growing team: You will become part of a motivated development team of approximately 10 people.
  • Impact: Your pipelines ensure that people have easier access to essential resources.

Loading...