Site Reliability Engineer at InterEx Group
Berlin, Berlin, Germany -
Full Time


Start Date

Immediate

Expiry Date

24 Nov, 26

Salary

0.0

Posted On

26 Aug, 26

Experience

0 year(s) or above

Remote Job

Yes

Telecommute

Yes

Sponsor Visa

Yes

Skills

Industry

Information Technology & Services

Description


Responsibilities:

  • Create awareness in other teams about methods and procedures we use to help them to prevent repetitive help requests.
  • Help application developers to understand the infrastructure / cluster / system
  • “We are the team that is in charge of understanding & explaining how the system fits into the customer’s ecosystem”
  • Share knowledge / mindset to other teams (dev/infra engineers)
  • Cross functional, share knowledge between infra engineers
  • Contribute towards building VCP as a Product which meets our standards of quality
  • Increase stability and reliability of VCP by automated testing and automation
  • Customer satisfaction and product reliability
  • Improve the functionality and reliability of VCP
  • Translate customer ecosystem needs to engineering deliverables
  • Find the broken pieces of the puzzle at system/cluster level
  • Combination of individual ‘stories’ in a complete book
  • Make the VCP reliable by improving system resilience (bug-fixing and beyond)
  • Resolve bugs in a sustaining way (implement regression test, design structural fixes)
  • Ambassador of predictable component lifecycle management
  • Technical roadmap maintenance (App life cycle management)
  • Support feature and service request from the field
  • Suggest improvements to our technical solutions and way of working, and implement them in alignment with your team and their stakeholders


Highly valued qualifications & experiences:

  • Experience with DC/OS
  • Experience with new technology introduction @ zero downtime including data migration
  • Fan of automatic testing and qualification, if can be part of CI/CD pipeline.
  • Affinity to dig deep into the details of networking issues
  • Available to work (remotely) outside regular office hours when it proves that attempt to build a fail-safe system was not yet successful. We really want this to be an exception, not a rule.


Required qualifications & experiences:

  • Knowledge of distributed computing systems, practical experience (must!)
  • Experienced in build and release infrastructure, Maven, Nexus, Bamboo, Github
  • Familiar with at least one scripting language (Python)
  • Experience with Ansible
  • Linux expert

How To Apply:

Incase you would like to apply to this job directly from the source, please click here

Responsibilities
Loading...