Site Reliability Engineer at Vontier
Mumbai, maharashtra, India -
Full Time


Start Date

Immediate

Expiry Date

07 Sep, 26

Salary

0.0

Posted On

09 Jun, 26

Experience

2 year(s) or above

Remote Job

Yes

Telecommute

Yes

Sponsor Visa

No

Skills

AWS, Terraform, Kubernetes, CI/CD, Jenkins, GitHub Actions, PowerShell, Bash, Python, Ansible, New Relic, AppDynamics, DataDog, Linux, Windows, Networking

Industry

electrical;Appliances;and Electronics Manufacturing

Description
Job Summary  As a Site Reliability Engineer, you will play a critical role in ensuring the availability and performance of our customer-facing platform. You will work closely with DevOps, DBA, and Development teams to provision and maintain infrastructure, deploy and monitor our applications, and automate workflows. Your contributions will have a direct impact on customer satisfaction and overall user experience.    Responsibilities and Deliverables  * Manage, monitor, and maintain highly available systems (Windows and Linux)  * Analyze metrics and trends to ensure performance and rapid scalability.  * Address routine service requests while identifying ways to automate and simplify.  * Create infrastructure as code using Terraform, ARM Templates, Cloud Formation.  * Maintain data backups and disaster recovery plans.  * Adhere to security best practices through all stages of the software development lifecycle  * Follow and champion ITIL best practices and standards.     Organizational Alignment  * Reports to the Senior SRE Manager  * This role involves close collaboration with DevOps, DBA, and security teams.    Technical Proficiencies  * Hands-on experience with AWS is a must-have.  * Proficiency analyzing application, IIS, system, security logs, and CloudTrail events.  * Experience with CI/CD tools such as Jenkins and GitHub Actions  * Experience maintaining and administering Windows, Linux, and Kubernetes.  * Experience in automation using scripting languages such as PowerShell, Bash, or Python.  * Good understanding of networking concepts (VPC, subnet, private link, peering).  * Familiarity with configuration management using Ansible, Azure Automation or similar.  * Familiarity with observability tools such as New Relic, AppDynamics, or DataDog.    Experience  * 3+ years of experience in SRE or System Administration role.  * Demonstrated ability building and supporting high availability Windows/Linux servers.  * 2+ years of experience working with cloud technologies including AWS, Azure.  * Comfortable using Scrum, Kanban, or Lean methodologies.    Education  * Bachelor’s Degree or College Diploma in Computer Science, Information Systems, or equivalent experience. 
Responsibilities
Ensure the availability and performance of the customer-facing platform by managing highly available Windows and Linux systems. Automate workflows using infrastructure as code and maintain disaster recovery plans and security best practices.
Loading...