Job Description
Roles & Responsibilities
This is a hands-on ownership role for someone who wants direct responsibility for production infrastructure, reliability, security operations and technology controls within a regulated environment.
The Role As Lead SRE / Technology Operations, you will own the reliability and operational readiness of the production technology environment. You will be the person responsible when production breaks and for ensuring there is a clear, evidenced process for understanding what happened and preventing recurrence.
Key responsibilities include:
- Own production infrastructure, monitoring, observability and access controls
- Manage incident response, RCA, runbooks and escalation procedures
- Own business continuity, backup/restore and disaster recovery testing
- Manage cloud infrastructure, Kubernetes and Infrastructure as Code
- Design and maintain IAM/PAM, least-privilege access and periodic access reviews
- Own CI/CD pipelines and deployment approval controls
- Coordinate security operations with outsourced CISO, SOC and managed security providers
- Coordinate penetration testing, vulnerability remediation and retesting
- Manage secrets, certificates and infrastructure security hardening
- Maintain operational and technical control evidence required for regulatory and audit purposes
- Support wallet/custody operational controls and reconciliation processes where applicable
This Role Will Suit Someone Who Has
- Personally owned production infrastructure rather than simply managing a technical team
- Is comfortable being hands-on when incidents occur
- Can balance reliability and delivery with security and regulatory requirements
- Understands that documentation and control evidence are part of operating a regulated platform
- Can work independently within a small, senior technology function
- Is comfortable working directly with the CTO, auditors, security partners and external vendors
This is not a purely managerial position. We are specifically looking for someone who has remained technically hands-on and has personally owned production incidents, infrastructure and operational controls.
Desired Candidate Profile
- 6+ years of hands-on experience across SRE, DevOps, DevSecOps, platform or infrastructure engineering
- 2+ years working within fintech, banking, exchange, brokerage or another genuinely regulated/audited environment
- Strong production experience with AWS or equivalent cloud platforms
- Hands-on Kubernetes experience
- Strong Infrastructure as Code experience, particularly Terraform, Ansible or equivalent
- Solid IAM/PAM and access-control experience
- Experience building and managing CI/CD pipelines
- Strong incident management, monitoring and observability experience
- Practical experience with backup/restore and disaster recovery, including RTO/RPO and failover testing
- Experience working with an outsourced CISO, SOC or managed security provider
- Experience managing vulnerabilities, penetration-test remediation and security findings
- Comfortable owning audit evidence and working within formal control environments
- Experience with crypto, digital assets, custody, wallet operations, Fireblocks, BitGo, blockchain infrastructure or KYT would be highly advantageous. However, this is not a hard requirement.
- Strong SRE/DevSecOps professionals coming from regulated banking, fintech, brokerage or other audited environments will also be considered.
- Personally owned production infrastructure rather than simply managing a technical team
- Is comfortable being hands-on when incidents occur
- Can balance reliability and delivery with security and regulatory requirements
- Understands that documentation and control evidence are part of operating a regulated platform
- Can work independently within a small, senior technology function
- Is comfortable working directly with the CTO, auditors, security partners and external vendors
Company Industry
Department / Functional Area
Incase you would like to apply to this job directly from the source, please click here