Job Requisition ID: 26WD99916
Position Overview
As a Senior DevOps Developer on the GCPay team, you will lead the evolution of our cloud‑native platform by advancing CI/CD, release engineering, infrastructure as code (IaC), system reliability, and security operations. You will support and enhance a core technology stack built on Java, MySQL, ElasticSearch, and AWS cloud technologies, enabling secure and reliable delivery across web applications, ERP integrations, APIs, and internal services.
This role requires a hands‑on developer with broad expertise across development and operations, including automation, infrastructure management, observability, and modern DevOps toolchains. You will play a critical role in strengthening platform resilience, improving security posture, increasing delivery velocity, and championing DevOps best practices across the organization.
Responsibilities
- Architect and implement scalable hosting solutions for dynamic SaaS web applications, ensuring reliability, performance, and resilience at scale.
- Lead the design, implementation, and evolution of secure, scalable, and reliable Infrastructure‑as‑Code solutions across GCPay environments.
- Establish, document, and drive adoption of engineering standards, operational processes, and DevOps best practices that improve consistency, maintainability, and delivery efficiency.
- Own and continuously strengthen security and compliance practices across infrastructure and applications, including hardening, access controls, and least‑privilege enforcement.
- Use modern tools such as Docker, Terraform, AWS cloud technologies, and related automation frameworks to provision, manage, and deploy infrastructure and services.
- Partner closely with development and testing teams throughout the development lifecycle to ensure quality, operational readiness, and successful delivery.
- Drive automation across build, deployment, infrastructure, and operational workflows, and evaluate new technologies that can improve reliability, scalability, security, or developer productivity.
- Define, monitor, and improve Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets to ensure reliability goals are measurable, understood, and achieved.
- Provide technical leadership in aligning platform strategy, infrastructure investments, and operational priorities with business and product requirements.
- Participate in on‑call rotation and lead incident response efforts, ensuring timely resolution, strong cross‑functional coordination, and clear communication during production events.
- Lead blameless post‑incident reviews to identify root causes, drive corrective and preventive actions, and improve overall platform reliability and team effectiveness.
- Contribute hands‑on coding, scripting, and automation to improve tooling, deployment workflows, infrastructure management, and operational efficiency.
- Own the operational health, scalability, and security posture of GCPay’s core platform, including systems built on Java, MySQL, ElasticSearch, and AWS cloud technologies.
- Serve as a senior technical resource and trusted partner across teams, using sound engineering judgment and business context to solve complex platform and operational challenges.
- Demonstrate strong ownership, a self‑starter mindset, and a commitment to continuous improvement, learning, and engineering excellence.
- Solve complex problems that require in‑depth evaluation of variable factors by taking a broad perspective to identify the best approach and innovative solutions.
- Effectively communicate technical challenges within and across teams.
- Keep yourself up‑to‑date with evolving technologies and showcase them with an implementation.
Minimum Qualifications
- 8+ years of experience in DevOps, Site Reliability Engineering, or related roles supporting cloud‑based applications.
- Advanced hands‑on experience with Linux administration, including system monitoring, troubleshooting, reliability, performance, and security.
- Strong experience managing large‑scale cloud infrastructure, with deep expertise in AWS preferred.
- Proven track record of architecting and operating large‑scale, highly available, and resilient systems.
- Strong scripting and automation skills using languages such as Bash, Python, or similar.
- Expert‑level knowledge of AWS services such as EC2, ECS, EKS, Lambda, ELB, S3, IAM, VPC, DynamoDB, and RDS.
- Hands‑on experience with Docker, Kubernetes, and modern container orchestration technologies.
- Proficiency with infrastructure‑as‑code tools such as Terraform and CloudFormation.
- Strong experience with CI/CD tools and software delivery workflows, including Git, Jenkins, Artifactory, or similar technologies.
- Experience with observability, log analysis, and monitoring tools such as CloudWatch, Splunk, Dynatrace, etc.
- Solid experience with relational databases such as MySQL or PostgreSQL, including strong SQL skills.
- Experience supporting SOC 2, ISO 27001, or similar compliance and audit requirements.
- Strong written and verbal communication skills, with the ability to communicate effectively across technical and non‑technical audiences.
- Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
Preferred Qualifications
- Interest or experience with incorporating AI Tools/Agents into DevOps/SRE practices to improve quality of work, ease of development and deployment.
- Relevant experience with payment technologies, fintech, or banking.
- Experience driving reliability initiatives, operational excellence, and service maturity in DevOps or SRE environments.
- Demonstrated ability to influence cross‑functional teams, provide technical leadership, and help set platform or operational direction.
- Excellent problem‑solving skills, with the ability to work independently and navigate complex technical challenges.
#J-18808-Ljbffr