Hands-on Engineering - 60%
•Design and implement secure, scalable, resilient, and highly available cloud and hybrid environments, primarily on AWS.
•Own and evolve AWS architecture across networking, compute, storage, identity, security, containers, and managed services.
•Build and maintain automated CI/CD pipelines using GitLab CI, Argo CD, Jenkins, GitHub Actions, or similar tools.
•Implement GitOps practices and automate application deployment, configuration, testing, and release processes.
•Provision and manage infrastructure using Terraform and/or AWS CloudFormation, following reusable and version-controlled Infrastructure as Code practices.
•Design, deploy, and operate containerised workloads using Docker and Kubernetes, preferably Amazon EKS.
•Improve Kubernetes resilience through effective workload configuration, autoscaling, resource management, health checks, ingress, secrets management, and disaster-recovery practices.
•Develop automation and operational tooling using Python, Bash, or PowerShell.
•Implement comprehensive observability using Prometheus, Grafana, OpenTelemetry, centralised logging, distributed tracing, alerting, and service-level indicators.
•Embed security throughout the software delivery lifecycle, including identity and access management, secrets management, vulnerability scanning, policy enforcement, encryption, and supply-chain security.
•Troubleshoot complex infrastructure, deployment, performance, and production reliability issues.
•Participate in incident response, root-cause analysis, capacity planning, and continuous service improvement.
•Identify opportunities to improve reliability, security, developer productivity, performance, and cloud cost efficiency.
Leadership and Management - 40%
•Lead, mentor, and develop DevOps, cloud, and platform engineers through coaching, technical guidance, and regular feedback.
•Establish and maintain engineering standards for cloud architecture, automation, CI/CD, security, reliability, and operational readiness.
•Plan and prioritise the team's work in partnership with engineering, product, security, and business stakeholders.
•Provide technical direction while balancing delivery timelines, operational risk, scalability, security, and cost.
•Review architectural proposals, infrastructure code, automation, and operational processes.
•Promote a culture of ownership, collaboration, documentation, continuous improvement, and blameless learning.
•Coordinate incident response and ensure corrective actions are documented, prioritised, and completed.
•Support recruitment, onboarding, performance management, career development, and capability planning.
•Communicate technical risks, dependencies, progress, and recommendations clearly to technical and non-technical stakeholders.
•Champion DevSecOps, Site Reliability Engineering, and cloud governance practices across the organisation.
Person Specification
Essential
•At least 8+ years of experience in DevOps, cloud engineering, platform engineering, Site Reliability Engineering, or a related field.
•Previous experience leading, mentoring, or managing engineers.
•Strong, recent, hands-on expertise in AWS, including production cloud architecture and operations.
•Working knowledge of Microsoft Azure and/or GCP and an understanding of multi-cloud concepts.
•Strong experience with CI/CD platforms, preferably GitLab CI, plus familiarity with Argo CD, Jenkins, or GitHub Actions.
•Advanced experience with Infrastructure as Code, particularly Terraform or AWS CloudFormation.
•Strong practical experience with Docker, Kubernetes, and Amazon EKS in production environments.
•Proficiency in at least one scripting or programming language, such as Python, Bash, or PowerShell.
•Experience implementing monitoring, logging, alerting, and tracing using tools such as Prometheus, Grafana, OpenTelemetry, CloudWatch, or equivalent platforms.
•Strong understanding of cloud networking, DNS, load balancing, identity and access management, secrets management, and security controls.
•Experience designing highly available systems and implementing backup, disaster-recovery, and business-continuity strategies.
•Experience supporting ISO 27001 and PCI DSS compliance audits, with hands-on expertise in implementing CIS Benchmarks and supporting penetration testing and remediation activities.
•Strong troubleshooting, problem-solving, documentation, stakeholder-management, and communication skills.
Desired
•AWS certifications such as AWS Certified Solutions Architect, DevOps Engineer, or Security - Specialty.
•Kubernetes certification such as CKA, CKAD, or CKS.
•Experience operating hybrid or multi-cloud environments.
•Experience with service meshes, policy-as-code, FinOps, or cloud cost-optimisation practices.
•Familiarity with SRE practices, including SLIs, SLOs, error budgets, incident management, and post-incident reviews.
•Experience supporting regulated, security-sensitive, or high-availability environments.
ModInclusion
Modulr is a leading fintech platform providing Payments Automation and enterprise payment solutions. We serve high-growth businesses across multiple verticals including Lending, Travel, Banking, and Core Business Payments, operating across India, the UK, Europe, and beyond. Our mission is to simplify and automate payments for our customers, and we're building a diverse, talented team to support that mission. Our people are central to our success, and we're committed to creating an inclusive workplace where everyone can thrive.