HomeSearchAi Engineer Jobs › Senior Systems Engineer I

Senior Systems Engineer I

RELX

India-Bengaluru (Helios Business Park) · senior
More Ai Engineer jobs: Ai Engineer jobsAi Engineer salary

Are you a Senior Systems Engineer looking for an innovative role? Do you enjoy being part of a team that works with a diverse range of technology? About our Team: Elsevier Health applies innovation, facilitates insights, and helps drive more informed decision-making for our customers across global health. We support health providers by providing accessible, trusted evidence-based information; prepare more medical and nursing students with effective tools and resources; provide insights that help clinicians improve patient outcomes; and supports a more personalized and localized healthcare experience.

All for the benefit of every patient. About the Role:   We are seeking a Senior Systems Engineer to join our team. This role is responsible for designing, building, and operating secure, scalable, and highly available AWS infrastructure that enables multiple product engineering teams to deliver software reliably. The ideal candidate has strong expertise in Terraform, AWS, GitHub Actions, ECS Fargate, and infrastructure automation , along with a practical operational mindset.

You will partner closely with software engineers, troubleshoot production issues, improve developer experience, and continuously modernize our cloud platform while maintaining high standards for security, reliability, and operational excellence. Responsibilities:   Design, build, and maintain Infrastructure as Code using Terraform following modular, reusable, and scalable practices. Operate and support production AWS environments across multiple accounts and regions.

Develop, maintain, and troubleshoot GitHub Actions CI/CD pipelines supporting application and infrastructure deployments. Build and manage containerized workloads on Amazon ECS Fargate. Administer and support Amazon RDS PostgreSQL environments, ensuring availability, performance, backup, and recovery readiness. Design and maintain secure AWS networking and IAM architectures. Respond to production incidents, perform root cause analysis, and implement preventive improvements.

Build automation using Bash, Python, AWS CLI, and related tooling to improve operational efficiency. Partner with multiple development teams to enable self-service infrastructure while reducing operational bottlenecks. Review infrastructure changes for operational risk, security impact, and deployment safety. Create and maintain technical documentation, operational runbooks, and disaster recovery procedures.

Requirements:    Qualifications:  Bachelor’s degree in computer science, Engineering, Information Technology, or an equivalent discipline. Typically 6-9 years of experience in DevOps, Cloud Engineering, or Systems Engineering. Proven experience operating and supporting production AWS environments in enterprise-scale organizations. Technical Skills:    Must Have Advanced Terraform: Expertise in modules, providers, state management, lifecycle controls, drift detection, safe refactoring, and remote state (S3, locking, cross-stack dependencies).

AWS Operations: Hands-on experience managing production, multi-account, multi-region AWS environments across ECS, RDS, ALB, VPC, IAM, Route53, ECR, S3, Lambda, DynamoDB, SQS, Secrets Manager, KMS, and CloudWatch. GitHub Actions CI/CD: Experience building and troubleshooting reusable workflows, OIDC authentication, approval gates, runners, Terraform deployments, application deployments, and migration pipelines.

ECS Fargate & Containers: Strong knowledge of Docker, ECR, ECS task definitions/services, IAM roles, health checks, autoscaling, ALB integration, and deployment rollbacks. RDS PostgreSQL: Experience with upgrades, Multi-AZ, parameter groups, backups/restores, RDS Proxy, performance tuning, and database migration coordination. AWS Networking & Security: Proficiency in VPCs, networking, ALBs, Route53, ACM/TLS, IAM, OIDC, Secrets Manager, KMS, and cloud security best practices.

Incident Response & Observability: Skilled in troubleshooting using logs, metrics, al

Search all live jobs — free, no account →