HomeSearchData Engineer Jobs › Software Engineer II Platform Data Reliability

Software Engineer II Platform Data Reliability

sonyinteractiveentertainmentglobal

United States, San Mateo, CA
More Data Engineer jobs: Data Engineer jobsData Engineer salary

Job at a glance

United States, San Mateo, CA
Location
sonyinteractiveentertainmentglobal
Employer
Data Engineer jobs
Category

Why Sony Interactive Entertainment? Sony Interactive Entertainment isn’t just the Best Place to Play — it’s also the Best Place to Work. Sony Interactive Entertainment (SIE) is the company behind the PlayStation brand. As a subsidiary of Sony Group Corporation, we’re part of a proud legacy of innovation and excellence. SIE is a dynamic technology company, delivering cutting-edge hardware and network services to more than 100 million people and an entertainment leader, home to some of the most beloved and recognizable intellectual properties (IP) in the world.

Our role at SIE is to create and nurture the experiences under the PlayStation brand, a name synonymous with entertainment excellence and creativity. Software Engineer II – Platform Data Reliability & Automation Ready to level up your career? Join PlayStation as a Software Engineer II focused on Platform Data Reliability and Automation and help build reliable, scalable experiences for millions of players around the globe.

At PlayStation, we’re known not only for delivering exceptional gaming experiences but also for fostering an engineering environment centered on innovation, creativity, collaboration, and technical excellence. We welcome passionate engineers who enjoy solving challenging problems and are excited about shaping the future of play. Role Overview We are seeking a Software Engineer II (Platform Data Reliability & Automation) to help build, automate, and operate scalable data platforms using Infrastructure as Code (IaC) and cloud technologies.

This role focuses on improving the reliability and automation of NoSQL, streaming, and caching services across AWS and GCP environments. You’ll develop automation, observability, and operational tooling supporting technologies such as Cassandra, Aerospike, Kafka, and Redis. Working alongside senior engineers, platform teams, and product teams, you’ll contribute to highly available infrastructure supporting billions of transactions and millions of players globally.

By applying software engineering and database reliability engineering principles, you’ll help reduce manual work, improve system uptime, and make data services easier and safer for engineering teams to use. Responsibilities Develop, maintain, and improve Infrastructure as Code and configuration-management automation using tools such as Terraform and Ansible to provision, configure, monitor, scale, and manage NoSQL, streaming, and caching platforms.

Build automation that enables repeatable and reliable deployment of data services across cloud and hybrid environments. Contribute to the reliability, availability, scalability, performance, and resiliency of platform data services. Contribute to defining, measuring, and improving service-level indicators, service-level objectives, and error budgets. Develop automation for operational activities such as scaling, failover, backup, recovery, upgrades, and routine maintenance.

Build and enhance observability solutions using metrics, logging, tracing, dashboards, and alerts. Troubleshoot issues affecting Cassandra, Aerospike, Kafka/MSK, Redis, and related platform services. Participate in on-call rotations and incident response, contributing to root-cause analysis and the implementation of permanent fixes. Write reliable, maintainable, and well-tested Go code for infrastructure automation, platform services, and operational tooling.

Collaborate with engineering, platform, security, and operations teams to integrate and deliver reliable data services. Create and maintain operational documentation, procedures, runbooks, and automation playbooks. Participate in code reviews, technical design discussions, and continuous improvement initiatives. Explore practical applications of AI-assisted automation, anomaly detection, automated remediation, and developer-productivity tooling where appropriate.

Skills and Qualifications Bachelor’s or Master’s degree in Computer Science or a related field, or equivalent practical ex

Search all live jobs — free, no account →