Develop and maintain Terraform and Ansible automation for provisioning, configuring, monitoring, scaling, and managing NoSQL, streaming, and caching platforms.
Build repeatable deployment automation across cloud and hybrid environments.
Improve the reliability, availability, scalability, performance, and resiliency of platform data services.
Define, measure, and improve service-level indicators, service-level objectives, and error budgets.
Automate scaling, failover, backup, recovery, upgrades, and routine maintenance.
Build observability capabilities using metrics, logging, tracing, dashboards, and alerts.
Troubleshoot Cassandra, Aerospike, Kafka/MSK, Redis, and related platform services.
Participate in on-call rotations, incident response, root-cause analysis, and permanent remediation.
Write reliable and tested Go code for infrastructure automation, platform services, and operational tooling.
Collaborate with engineering, platform, security, and operations teams; maintain documentation, runbooks, and automation playbooks.
Participate in code reviews, technical design discussions, and practical applications of AI-assisted automation and anomaly detection.
Requirements
Bachelor’s or master’s degree in computer science or a related field, or equivalent practical experience.
At least 3 years of experience in software engineering, database reliability engineering, site reliability engineering, platform engineering, or a related field.
Production Go development experience, including testing, concurrency, error handling, and maintainability.
Hands-on experience with Infrastructure as Code or configuration-management tools such as Terraform or Ansible.
Experience deploying or operating workloads on Kubernetes.
Experience with AWS or GCP and managed services such as MSK, DynamoDB, ElastiCache, or Memorystore.
Working knowledge of one or more NoSQL, caching, or streaming technologies such as Cassandra, Aerospike, Kafka, AWS MSK, or Redis.
Understanding of distributed systems concepts and working knowledge of Linux, networking, storage, and system troubleshooting.
Familiarity with observability practices including metrics, logging, tracing, alerting, and dashboard creation.
Ability to diagnose and resolve technical problems independently within a defined scope and collaborate effectively across teams.
Benefits
Medical, dental, and vision coverage; matching 401(k); paid time off; wellness program; and employee discounts on Sony products.
Potential eligibility for a bonus package.
Hybrid working policy applies, and base pay may vary based on job-related factors including location, knowledge, skills, and experience.
Background checks are conducted at the offer stage for new employees.
Salary: $150k - $225k/yr
PlayStation Global
At PlayStation Global, we focus on creating user-friendly solutions that simplify processes, improve efficiency, and empower teams to succeed.