Roles & Responsibilities
5 Work Days Per Week
Office at Tai Seng Exchange Tower B
Near Tai Seng MRT, Singapore
Insurance Coverage
Entitled to Yearly Bonus & Performance Bonus
About The Role
Focused on maintaining high availability and stability of application systems, ensuring compliance with privacy and data protection laws, managing changes and releases, and driving operational automation and infrastructure optimization.
High Availability and Stability Maintenance of Application Systems:
Includes daily monitoring, alert response, emergency handling, on-call duties, regular system health checks, and performance optimization.
Compliance and Secure Access Construction for Application Systems
Ensure operational design, processes, and data management comply with relevant privacy and data protection laws.
Ensure compliance with full auditing and regulatory checks and provide auditing materials as required.
Change and Release Management:
Best practices for application system changes, including change control, version management, and rollback strategies, while ensuring operational duties during release windows.
Automation and Infrastructure Optimization
Drive operational automation by designing and implementing automated tools and processes, ensuring resource allocation is optimized and supporting business scalability.
Other Operational Practices and Work Arrangements:
Provide feedback and suggestions for business architecture design and continuously produce operational technical documentation.
Qualifications
Bachelor's Degree or above; a degree in computer science or a related field is preferred.
Experiences as Senior SRE or leading a small team is preferable.
At least 5 years of experience in cloud services products and application operations.
Familiar with cloud platform deployment and management (AWS, Azure, Google Cloud, etc.).
Experience with automation operations and container technologies (Docker, Kubernetes).
Familiarity with CI / CD processes and tools (e.g Jenkins, GitLab CI).
Tell employers what skills you have
Stability Testing
RDS
Microsoft Azure
Jenkins
Kubernetes
Ubuntu
AWS
Field Support
Cloud Deployment
Team Leadership
Networking
Technical Consultation
Site Reliability Engineering
Routing Protocols
Google Cloud
Docker
GitLab
Network Security
C++
Container Operations
Senior Site Reliability Engineer / SRE Leader • D19 Hougang, Sengkang, Serangoon Garden, Punggol, SG