Company Description
Trade W is a leading multi-asset trading platform with over seven years of industry experience, providing global users with secure, convenient, and efficient access to the financial markets. We offer CFD trading across a wide range of asset classes — including forex, cryptocurrencies, stocks, indices, metals, and commodities — through our intuitive app and web platform.
Launched in 2018 as the flagship brand of Tradewill Global LLC, Trade W was built on a customer-first philosophy and a vision to make trading success more accessible. Today, we continue to grow as a trusted platform, committed to empowering traders worldwide with equal opportunities for success.
Location: Kuala Lumpur Work Type: Full-Time | Onsite
About the Role
We are seeking an experienced DevOps / SRE Engineer (AWS & GCP) to support and scale our high-availability, multi-cloud infrastructure across global markets. This role plays a critical part in ensuring system reliability, performance, security, and automation excellence.
You will work closely with Engineering, Security, and Product teams to design resilient systems, drive automation, and support a mission-critical fintech platform operating at global scale.
Key Responsibilities
Production Operations & System Stability
- Manage, monitor, and optimize large-scale production environments across AWS, GCP, Linux/Windows systems, and middleware.
- Perform proactive system health checks, performance tuning, capacity planning, and version upgrades.
- Lead incident response, root cause analysis, and service recovery during system outages.
Multi-Cloud Architecture & SRE Practices
- Design, deploy, and maintain multi-cloud and multi-cluster SRE frameworks supporting distributed global workloads.
- Build and enhance observability systems, including metrics, logging, tracing, automated alerts, and incident classification.
Automation & Infrastructure as Code
- Develop automation scripts, internal tools, and operational frameworks to reduce manual workload and improve efficiency.
- Drive Infrastructure as Code (IaC) adoption using Terraform, Ansible, or equivalent tools.
Security, Compliance & Business Continuity
- Oversee data backups, disaster recovery (DR), log auditing, and cloud security enforcement.
- Collaborate closely with Security teams to ensure compliance with internal policies and regional regulatory standards.
Incident Response & On-Call Support
- Participate in 24/7 on-call rotations to support high-availability systems.
- Improve emergency response processes, escalation workflows, and on-call readiness.
Continuous Improvement
- Conduct post-incident reviews, documentation, and preventive improvement initiatives.
- Partner with Engineering and Product teams to optimize CI/CD pipelines and deployment workflows.
Requirements
- Bachelor's degree in Computer Science, Information Systems, Cybersecurity, or a related field.
- 5+ years of hands-on experience in DevOps, SRE, or Cloud Operations supporting large-scale systems.
- Strong expertise in AWS and GCP multi-cloud architecture, deployment, scaling, and operations.
- Solid understanding of SRE principles, CI/CD pipelines, DevOps / DevSecOps practices, and IaC.
- Strong knowledge of Linux systems, networking fundamentals (TCP/IP, routing, load balancing), and troubleshooting.
- Excellent communication skills with the ability to collaborate cross-functionally.
- Comfortable operating in high-pressure environments and participating in on-call rotations.
- Proficient in Mandarin and English for regional collaboration.
- Certifications such as CKA, CKS, CNCF, AWS/GCP Professional are strong advantages.
Why Join Trade W
- Work with international engineering teams on large-scale, multi-region infrastructure projects.
- Grow your expertise in a fast-paced, innovation-driven fintech environment with structured career pathways.
- Enjoy a competitive compensation package, including quarterly performance bonuses, comprehensive benefits, and additional perks upon confirmation.
- Be part of a collaborative, learning-focused culture that values engineering excellence and continuous improvement.