{"id":1274972,"url":"https://alion.io/job/paynet-principal-engineer-platform-engineering","title":"Principal Engineer (Platform Engineering)","company":{"id":2112342,"name":"PayNet","domain":"paynet.my","url":"https://alion.io/company/paynet-my","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"BrioHR","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"lead","employment_type":null,"work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Malaysia"],"countries":["MY"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":20000,"max_usd":49000,"period":"year","method":"global_role_cell_scaled_by_country","sample_n":602},"experience_years_min":8,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Atlantis","optional":false},{"name":"AWS","optional":false},{"name":"CI/CD","optional":false},{"name":"GitLab CI","optional":false},{"name":"Helm","optional":false},{"name":"IAM","optional":false},{"name":"Kubernetes","optional":false},{"name":"Machine Learning","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Python","optional":false},{"name":"Terraform","optional":false}],"status":"live","first_seen_at":"2026-09-26T01:11:32Z","employer_posted_date":"2026-09-26","last_verified_at":"2026-09-29T16:48:13Z","board_verified":false,"closed_at":null,"days_open":6,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":6},"description":"Why PayNet / Why Now\nBuild the secure platform foundations behind PayNet’s fraud intelligence, data, microservices, and machine learning capabilities.\nShape a critical stage of the PayNet Secure Project as workloads, data volumes, and production demands continue to grow.\nWork where platform decisions must balance speed, resilience, security, regulatory expectations, and long-term sustainability.\nJoin a role with the mandate to make sound technical trade-offs, challenge assumptions, and turn architecture into dependable production outcomes.\nTL;DR\nOwn the secure AWS and Kubernetes platform that supports fraud, data, microservices, and machine learning workloads.\nLead hands-on decisions across Terraform, Atlantis, GitLab CI/CD, Helm, Python automation, observability, and release controls.\nDrive reliability, scalability, security, auditability, and cost-conscious use of platform resources across Development, UAT, and Production.\nEnable production-grade MLOps while reducing dependency on senior architects through clear judgment and accountable execution.\nWhy This Role Matters\nPlatform reliability directly affects the stability and responsiveness of systems supporting fraud intelligence and live model operations.\nSecure, traceable infrastructure and deployment decisions are essential in a regulated and security-sensitive environment.\nGrowing data workloads require deliberate architecture that scales without adding unnecessary operational complexity.\nStrong technical ownership will create faster decisions, clearer accountability, and more sustainable delivery across engineering and data teams.\nWhat You Will Actually Do\nOwn and evolve secure, highly available AWS platforms across segregated Development, UAT, and Production environments.\nArchitect and operate Kubernetes clusters, deployment patterns, workload scaling, resource optimisation, containers, and Helm releases.\nBuild controlled infrastructure workflows with Terraform and Atlantis, preserving traceability of infrastructure and configuration changes.\nLead GitLab CI/CD design for automated build, test, security validation, promotion, deployment, rollback, and release governance.\nDrive observability, incident response, root-cause analysis, performance improvement, resilience, and cost-effective resource use.\nPartner with Data Science, ML, security, networking, and infrastructure teams to enable secure model deployment and lifecycle operations.\nExamples of This Role in Practice\nA deployment introduces instability in Production: lead diagnosis, decide the rollback path, and drive a sustainable fix rather than a temporary workaround.\nA data workload grows beyond 100GB per week: evaluate scaling options, make the trade-off explicit, and evolve the platform without over-engineering it.\nA new tool could accelerate delivery but open-source adoption is restricted: assess the security and governance implications and recommend a viable path.\nA live model needs lower latency and stronger monitoring: align platform, ML, and observability decisions to improve production performance and control.\nAn existing architecture decision is unclear: review the standards and decision records, challenge assumptions constructively, and document the chosen direction.\nWhat Will Help You Succeed\nAt least 8 years of relevant experience in DevOps, Cloud Engineering, Platform Engineering, Site Reliability Engineering, or related infrastructure roles.\nDeep production experience with AWS, Kubernetes, Terraform, GitLab CI/CD, Helm, Python automation, and mission-critical systems.\nStrong judgment in cloud architecture, networking, IAM, secrets management, environment segregation, deployment controls, and security trade-offs.\nPractical strength in monitoring, alerting, troubleshooting, performance optimisation, incident response, and root-cause analysis.\nAbility to explain why systems are designed as they are, question assumptions, decide with incomplete information, and communicate clearly with stakeholders.\nUseful exposure includes regulated environments, large-scale or distributed data processing, and production MLOps tools or lifecycles.","description_format":"text","description_chars":4128,"description_truncated":false,"requirements":{"experience_years_min":8,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"phd","optional":false},"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-09-26T01:11:32Z"}],"liveness":{"score":86,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.859,"p_room":1,"age_days":5,"expected_fill_days":24,"reasons":["conf:36","velocity","win:early"],"computed_at":"2026-10-01T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/paynet-principal-engineer-platform-engineering","json_url":"https://alion.io/job/paynet-principal-engineer-platform-engineering.json","meta":{"generated_at":"2026-10-02T03:05:33Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":4267,"day_limit":5000,"remaining_today":733,"minute_limit":60,"resets_at":"2026-10-03T00:00:00Z"}}}