800+ Devops Jobs - September 2026 - High Salaries

Showing 803 jobs results for "devops"
Never miss any updates for Devops jobs

KL City

  • You will be working in a newly set-up technology centre located in Kuala Lumpur as part of Technology and Operations to deliver innovative financial technology solutions that enable business growth and technology transformation.
  • Manage application release and deployment activities for Microsoft Dynamics 365 Customer Onboarding, Leads, and Lending applications.
  • Plan, coordinate, and execute application deployments across environments. ...
Posted
6 days ago

Singapore

  • Experience with well-architected framework pillars (especially reliability, security, cost optimization).
  • Designing fault-tolerant and horizontally scalable systems
  • Advanced proficiency in Terraform, CloudFormation, or CDK ...
Posted
10 days ago

Hong Kong

  • Coordinate incident management, root cause analysis, and issue resolution.
  • Participate in after-hours production support and critical incident response when required.
  • Design and enhance platform observability, monitoring, and alerting capabilities. ...
Posted
10 days ago

Singapore

  • Support User Acceptance Testing (UAT) including coordination with testers and issue tracking
  • Manage and track defects/issues during UAT and hypercare period
  • Document outcomes, highlight risks/issues, and support go-live activities ...
Posted
10 days ago

KL City

Posted
10 days ago

Singapore

  • Experience with well-architected framework pillars (especially reliability, security, cost optimization).
  • Designing fault-tolerant and horizontally scalable systems
  • Advanced proficiency in Terraform, CloudFormation, or CDK ...
Posted
10 days ago

Singapore

  • Good understanding of Platform Engineering principles and practices.
  • Strong troubleshooting and debugging skills.
  • Experience with Java/Spring Boot, REST APIs, and Microservices. ...
Posted
10 days ago

Singapore

  • Experience with well-architected framework pillars (especially reliability, security, cost optimization).
  • Designing fault-tolerant and horizontally scalable systems
  • Advanced proficiency in Terraform, CloudFormation, or CDK ...
Posted
10 days ago

Singapore

  • Experiment with prompt engineering, fine-tuning, or orchestration of AI workflows
  • Cloud Operations Optimization
  • Analyze existing cloud infrastructure processes and identify opportunities for automation and optimization. ...
Posted
10 days ago

Singapore

  • Support data-driven decision-making through insights and reporting
  • Dashboard Development & Modernization
  • Build and optimize interactive dashboards for decision-making ...
Posted
10 days ago

Singapore

  • Manage backup operations and recoverability, including backup verification, scheduled restore testing and recovery checks.
  • Administer virtualisation platforms, including host lifecycle, capacity management, VM provisioning, inventory management and platform migrations.
  • Provide operational support for Active Directory and related identity services, including directory changes, migrations and regional cutovers led by the Engineering team. ...
Posted
10 days ago

Singapore

  • Undergraduate, or Postgraduate who is currently pursuing a degree/master in computer science, software engineering or related fields
  • Experience in Java/Golang (go) programming language and familiar with common Linux commands;
  • Familiar with Mysql or any relational database, with certain SQL writing skills and optimization experience; ...
Posted
10 days ago

JONES LANG LASALLE PROPERTY CONSULTANTS PTE LTD

Singapore

  • Lead the development of emergency response plans and coordinate regular EHS training for team members of the managed site
  • Implement and maintain Computerized Maintenance Management Systems (CMMS) for tracking work orders and assets
  • Provide accountability for helpdesk administration, end-user forums, and complaint management systems ...
Posted
10 days ago

ALLEGIS GROUP SINGAPORE PRIVATE LIMITED

Singapore

  • Build and maintain CI/CD pipelines and deployment automation
  • Develop automation solutions using Python and Shell scripting
  • Support AWS-based cloud environments and platform modernisation initiatives ...
Posted
10 days ago

Singapore

  • Responsible for capacity planning of Compute infrastructure, including licensing and vendor support related engagements. Also includes obsolescence management
  • Work with Change Management to propose and implement changes, especially for patching and migrations.
  • Works with other groups to implement new builds ...
Posted
10 days ago

Singapore

  • The ideal candidate possesses strong technical expertise, is proactive, and is passionate about engineering excellence. You will work closely with platform, operations, application, and cross-functional teams to design and deliver automation solutions that enhance efficiency, reliability, and security across the organization. This role encompasses both BAU operations and project-based initiatives, requiring the ability to balance operational support with continuous improvement and delivery.
  • This role also embraces AI-driven innovation, including the use of LLMs, intelligent automation, and agentic AI, to improve operational efficiency, reduce operational toil, strengthen security posture, and enable scalable, self-service platforms.
  • Key ResponsibilitiesDevSecOps Engineering (CI/CD, Tools & Practices)•    Strong understanding of DevSecOps principles.•    Design, build, and maintain end to end CI/CD pipelines for application and platform deployments.•    Strong hands on experience with CI/CD tooling such as Jenkins, Redhat Ansible, Octopus, AWS CodePipeline, and source control platforms including Bitbucket or AWS CodeCommit.•    Automate build, test, security scanning, packaging, and deployment workflows across multiple environments.•    Develop and maintain reusable pipeline templates and automation frameworks to standardize delivery practices.•    Continuously improve delivery speed, reliability, quality, and developer experience through automation and best practices.•    Hands on experience with Red Hat Ansible for configuration management, provisioning, and automation.•    Manage artifacts, binaries, and dependencies using repositories such as JFrog Artifactory and Sonatype Nexus Repository.•    Strong understanding of artifact lifecycle management, dependency control, and versioning best practices.•    Embed security and quality controls into CI/CD pipelines, including: o    Static Application Security Testing (SAST) using tools such as Fortify.o    Code quality and secrets scanning using SonarQube.o    Software Composition Analysis (SCA), 3rd party / OSS libraries using Sonatype Nexus.•    Collaborate closely with Security, Risk, and Compliance, application teams to meet security standards and regulatory requirements.•    Automate security validations, policy enforcement, guardrails, and audit evidence collection across pipelines and cloud environments.•    Support audit, risk, and compliance activities through repeatable and automated controls.________________________________________Cloud & Platform Engineering (AWS)•    Build and operate cloud native CI/CD pipeline and automation workflows/solution on AWS.•    Hands-on experience in Terraform to build for provisioning, managing and automating cloud infrastructure using Infrastructure as Code (IaC) best practices.•    Design solution using multiple AWS services including: o    EC2, ECS/EKS, Lambda, IAM, VPC, ALB/NLB, S3, RDS, CodePipeline, CodeBuild, AWS Inspector, CloudWatch, Glue, etc.•    Design secure, scalable, and resilient cloud architectures aligned with operational and compliance requirements.•    Implement cloud automation for provisioning, configuration, scaling, and lifecycle management.________________________________________Automation, Reliability & Operational Excellence•    Design automation with reliability, security, and scalability in mind (idempotency, error handling, retries, rollback).•    Automate manual operational tasks to reduce toil and improve system reliability. •    Implement logging, monitoring, and alerting for automation to ensure visibility and maintainability. •    Apply Infrastructure / Operations automation principles (e.g. configuration-driven automation, version control, code reviews).•    Work together with various team to identify automation opportunities.•    Proficiency in automation and scripting, particularly using Python and Unix shell scripting.________________________________________AI Driven & Future Capabilities•    Drive adoption of AI across teams by identifying practical use cases and integrating AI capabilities into everyday workflows to improve productivity, quality and operational efficiency by leveraging AI Assisted tools.•    Identify, prototype, and scale AI driven solutions that streamline workflows, reduce manual effort, and accelerate delivery. •    Work with multiple teams to promote effective, responsible, and secure usage of AI technologies.________________________________________General•    Collaborate with various team, eg. application, infrastructure, security, operations, product, IT Service Management team, IT Governance, Risk & Compliance, or Auditor.•    Some experiences in handling audit and risk activities from MAS, Group audit, Risk & Compliance, and 3rd party auditor.•    Prepare and update Automation and DevSecOps documentations, knowledgebase, SOP, best practices as required.•    Able to mentor automation / devsecops engineer peers as required.•    Comfortable liaising with vendor to discuss on the requirements, solutioning and validating the implementation. ...
Posted
10 days ago

Singapore

  • Experience with well-architected framework pillars (especially reliability, security, cost optimization).
  • Designing fault-tolerant and horizontally scalable systems
  • Advanced proficiency in Terraform, CloudFormation, or CDK ...
Posted
10 days ago
Posted
10 days ago

Singapore

  • Deliver solutions that blend new features with BAU tasks (e.g., bug fixes, minor refactoring), maintaining a stable foundation while remaining hands-on.
  • Surface observations and learnings from your work to your senior EM for potential contribution back to the SWE Practice.
  • Apply SWE Practice guidance when establishing DevOps/SRE, quality, and engineering practices in your area - avoid defining norms independently. ...
Posted
14 days ago

Singapore

  • Understand MLOps/AIOps basics: deployment, rollback, drift monitoring.
  • Build or extend an agent skill for automation or support.
  • Observability & Automation – Set up dashboards/alerts; script repetitive ticket steps. ...
Posted
14 days ago

Singapore

  • Master's degree in Computer Science, Artificial Intelligence or a closely related discipline.
  • 2+ years of backend engineering experience, working on databases, AI applications, agentic systems, operations platforms, or comparable infrastructure products.
  • Proven ability to lead the design and implementation of core modules or system-level architectures, with a strong track record of delivering high-quality, robust software. ...
Posted
15 days ago

Singapore

Posted
15 days ago

AVATAR MODERN TECHNO SERVICES PTE. LTD.

Outram

  • Define and track KPIs, SLIs,and SLOs to support service reliability.
  • Integrate Datadog with incident management tools such as ServiceNow, PagerDuty, and OpsGenie.
  • Collaborate with development, infrastructure, and SRE teams to troubleshoot and resolve performance issues. ...
Posted
15 days ago

Tuen Mun

  • 保障从基础设施到线上服务的SLA,负责服务器硬件资源、虚拟化平台、Kubernetes集群及容器化应用的稳定性、性能和可靠性2.负责底层服务器资源规划、管理和优化,包括但不限于物理机、虚拟机(基于KVM/VMware等)的生命周期管理3.设计和实施高可用、容灾、弹性伸缩、资源调度等运维解决方案,提升系统容错与自愈能力4.进行操作系统(Linux)内核参数调优、性能瓶颈分析及故障深度排查,解决底层资源与系统级问题5.推动自动化运维体系建设,开发和维护基础设施及应用层的运维工具与平台,提升部署、监控、告警、故障处理效率6.进行系统性能分析、成本优化与容量规划,推动架构持续演进7.跟踪云原生和运维技术发展趋势,引入适合的新工具、新方法,提升整体运维水平
  • 任职要求:
  • 熟练掌握至少一门脚本或编程语言(Bash/Python/Go),具备自动化开发经验2.具备扎实的Linux操作系统知识,有系统调优、故障排查及内核参数优化经验3.熟悉Kubernetes架构、核心组件及生态工具,具备生产环境集群管理、故障排查和性能调优经验4.熟悉主流虚拟化技术(如KVM、VMware)及其管理工具,有私有云或资源池管理经验者优先5.深入理解Docker原理与实践,具备镜像构建、优化和安全加固经验6.熟悉微服务架构下的常见中间件与基础设施,如:a.负载均衡:LVS/Nginx/Haproxyb.应用服务器:Tomcat/Spring Bootc.中间件:Kafka/RabbitMQ/Apollo等;7.具备熟练的自动化运维经验,熟悉Ansible/Terraform或类似工具8.熟悉监控预警体系构建,熟练使用Prometheus、Grafana、Alertmanager,具备Zabbix或其他监控工具使用经验9.熟悉CI/CD流程与工具链,如Jenkins、GitLab CI或ArgoCD,具备流水线设计与优化经验10.具备公有云(AWS/Aliyun/Tencent Cloud)实践经验,熟悉常见云服务(VPC/ECS/S3/CLB等)的使用与管理11.该职位工作地在香港屯门 或 上水 ...
Posted
16 days ago

KL City

  • Incident Management: Respond to and manage incidents to minimize downtime and resolve issues quickly, including on-call support.
  • System Performance: Measure, analyze, and tune system performance to ensure efficiency and stability.
  • Infrastructure Management: Provision and manage cloud infrastructure, sometimes using Infrastructure as Code (IaC), and assist in platform management and capacity planning. ...
Posted
16 days ago

KL City

  • Stakeholder & Vendor Collaboration: Coordinate with business analysts, clients, and testing managers to drive functional alignment, verify SIT/UAT/performance testing, and deliver high-quality technical results.
  • Implementation & Support: Oversee the end-to-end technical implementation plan to ensure seamless production cutovers and structured post-implementation support.
  • Documentation & Reporting: Maintain standard project documentation and provide regular delivery status updates to senior leadership, proactively escalating schedule-impacting issues with proposed workarounds. ...
Posted
16 days ago

Singapore

  • Experience with well-architected framework pillars (especially reliability, security, cost optimization).
  • Designing fault-tolerant and horizontally scalable systems
  • Advanced proficiency in Terraform, CloudFormation, or CDK ...
Posted
16 days ago

Singapore

  • Experience integrating systems with enterprise monitoring, alerting, SIEM, or Incident Response workflows.
  • Experience defining and implementing runbooks, operational procedures, escalation paths, and production support models.
  • Role: Site Reliability Engineer (SRE) ...
Posted
16 days ago

Singapore

  • Define and track KPIs, SLIs, and SLOs to support service reliability.
  • Integrate Datadog with incident management tools such as ServiceNow, PagerDuty, and OpsGenie.
  • Collaborate with development, infrastructure, and SRE teams to troubleshoot and resolve performance issues. ...
Posted
16 days ago

Geylang

  • Configure and maintain application deployment artefacts, environment configurations, and infrastructure definitions.
  • Collaborate with solution architects, platform teams, cybersecurity stakeholders, and software development teams to resolve deployment, security, compliance, and operationalissues.
  • Automate operational and deployment-related tasks through scripting and approved automation tools where appropriate. ...
Posted
16 days ago