DevOps Engineer - CI/CD & Monitoring (Remote, China)

🏢 Bjak · all Bjak jobs
📍 China
📅 Posted 2026-06-28 · via Himalayas
🏷 DevOps-Engineer,DevOps-Engineer-II,Infrastructure-DevOps-Engineer
Apply on original site ↗

BJAK’s automation systems power end-to-end insurance journeys across quote generation, policy issuance, renewals, endorsements, claims, payments and insurer integrations. These systems are business-critical, where deployment stability, monitoring and fast recovery directly impact customers and operations.

We're looking for a DevOps Engineer based in China to strengthen CI/CD systems, monitoring infrastructure and production visibility across BJAK’s AI automation platform, ensuring engineers can ship safely and systems remain highly observable and reliable.

This is a fully remote position where you'll collaborate closely with our Malaysia-based engineering, product and operations teams to improve deployment safety and system observability at scale.
The Mission

Build and maintain reliable CI/CD pipelines and monitoring systems that enable fast, safe and observable deployments across BJAK’s AI automation platform, reducing production risk while improving system visibility and operational confidence.
What You’ll Own

-
Design and maintain CI/CD pipelines for multiple services across the platform.

-
Improve deployment automation, release strategies and rollback mechanisms.

-
Build and enhance monitoring, alerting and observability systems across production services.

-
Ensure system health visibility through metrics, logs, traces and dashboards.

-
Work with engineers to reduce deployment risk and improve release confidence.

-
Implement safe deployment strategies such as canary, blue-green or phased rollouts.

-
Improve incident detection speed and reduce mean time to recovery (MTTR).

-
Support infrastructure reliability for business-critical insurance workflows.

-
Standardize deployment and monitoring practices across engineering teams.

-
Continuously improve CI/CD performance, stability and maintainability.

What We're Looking For

-
Experience in DevOps, SRE, platform engineering or infrastructure roles.

-
Strong understanding of CI/CD pipelines, deployment automation and release engineering.

-
Experience with monitoring, logging and observability systems in production environments.

-
Ability to troubleshoot deployment and production issues in a structured and calm manner.

-
Strong understanding of system reliability, uptime and operational risk.

-
Experience supporting production systems with high availability requirements.

-
Hands-on ownership mindset during incidents and deployment failures.

-
Practical judgment on release safety, performance and system stability.

-
Strong collaboration with engineering teams in fast-paced environments.

-
Low ego and disciplined approach to production operations.

Bonus Points

-
Experience with Jenkins, GitHub Actions, GitLab CI or similar CI/CD tools.

-
Experience with Kubernetes, Docker or container-based deployments.

-
Experience with observability stacks (Prometheus, Grafana, ELK, Datadog, etc.).

-
Experience with infrastructure-as-code tools (Terraform, Ansible, etc.).

-
Experience with zero-downtime deployments and progressive delivery strategies.

-
Experience with cloud platforms (AWS, GCP, Azure).

-
Experience in fintech, insurance or other high-availability industries.

-
Experience improving deployment velocity and reliability at scale.

-
Contributions to CI/CD or monitoring system improvements.

The Kind of Builder We Want

-
Thinks in deployment safety, system visibility and operational reliability.

-
Hands-on engineer who understands both pipelines and production systems deeply.

-
Calm and structured when handling deployment failures or production incidents.

-
Strong focus on observability, automation and release confidence.

-
Proactive in preventing issues rather than reacting to them.

-
Careful and deliberate when making production changes.

-
Builds systems engineers trust to deploy frequently and safely.

This Role Is Not For

-
Engineers who only react to deployment failures instead of preventing them.

← All remote jobs