Bachelors Degree or above in Computer Science Information Technology Software Engineering or a related field.
Minimum 5 years of monitoring and O&M experience.
Strong experience in the full lifecycle operation and maintenance of large-scale private monitoring platforms.
Experience operating and maintaining monitoring clusters and business/database monitoring environments.
Proficient in the independent operation configuration and tuning of Zabbix Prometheus Grafana and ELK clusters.
Proficient in PromQL and able to troubleshoot monitoring issues such as metric interruptions log loss and cluster abnormalities.
Experience monitoring Linux Kubernetes MySQL Oracle Redis Kafka and Flink environments.
Strong understanding of database monitoring risk identification and anomaly handling.
Practical experience with ETL/CDC task monitoring and cross-system multi-source data reconciliation.
Proficient in Shell and Python scripting.
Able to develop automation tools and independently troubleshoot monitoring and business-related issues.
Mandatory: Fluent in Chinese (Listening Speaking Reading and Writing) for daily communication and documentation.
Nice to Have
Familiarity with AI tools for IT operations including intelligent fault analysis log anomaly identification script generation and O&M efficiency optimization.
Experience with big data link monitoring and alert governance for large-scale clusters.
Experience in monitoring system upgrades migrations and integration with internal tools.
The Package:
Attractive Salary (RM 4000 to RM 8000).
Performance related bonus for confirmed staff.
Annual Leave 15 days.
Medical Leave 14 days.
Medical and hospitalization coverage.
Employment Type : FULL_TIME Experience: years Vacancy: 1
Buat amaran kerja untuk carian ini
Service Engineer Monitoring O&M Engineer • Kuala Lumpur, Federal Territory, Malaysia