Back to jobsSign in and upload your CV to see how well you match this job.
Sign inJob overview
- Location
- Cairo, Egypt
- Workplace
- Hybrid
- Employment type
- Full-time
- Experience level
- Lead
- Date posted
- Oct 6, 2026
- Last checked at the source
- Oct 7, 2026
- Job source
- via Workable
Responsibilities
• Ensure the availability, performance, and resilience of production environments by proactively monitoring systems and resolving complex operational issues.
• Administer and optimize CI/CD pipelines, containerized platforms, IBM Cloud Pak Stacks, Confluent, Elasticsearch, and supporting infrastructure to improve operational efficiency.
• Lead incident response, perform root cause analysis, and implement preventive actions to reduce recurring issues and improve service reliability.
• Develop and enhance monitoring, logging, alerting, automation, and operational processes to improve system observability and reduce manual effort.
• Provide technical leadership and mentorship to junior engineers while supporting deployments, release management, and production readiness activities.
• Maintain system security, backup, disaster recovery, and compliance standards while ensuring accurate operational documentation and runbooks.
• Collaborate with Development, Platform Engineering, QUALITY, Network, and Security teams to resolve complex technical challenges and continuously improve production operations.
• Bachelor's degree or Diploma in Computer Science, Engineering, or a related field.
• 4–6 years of experience in Technical Operations, Production Support, Site Reliability Engineering (SRE), DevOps, or System Administration.
• Strong experience with Linux/Windows administration, Docker, Kubernetes, OpenShift, CI/CD pipelines, automation tools, scripting (Bash, PowerShell, Python), and configuration management.
• Hands-on experience with IBM Cloud Pak Stacks (CP4BA, CP4I, CP4D), Confluent, Elasticsearch, Databases (SQL, DB2, MongoDB, etc.), JVM troubleshooting, monitoring, and observability platforms.
• Strong understanding of networking, security best practices, IAM, production support, backup, disaster recovery, scalability, performance tuning, and service reliability principles.
• Proven experience in incident management, troubleshooting, root cause analysis, and implementing operational improvements within enterprise production environments.
• Excellent technical leadership, analytical, communication, mentoring, and collaboration skills with the ability to manage multiple priorities effectively.
Requirements
• Bachelor's degree or Diploma in Computer Science, Engineering, or a related field.
• 4–6 years of experience in Technical Operations, Production Support, Site Reliability Engineering (SRE), DevOps, or System Administration.
• Strong experience with Linux/Windows administration, Docker, Kubernetes, OpenShift, CI/CD pipelines, automation tools, scripting (Bash, PowerShell, Python), and configuration management.
• Hands-on experience with IBM Cloud Pak Stacks (CP4BA, CP4I, CP4D), Confluent, Elasticsearch, Databases (SQL, DB2, MongoDB, etc.), JVM troubleshooting, monitoring, and observability platforms.
• Strong understanding of networking, security best practices, IAM, production support, backup, disaster recovery, scalability, performance tuning, and service reliability principles.
• Proven experience in incident management, troubleshooting, root cause analysis, and implementing operational improvements within enterprise production environments.
• Excellent technical leadership, analytical, communication, mentoring, and collaboration skills with the ability to manage multiple priorities effectively.
Skills
- Python
- SQL
- Bash
- MongoDB
- Elasticsearch
- Docker
- Kubernetes
- CI/CD
- Linux
- DevOps
- Networking
- Operations
- Communication
- Leadership
Visa and relocation
?The posting doesn't mention visa sponsorship. Check the original posting or ask the company.
?The posting doesn't mention relocation.
Job description
Our Staff TechOps & Support Engineer Ensures the stability, performance, and availability of production systems by managing infrastructure, monitoring environments, responding to incidents, automating operational tasks, supporting deployments, and maintaining security, backups, and documentation
Responsibilities
• Ensure the availability, performance, and resilience of production environments by proactively monitoring systems and resolving complex operational issues.
• Administer and optimize CI/CD pipelines, containerized platforms, IBM Cloud Pak Stacks, Confluent, Elasticsearch, and supporting infrastructure to improve operational efficiency.
• Lead incident response, perform root cause analysis, and implement preventive actions to reduce recurring issues and improve service reliability.
• Develop and enhance monitoring, logging, alerting, automation, and operational processes to improve system observability and reduce manual effort.
• Provide technical leadership and mentorship to junior engineers while supporting deployments, release management, and production readiness activities.
• Maintain system security, backup, disaster recovery, and compliance standards while ensuring accurate operational documentation and runbooks.
• Collaborate with Development, Platform Engineering, QUALITY, Network, and Security teams to resolve complex technical challenges and continuously improve production operations.
• Bachelor's degree or Diploma in Computer Science, Engineering, or a related field.
• 4–6 years of experience in Technical Operations, Production Support, Site Reliability Engineering (SRE), DevOps, or System Administration.
• Strong experience with Linux/Windows administration, Docker, Kubernetes, OpenShift, CI/CD pipelines, automation tools, scripting (Bash, PowerShell, Python), and configuration management.
• Hands-on experience with IBM Cloud Pak Stacks (CP4BA, CP4I, CP4D), Confluent, Elasticsearch, Databases (SQL, DB2, MongoDB, etc.), JVM troubleshooting, monitoring, and observability platforms.
• Strong understanding of networking, security best practices, IAM, production support, backup, disaster recovery, scalability, performance tuning, and service reliability principles.
• Proven experience in incident management, troubleshooting, root cause analysis, and implementing operational improvements within enterprise production environments.
• Excellent technical leadership, analytical, communication, mentoring, and collaboration skills with the ability to manage multiple priorities effectively.
GetGlobalJob is not the employer or a recruiting agency. You apply on the original publisher's site: always check the posting before sharing your details, and never pay for a job.
Check your fit for this job
Create your free account and upload your CV to see how well you match this job and which skills you're missing.
Check my fit