Belong, connect, grow, with KBR!
The KBR team of teams delivers future-forward science, technology and engineering solutions and mission-critical services that help governments and companies around the world accomplish their most important objectives, while also helping achieve their sustainability goals.
KBR Sustainable Technology Solutions provides holistic and value-added solutions across the entire asset life cycle. These include world-class licensed process technologies, differentiated advisory services, deep technical domain expertise, energy transition solutions, high-end design and engineering capabilities, and smart solutions to optimize planned and operating assets.
Job purpose
The IT HCS & VMware Platform Supervisor is responsible for supervising the operation, administration, governance, and continuous improvement of the Huawei Cloud Stack (HCS) platform and related private or hybrid cloud services across the IT data center environment.
The role leads HCS platform support activities to ensure secure, available, scalable, and recoverable cloud services, while coordinating with infrastructure, network, security, application, vendor, and service management teams to meet business continuity and operational requirements.
Duties and responsibilities
- Supervise daily administration of Huawei Cloud Stack services including ManageOne, FusionCompute, resource pools, projects or tenants, compute, storage, virtual networking, identity and access, monitoring, and platform health.
- Lead and guide platform engineers and support teams in incident handling, service requests, change implementation, preventive maintenance, lifecycle upgrades, and compliance with approved procedures.
- Monitor HCS capacity, utilization, performance, availability, licensing, quotas, and service levels; prepare forecasts and implement optimization plans for efficient resource usage.
- Oversee integration of HCS with VMware platforms, enterprise storage, SAN or NAS connectivity, backup systems, Active Directory, DNS, networking, security, monitoring, and IT service management tools.
- Ensure backup, replication, snapshot, restore, and disaster recovery processes are properly configured and tested; coordinate recovery exercises against approved RPO and RTO targets.
- Review and approve technical change plans, method statements, risk assessments, rollback plans, patching activities, configuration baselines, and security hardening requirements.
- Troubleshoot and supervise resolution of complex platform incidents by analyzing logs, alerts, performance data, and root causes; coordinate corrective actions with Huawei, vendors, and internal teams.
- Maintain platform documentation, architecture diagrams, runbooks, SOPs, asset records, KPI reports, and continuous improvement actions; promote automation using PowerShell, PowerCLI, Python, APIs, or approved orchestration tools.
Qualifications, training & certifications
- Bachelor’s degree in Computer Engineering, Computer Science, Information Technology, Network Engineering, or a related discipline.
- At least one current professional certification in cloud computing, virtualization, systems infrastructure, storage, backup, or IT service management is required.
- Huawei HCIP-Cloud Computing, HCIP-Cloud Service, or equivalent Huawei Cloud Stack professional certification is strongly preferred; HCIE-Cloud Computing is an advantage.
- Huawei Cloud Stack administration, ManageOne, FusionCompute, FusionStorage, or HCS operations training is strongly preferred.
- VMware Certified Professional - VCF Administrator, VVF Administrator, or Data Center Virtualization certification is preferred.
- Microsoft Certified: Azure Administrator Associate (AZ-104) or Azure Solutions Architect Expert (AZ-305 path) is an advantage for hybrid cloud responsibilities.
- Cisco Certified Network Associate (CCNA) or equivalent networking certification is preferred.
- ITIL Foundation (current version) is preferred; service management, change management, or project coordination training is an advantage.
- Veeam Certified Engineer Plus (VMCE+) or an equivalent enterprise backup and recovery certification is an advantage.
- Strong knowledge of HCS operations, virtualization, storage, backup, disaster recovery, Windows and Linux systems, networking, security, monitoring, and automation scripting.
- Strong leadership, professional communication, analytical troubleshooting, technical reporting, architecture documentation, and stakeholder coordination skills.
Experience
- 7-10 years of professional experience in cloud, virtualization, systems, data center, or enterprise infrastructure operations.
- Minimum 4-6 years of hands-on experience with Huawei Cloud Stack, ManageOne, FusionCompute, VMware, enterprise storage, backup, replication, and disaster recovery operations, including at least 2-3 years in a lead or supervisory role.
Language proficiency, computer and software skills
Core platforms and tools:
- Huawei Cloud Stack, ManageOne, FusionCompute, FusionStorage, and HCS monitoring tools
- VMware vSphere, vCenter, ESXi, VCF or VVF, templates, clusters, HA and DRS
- SAN/NAS storage, backup and replication tools, PowerShell, PowerCLI, Python, APIs, monitoring platforms, and ITSM ticketing tools
- English - upper intermediate or above (verbal and written)
Skills
- Strong leadership and supervisory capability across HCS platform operations, incident response, change delivery, vendor coordination, and team workload management.
- Advanced troubleshooting across HCS, virtualization, storage, backup, networking, security, and infrastructure services, with strong availability and risk awareness.
- Ability to manage incidents, problems, changes, risks, capacity, lifecycle activities, audit findings, and service improvement plans within agreed service levels.
- Strong documentation, reporting, automation, governance, and communication skills with readiness to support critical incidents, migrations, upgrades, and recovery activities.
انتمِ، اتصل، نمُوِّ مع كِب آر!
فريق كِب آر يقدِّم حلولاً علمية وتكنولوجية وهندسية متقدمة ومُهِم الخدمات الحيوية التي تساعد الحكومات والشركات حول العالم في إنجاز أهدافها الأكثر أهمية، مع المساهمة أيضاً في تحقيق أهداف الاستدامة.
توفر حلول التكنولوجيا المستدامة من كِب آر حلولاً شمولية وقيمة مضافة عبر دورة حياة الأصول كاملة. وتشمل تقنيات عمليات مرخصة من الطراز العالمي، وخدمات استشارية مميزة، وخبرة فنية عميقة، وحلول تحوّل الطاقة، وقدرات تصميم وهندسة عالية المستوى، وحلول ذكية لتحسين الاستغلال لكل من الأصول المخططة والتشغيلية.
هدف الوظيفة
المشرف على منصة Huawei Cloud Stack (HCS) هو المسؤول عن إشراف تشغيل، وإدارة، وحوكمة، وتحسين مستمر لمنصة Huawei Cloud Stack وخدماتها السحابية الخاصة أو híbrida المرتبطة عبر بيئة مركز بيانات تقنية المعلومات.
يقود دور دعم منصة HCS لضمان خدمات سحابية آمنة ومتاحة وقابلة للتوسع وقابلة للاسترداد، مع التنسيق مع فرق البنية التحتية، والشبكات، والأمن، والتطبيق، والبائعين، وإدارة الخدمات لتحقيق متطلبات استمرارية العمل والتشغيل.
الواجبات والمسؤوليات
- اشراف الإدارة اليومية لخدمات Huawei Cloud Stack بما في ذلك ManageOne، FusionCompute، تجمعات الموارد، المشاريع أو المستأجرين، الحوسبة، التخزين، الشبكات الافتراضية، الهوية والوصول، المراقبة، وصحة المنصة.
- قيادة وتوجيه مهندسي المنصة وفرق الدعم في معالجة الحوادث، وطلبات الخدمة، وتنفيذ التغييرات، والصيانة الوقائية، وترقيات دورة الحياة، والامتثال للإجراءات المعتمدة.
- مراقبة سعة HCS والاستخدام والأداء والتوافر والترخيص والحصص ومستويات الخدمة؛ إعداد التوقعات وتنفيذ خطط التحسين لاستخدام الموارد بكفاءة.
- الإشراف على تكامل HCS مع منصات VMware والتخزين المؤسسي والتواصل SAN أو NAS وأنظمة النسخ الاحتياطي وActive Directory وDNS والشبكات والأمن وأدوات المراقبة وإدارة خدمات تكنولوجيا المعلومات.
- ضمان تكوين واختبار عمليات النسخ الاحتياطي والتكرار والنقاط والتعافي من الكوارث؛ تنسيق تمارين الاستشفاء وفق أهداف RPO وRTO المعتمدة.
- مراجعة واعتماد خطط التغيير التقنية وبيانات الأساليب وتقييم المخاطر وخطط التراجع وأنشطة التحديث والاعتيان على المعايير الأمنة.
- استكشاف الأعطال وإشراف حلها في حوادث المنصة المعقدة من خلال تحليل السجلات والتنبيهات وبيانات الأداء والجذور؛ تنسيق الإجراءات التصحيحية مع Huawei والبائعين والفرق الداخلية.
- الحفاظ على توثيق المنصة، ومخططات الهندسة، ودفاتر التشغيل، وإجراءات التشغيل القياسية، وسجلات الأصول وتقارير KPI وإجراءات التحسين المستمر؛ تعزيز الأتمتة باستخدام PowerShell وPowerCLI وPython وواجهات برمجة التطبيقات أو أدوات التنظيم المعتمدة.
المؤهلات والتدريب والشهادات
- درجة البكالوريوس في هندسة الحاسوب أو علوم الحاسوب أو تكنولوجيا المعلومات أو هندسة الشبكات أو تخصص ذي صلة.
- على الأقل شهادة مهنية حالية واحدة في الحوسبة السحابية، الافتراضية، بنية الأنظمة التحتية، التخزين، النسخ الاحتياطي، أو إدارة خدمات تكنولوجيا المعلومات مطلوبة.
- يفضل بشدة شهادة Huawei HCIP-Cloud Computing أو HCIP-Cloud Service أو ما يعادلها من شهادات Huawei Cloud Stack؛ وتُعد HCIE-Cloud Computing ميزة.
- يفضل بشدة تدريب إدارة Huawei Cloud Stack، ManageOne، FusionCompute، FusionStorage، أو عمليات HCS.
- يفضل شهادة VMware Certified Professional - VCF Administrator أو VVF Administrator أو شهادة افتراضية مراكز البيانات.
- يفضل Microsoft Certified: Azure Administrator Associate (AZ-104) أو Azure Solutions Architect Expert (AZ-305 مسار).
- يفضل Cisco Certified Network Associate (CCNA) أو شهادة شبكات مكافئة.
- يفضل ITIL Foundation (النسخة الحالية)؛ ويفضل تدريب إدارة الخدمات، إدارة التغيير، أو تنسيق المشاريع.
- يفضل Veeam Certified Engineer Plus (VMCE+) أو شهادة نسخ احتياطي وتحليل استرداد مؤسسة مكافئة.
- معرفة قوية بعمليات HCS، الافتراضية، التخزين، النسخ الاحتياطي، استرداد الكوارث، أنظمة Windows وLinux، الشبكات، الأمن، المراقبة، وبرمجة الأتمتة.
- قيادة قوية، تواصل مهني، تحليل واستقصاء تقني، تقارير فنية، توثيق الهندسة، وتنسيق أصحاب المصلحة.
الخبرة
- 7-10 سنوات خبرة مهنية في عمليات السحابة، الافتراضية، الأنظمة، مراكز البيانات، أو بنية تحتية مؤسسية.
- لا يقل عن 4-6 سنوات خبرة عملية مع Huawei Cloud Stack وManageOne وFusionCompute وVMware والتخزين المؤسسي والنسخ الاحتياطي والتكرار وعمليات استرداد الكوارث، بما فيها سنتين إلى ثلاث سنوات على الأقل في دور قيادي أو إشرافي.
اللغة والمهارات الحاسوبية والبرمجية
المنصات والأدوات الأساسية:
- Huawei Cloud Stack، ManageOne، FusionCompute، FusionStorage، وأدوات مراقبة HCS
- VMware vSphere، vCenter، ESXi، VCF أو VVF، القوالب، العنقود، HA وDRS
- تخزين SAN/NAS، أدوات النسخ الاحتياطي والتكرار، PowerShell، PowerCLI، Python، APIs، منصات المراقبة، وأدوات تذاكر ITSM
- الإنجليزية - مستوى مرتفع أو أعلى (شفوياً وكتابياً)
المهارات
- قدرات قيادية وإشراف قوية عبر عمليات منصة HCS، استجابة الحوادث، تنفيذ التغييرات، تنسيق الموردين، وإدارة عبء عمل الفريق.
- استكشاف أخطاء متقدمة ضمن HCS، الافتراضية، التخزين، النسخ الاحتياطي، الشبكات، الأمن، وخدمات البنية التحتية، مع وعي عالٍ بالتوافر والمخاطر.
- قدرة على إدارة الحوادث والمشكلات والتغييرات والمخاطر والسعة وأنشطة دورة الحياة وتقرير التدقيق وخطط تحسين الخدمة ضمن مستويات الخدمة المتفق عليها.
- توثيق قوي، تقارير، أتمتة، حوكمة، واتصالات مع جاهزية لدعم الحوادث الحرجة، والعمليات الانتقالية، والترقيات، وأنشطة الاسترداد.