Overview
We’re looking for a highly skilled, hands-on Cloud & Network Infrastructure Engineer who can confidently operate and scale network, virtualization, storage, and hosting infrastructure across global data centers-and who can solve issues fast, automate aggressively, and contribute new ideas that improve our platform and product offerings.
- Design, deploy, and manage scalable cloud VPS infrastructure across multiple data centers.
- Build and operate Proxmox clusters (primary platform), ensuring HA, performance, and stability.
- Work with and/or support additional virtualization and cloud platforms when needed
- Implement resource planning, capacity forecasting, and performance optimization across virtualization stacks.
- Develop and improve provisioning pipelines, templates, images, and deployment standards.
- Install, manage, and optimize Ceph (MON/OSD/MGR/RGW) for high-performance and fault-tolerant storage.
- Improve storage performance, balancing, replication, recovery, and cluster scalability.
- Implement backup and snapshot strategies, and ensure storage availability for VPS and cloud services.
- Configure, monitor, and troubleshoot Juniper switching/routing hardware.
- Own network performance optimization: routing stability, throughput, latency, packet loss, and capacity planning.
- Work with routing and HA technologies (BGP/OSPF/VRRP, VLANs, LACP, etc.).
- Design secure and reliable interconnects between data centers and cloud infrastructure.
- Configure and manage firewalls, VPNs, access policies, and network security controls.
- Implement security best practices: segmentation, ACL policies, patching strategy, vulnerability mitigation, and incident response readiness.
- Work with secure remote access, bastion models, and least-privilege access.
- Manage production hosting/server environments on Linux and Windows Server
- Maintain and optimize web server stacks and hosting infrastructure
- Monitor infrastructure performance and proactively resolve incidents to minimize downtime.
- Build alerting and observability (metrics/logs/traces) and improve real-time monitoring workflows.
- Create and implement disaster recovery plans, failover designs, and redundancy strategies.
- Participate in on-call/shift operations as required for a global hosting platform.
- Maintain detailed infrastructure documentation: configurations, diagrams, policies, and runbooks.
- Collaborate with DevOps, Support, Product, and Leadership teams to ensure reliability, availability, and scalability.
- Recommend architecture improvements and propose new product/feature ideas for Ultahost’s cloud offerings.
- 5+ years of experience in cloud VPS management, hosting infrastructure, data center operations, or similar roles.
- Strong, hands-on experience with Proxmox clusters + KVM/QEMU
- Strong, hands-on experience with Ceph storage (deployment, troubleshooting, tuning, scaling)
- Strong, hands-on experience with Juniper networking (routing/switching/troubleshooting)
- Solid expertise across Virtualization & HA concepts (cluster HA, migration, replication, failover)
- Solid expertise across Networking protocols and technologies (BGP, OSPF, VLAN, VRRP, LACP, etc.)
- Solid expertise across Performance tuning and root-cause analysis under pressure
- Solid expertise across Security fundamentals (firewalls, VPNs, hardening, patch management)
- Solid expertise across Monitoring and incident response best practices
- Strong English written & verbal communication skills.
- Comfortable with fast-paced environments, autonomy, and shift/on-call requirements.
- Bachelor’s degree in a related field (or equivalent experience).
- Experience with OpenStack, VMware, Hyper-V, oVirt/RHV, Xen, LXD/LXC.
- Automation experience with Ansible / Terraform / Python / Bash / CI-CD.
- Kubernetes knowledge (deploying/operating clusters for customers).
- Hosting billing/provisioning integration experience (WHMCS or custom systems).
- Background in web hosting or cloud industry (preferred).
- Strong product mindset: ability to propose and shape new cloud/hosting offerings.
- Virtualization
- Storage Management
- Network Engineering
- Performance Optimization
- Security Hardening
- Incident Response
- Capacity Planning
- Disaster Recovery
- Proxmox
- KVM
- QEMU
- VMware vSphere/ESXi
- Microsoft Hyper-V
- OpenStack
- oVirt
- RHV
- Xen
- LXD/LXC
- Virtuozzo
- Ceph
- Juniper
- BGP
- OSPF
- VRRP
- VLAN
- LACP
- Ubuntu
- Debian
- AlmaLinux
- Rocky Linux
- CentOS
- RHEL
- Windows Server
- Nginx
- Apache
- LiteSpeed
- OpenLiteSpeed
- PHP-FPM
- MySQL
- MariaDB
- Redis
- HAProxy
- Keepalived
- Ansible
- Terraform
- Python
- Bash
- CI-CD
- Kubernetes
✨ Our intelligent job search engine discovered this job and republished it for your convenience.
Please be aware that the job information may be incorrect or incomplete. The job announcement remains the property of its original publisher. To view the original job and its full details, please visit the job's URL on the owner’s page.
Please clearly mention that you have heard of this job opportunity on https://ijob.am.

