Principle Cloud Engineer
We are seeking a highly skilled and experienced
Cloud Engineer
to manage, optimize, and secure our Microsoft Azure cloud environments. You will play a key role in designing, implementing, and maintaining reliable, scalable, and secure Azure infrastructure to support multiple client cloud environments.
Azure Infrastructure & Architecture
Compute & Storage Management:
Design, deploy, and manage Azure compute resources (Virtual Machines, VM Scale Sets, Availability Sets) and storage solutions (Blob Storage, File Storage, Managed Disks) to support scalable, highly available environments.
Virtual Networking:
Architect and manage Azure Virtual Networks (VNets
), including subnets, VNet
peering, Network Security Groups (NSGs), VPN Gateways, and ExpressRoute connections.
Role-Based Access Control (RBAC):
Design and implement RBAC policies to enforce least-privilege access across subscriptions, resource groups, and individual resources.
Backup & Disaster Recovery:
Configure and maintain Azure Backup and Azure Site Recovery to protect workloads, ensure data durability, and support business continuity planning.
Identity & Access Management
Entra ID Administration:
Manage Microsoft Entra ID (Azure AD), including users, groups, enterprise applications, and conditional access policies, and understand how Entra ID interacts with Azure RBAC and resource access.
Hybrid Identity:
Support hybrid identity scenarios, including Entra Connect synchronization between on-premises Active Directory and Microsoft Entra ID where applicable.
Security, Monitoring, and Cost Optimization
Cloud Security:
Implement and manage robust security policies, including
Identity and Access Management (Microsoft Entra ID, RBAC)
, network security (VNet
, Network Security Groups/NSGs), encryption, and compliance controls across the Azure environment.
Monitoring and Logging:
Configure and manage comprehensive monitoring, alerting, and logging solutions using platform-native tools (e.g.,Azure Monitor and Azure Log Analytics
) to ensure continuous operational health.
Cost Management:
Continuously monitor and analyze cloud spending, recommending and implementing optimization strategies (e.g., resource right-sizing, reserved instances, storage tiering) to ensure cost-efficiency.
Incident Response:
Participate in a 24/7 on-call rotation to respond to, troubleshoot, and resolve critical incidents in a timely manner.