Back to search
Larsen & Toubro Linkedin · Posted 22d ago

CloudOps Engineer (L1)

Mumbai

Linkedin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

Job Purpose

The Frontline Infrastructure Support Engineer (L1) serves as the first point of contact for infrastructure-related incidents and service requests. The role is responsible for continuous monitoring of enterprise IT infrastructure, performing initial diagnostics, executing predefined operational tasks, restoring services using standard operating procedures (SOPs), and escalating unresolved issues to L2/L3 teams while ensuring SLA compliance.

Roles & Responsibilities

Infrastructure Monitoring

  • Continuously monitor the health, availability, and performance of servers, virtualization, Kubernetes, storage, backup, network, and GPU infrastructure using enterprise monitoring tools.
  • Detect alerts, perform initial health checks, validate service availability, and initiate incident notifications as per operational procedures.

Incident Management

  • Log, categorize, prioritize, and troubleshoot Level 1 incidents using SOPs and knowledge articles while ensuring SLA compliance.
  • Restore services where possible and escalate unresolved or critical incidents to L2/L3 teams with complete documentation.

Windows & Active Directory Support

  • Perform user account administration, password resets, account unlocks, and basic Windows server health verification.
  • Validate Windows services and escalate server or Active Directory issues requiring advanced administration.

Linux Operations Support

  • Monitor Linux server availability, service status, and system logs to identify operational issues.
  • Execute approved service restarts and escalate operating system issues beyond standard support procedures.

Virtualization Support

  • Monitor virtual machine availability, health, and resource utilization across virtualization platforms.
  • Perform basic VM recovery activities and escalate hypervisor or platform-related issues to specialized teams.

Kubernetes Operations Support

  • Monitor Kubernetes clusters, nodes, pods, and services to ensure platform availability.
  • Perform approved pod/service restarts and escalate cluster or orchestration-related issues.

Storage & Backup Support

  • Monitor storage health, capacity, and backup job execution to ensure operational continuity.
  • Perform authorized backup recovery actions and escalate storage or backup infrastructure failures.

Network & Infrastructure Diagnostics

  • Perform basic network diagnostics, connectivity verification, and infrastructure service validation.
  • Monitor network health and escalate complex connectivity or performance issues to network support teams.

GPU Infrastructure Support

  • Monitor GPU server health, utilization, and hardware alerts to maintain operational readiness.
  • Coordinate hardware replacement activities and vendor support for GPU-related incidents.

Hardware Support

  • Perform basic hardware health checks and assist in diagnosing infrastructure component failures.
  • Support hardware replacement activities, post-maintenance validation, and asset record updates.

Documentation & Reporting

  • Maintain accurate incident records, operational logs, and shift handover documentation.
  • Follow SOPs, contribute to knowledge management, and report recurring issues for continuous service improvement.

Experience & Educational Requirement

BE/B-Tech or equivalent with Computer Science or Electronics & Communication

Relevant Experience

  • 1–3 years in an IT Infrastructure Operations, NOC, Service Desk, or Data Center Operations environment.
  • 24×7 support operations, enterprise monitoring tools, ITSM platforms (ManageEngine, ServiceNow, BMC Remedy, Jira), and basic cloud or virtualization environments.
  • 1–3 years of hands-on experience in Infrastructure Monitoring, Incident Management, Windows/Linux Administration, Active Directory, Virtualization (VMware/Hyper-V), Kubernetes, Storage & Backup Monitoring, Network Diagnostics, and ITSM tools (e.g., ServiceNow).
  • 1–3 years of experience in Hardware Support, Documentation, ITIL/SOP adherence, Shift Operations, Asset & Vendor Coordination, Customer Support, and Operational Reporting
Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search