Datacenter Operations TechnicianLumos Technology Services
The Datacenter Operations Technician provides advanced support for datacenter infrastructure, including Hyper-V and VMware virtualization, enterprise networking, and SAN storage such as OSNexus QuantaStor. The role leads troubleshooting, lifecycle management, capacity planning, documentation, and mentoring to maintain secure, reliable, highly available services for internal and client environments.
Supervisor: CEO
Travel Required: In office; occasional local travel and after hours support
Datacenter Operations
Provide senior-level operational support for physical and virtual datacenter infrastructure, including compute, storage, networking, power, environmental systems, and supporting management platforms
Monitor infrastructure health, availability, capacity, performance, security alerts, and respond in accordance with established service-level and escalation requirements
Lead diagnosis and resolution of complex hardware, operating system, virtualization, networking and storage incidents, including coordination with manufacturers, carriers, and third-party vendors
Plan and perform installation, configuration, maintenance, upgrades, migrations, moves, additions, changes, and decommissioning of datacenter equipment
Execute preventive maintenance, firmware updates, operating system patching, hardware lifecycle replacements, and configuration reviews in accordance with change-management procedures
Participate in scheduled after-hours maintenance and an on-call rotation as required to support business continuity and customer commitments
Maintain accurate rack elevations, asset records, warranty information, cabling documentation, inventories, and equipment disposition records
Virtualization and Compute
Administer and optimize Microsoft Hyper-V clusters and VMware vSphere environments including hosts, clusters, virtual machines, templates, virtual switches, storage connectivity, and management services
Perform virtual machine provisioning, migrations, resource allocation, snapshot management, backup coordination, replication support, and recovery testing
Troubleshoot cluster, host, guest, storage-path, and virtual-networking issues while minimizing service disruption
Review CPU, memory, storage, and licensing utilization and provide recommendations for consolidation, expansion, refresh, and standardization
Validate backup, restore, high-availability, and disaster-recovery capabilities in coordination with security, backup, and operations teams
Networking
Configure and troubleshoot datacenter switching, VLANs, trunks, routing, port channels, network interface teaming, firewall connectivity, VPNs, DNS, DHCP, and out-of-band management networks
Analyze latency, packet loss, throughput, interface errors, and connectivity issues using appropriate monitoring and diagnostic tools
Maintain secure network segmentation, redundant paths, documented configurations, and configuration backups
Coordinate network changes with internal teams, carriers, vendors, and customers to reduce risk and downtime
Storage and SAN
Administer OSNexus QuantaStor and other SAN or shared-storage platforms, including storage pools, RAID groups, volumes, targets, initiators, snapshots, replication, alerts, and access controls
Configure and troubleshoot iSCSI connectivity, multipathing, host mappings, storage networks, and virtualization-platform integration
Monitor storage capacity, performance, latency, drive health, replication status, and data-protection jobs, and address exceptions before service is affected
Plan storage growth, hardware refreshes, migrations, firmware updates, and maintenance activities while protecting data availability and integrity
Work with backup and disaster-recovery teams to test restores, replication, failover, and recovery procedures
Service Management, Security, and Compliance
Create, manage, and update tickets in the company PSA for incidents, service requests, problems, changes, projects, and maintenance activities
Assign appropriate priority and impact, meet SLA requirements, document troubleshooting and resolution details, and provide timely status updates to stakeholders
Follow access-control, change-management, vulnerability-management, logging, physical-security, and incident-response policies applicable to datacenter operations
Escalate suspected security incidents immediately and preserve relevant system information and logs in accordance with company procedures
Contribute operational evidence, documentation, and corrective actions supporting audits, risk assessments, and compliance requirements
Documentation, Projects, and Leadership
Create and maintain standard operating procedures, diagrams, build standards, recovery procedures, maintenance plans, knowledge-base articles, and customer-specific documentation
Lead or participate in datacenter projects, define technical requirements, identify dependencies and risks, develop implementation and rollback plans, and communicate progress
Perform root cause analysis for significant or recurring incidents and implement corrective and preventive actions
Mentor junior technicians, review technical work and ticket documentation, share knowledge, and assist with skills development
Evaluate products, processes and systems, recommend improvements and contribute to infrastructure roadmaps, budgets, renewals, and capacity plans
Communicate professionally with coworkers, management, clients, and vendors
Other duties as needed
Skills
Advanced administration and troubleshooting of Microsoft Hyper-V, failover clustering, live migration, cluster shared volumes, and Windows Server virtualization environments
Advanced administration and troubleshooting of VMware vSphere, ESXi, vCenter Server, snapshots, templates, virtual networking, high availability, and distributed resource scheduling
Strong enterprise networking knowledge, including TCP/IP, VLANs, routing, switching, link aggregation, firewalls, VPNs, DNS, DHCP, SNMP, and network performance analysis
Hands-on experience with SAN and shared-storage technologies, including OSNexus QuantaStor, iSCSI, Fibre Channel concepts, multipathing, RAID, snapshots, replication, storage pools, and capacity management
Strong knowledge of Windows Server and working knowledge of Linux administration, Active Directory, group policy, PowerShell, patching, and infrastructure monitoring
Experience with server hardware, firmware, racks, cabling, power distribution, environmental monitoring, and datacenter change-control practices
Ability to lead complex incident response, root cause analysis, vendor escalation, maintenance activities, and infrastructure projects
Strong documentation, client communication, prioritization, mentoring, and professional writing skills
Certifications
VMware VCP, Microsoft Windows Server or Azure, OSNexus QuantaStor or equivalent SAN/storage, Cisco CCNA or equivalent networking certification, CompTIA Network+ or Server+, and ITIL Foundation.
Education
Bachelor's in IT related field or equivalent combination of certification and experience.
Physical Requirements
Must be able to lift 50 pounds, install or remove equipment in standard datacenter racks, work in confined equipment areas, use ladders or step stools when required, and remain seated or standing for extended periods. Must be able to wear company-provided safety equipment and a headset as needed.
Lumos Technology Services is a cybersecurity and managed services firm serving small to mid-market businesses across industries such as healthcare, nonprofits, finance, insurance, legal services, energy, and retail. Our mission is to equip and transform small businesses to face the technology challenges of the 21st century, providing a secure and reliable technology foundation. Guided by our core principles of integrity, value, humility, persistence, and excellence, we build custom, values-driven IT solutions and take pride in earning our clients' trust with the systems and data they depend on.
- Information Technology








