Site Reliability Engineer (SRE) Job at Long Finch Technologies, Jersey City, NJ

Q2liZnpIWFZhcno3V0lQS0lnMnpUV2ljaWc9PQ==
  • Long Finch Technologies
  • Jersey City, NJ

Job Description

Overview

We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.

The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.

 

Key Responsibilities

  • Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
  • Optimize, and support highly available VDI environments on Hyper-V.
  • Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
  • Disaster recovery, backup, patch management, and business continuity strategies.
  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
  • Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
  • Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
  • Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
  • Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
  • Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
  • Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
  • Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
  • Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
  • Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.


Experience & Qualifications

  • 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
  • Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
  • Proven experience implementing automation to reduce operational overhead and improve service reliability.
  • Experience supporting enterprise private cloud and VDI environments.
  • Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
  • Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
  • Experience in Banking or Financial Services environments is advantageous.

    Preferred Skills

    • Windows Server 2016/2019/2022 administration.
    • Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
    • Exposure to hybrid cloud and private cloud platforms.
    • Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
    • Experience supporting enterprise VDI environments.
    • Understanding of ITIL Incident, Problem, Change, and Release Management.
    • Experience working in regulated industries such as Banking or Financial Services.

Job Tags

Contract work

Similar Jobs

Tesla

Tours Coordinator, Megafactory Job at Tesla

 ...What To Expect The Tours Coordinator plays a pivotal role in educating stakeholders about Tesla's innovative products, fostering confidence...  ..., Operations, and beyond Logistics & Guide Leadership: Manage all tour logistics including check-in, AV setup, safety briefings... 

HRT Solutions

Plant Environmental Health and Safety Administrator III Job at HRT Solutions

 ...opportunity to join our EHS team as a Plant Environmental Health and Safety Administrator III. In this exciting role you will be responsible...  ...overall business performance by ensuring compliance with occupational safety and health regulations, environmental regulations... 

Tris Pharma

Manager, Medical Communication Job at Tris Pharma

Tris Pharma, Inc. ( is a leading privately-owned US biopharmaceutical company focused on development and commercialization of innovative medicines in ADHD, spectrum disorders, anxiety, pain and addiction addressing unmet patient needs. We have >150 US and International ...

Texas Hotel Management

Hotel Night Auditor Job at Texas Hotel Management

 ...The Night Auditor is responsible for handling nightly audits, reconciling hotel financial transactions, and providing front desk services during the overnight shift, ensuring both accounting accuracy and excellent guest service at the hotel. Key Responsibilities:... 

Cipla

Assistant Documentation Specialist Job at Cipla

Job Title Assistant Documentation Specialist Location Central Islip or Hauppauge NY locations Employment Type Hourly - Full Time Hourly Range $20.00-$26.00 Work Hour / Shift 1st shift 7:00 AM - 3:30 PM Responsibilities/ Accountabilities Prepare and review Change Controls...