SRE / Infrastructure Engineer (LABGEN)

Medfar

$110K — $130K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years of experience in Site Reliability Engineering, Infrastructure, DevOps, or similar roles, preferably in SaaS or regulated environments.
  • Hands-on administration experience with Linux (RHEL, CentOS, Ubuntu) and Windows Server.
  • Proficiency in Apache web server administration and high-availability Linux environments.
  • Strong knowledge of TCP/IP, DNS, VPNs, and firewalls, with experience in network segmentation.
  • Familiarity with cloud technologies: Azure, AWS, or GCP, along with disaster recovery and backup management.
  • Solid understanding of SQL and relational databases, including performance analysis and query troubleshooting.
  • Awareness of regulatory standards like ISO 27001, HIPAA, and HITECH is a plus.

Responsibilities

  • Manage lifecycle tasks for Windows and Linux production servers including provisioning and decommissioning.
  • Design high-availability configurations to prevent single points of failure.
  • Monitor critical metrics such as performance, uptime, and infrastructure costs.
  • Implement and maintain security controls and compliance requirements across environments.
  • Own and monitor production backup policies, ensuring integrity and recovery procedures.
  • Coordinate production deployments and manage application hosting infrastructure.
  • Support infrastructure modernization and evaluate options for cloud migration.

Benefits

  • Generous health, vision, and dental group insurance coverage after probation.
  • Two weeks of paid time off.
  • Dynamic and multicultural work environment promoting innovation and efficiency.
  • Strong focus on entrepreneurial culture in the workplace.
  • Accessibility by public transit, with a convenient location near Long Island Rail Road.
Full Job Description
Job Description

You will work on the Labgen environment. We are looking for a hands-on Site Reliability Engineer / Infrastructure Engineer to own the reliability, security, and continuity of Comtron's Windows- and Linux-based production environment.

We are looking for a hands-on Site Reliability Engineer / Infrastructure Engineer to own the reliability, security, and continuity of Labgen's Windows- and Linux-based production environment.

You will be responsible for three primary areas:
  • Production infrastructure, availability, and security
  • Backups and disaster recovery
  • Production deployments and hosting environments

You will also support infrastructure modernization, technology upgrades, and the potential transition of applicable systems to public or hybrid-cloud environments.

The role reports to the Software Development Manager and works closely with the IT Director, Software Development, Quality Assurance, Security, Support, and Finance teams.

Key Responsibilities

1) Infrastructure and Reliability
  • Manage the lifecycle of Windows and Linux production servers, including provisioning, configuration, patching, hardening, and decommissioning.
  • Design and maintain high-availability configurations to reduce single points of failure and ensure continuous access to the Labgen platform.
  • Monitor uptime, performance, capacity, resource utilization, and infrastructure costs.
  • Maintain production access controls, firewall rules, certificates, networking, storage, and database connectivity.
  • Maintain network segmentation between production, development, and corporate environments.
  • Coordinate vulnerability assessments, penetration testing, and remediation activities.
  • Manage privately hosted infrastructure and relationships with hardware, software, connectivity, and data center vendors.
  • Support infrastructure purchasing, contract renewals, licensing, and capacity planning.
  • Evaluate infrastructure tools and public or hybrid-cloud solutions, including Azure, AWS, or GCP.

2) Backups and Disaster Recovery
  • Own production backup policies, including scope, frequency, retention, monitoring, and integrity validation.
  • Ensure backups are completed successfully and recovery procedures are regularly tested and documented.
  • Establish and maintain recovery time and recovery point objectives.
  • Maintain and regularly test disaster recovery procedures for hosted systems.
  • Lead improvements following disaster recovery exercises and production incidents.

3) Deployments and Technology Lifecycle
  • Manage staging, UAT, and production hosting environments.
  • Execute production deployments and rollback procedures in coordination with the Software Development team.
  • Coordinate deployment windows, infrastructure changes, and change-management activities.
  • Manage application hosting infrastructure, including Windows Server, Apache, application services, scheduled processes, networking, storage, certificates, and database connectivity.
  • Support CI/CD tooling, deployment automation, and pipeline reliability.
  • Plan and coordinate operating system, database, framework, library, and application dependency upgrades.
  • Assess cloud migration options, application dependencies, risks, and phased modernization strategies.
  • Support proof-of-concept initiatives and ensure proposed solutions meet security, reliability, scalability, compliance, and disaster recovery requirements.


Qualifications

What we're looking for:
  • 5+ years of experience in Site Reliability Engineering, Infrastructure, DevOps, or a similar role, ideally within a SaaS or regulated environment.
  • Strong hands-on Linux administration experience, including RHEL, CentOS, Ubuntu, or similar environments.
  • Experience administering Windows Server and application-hosting environments.
  • Strong Apache web server administration experience.
  • Experience designing or managing high-availability Linux production environments.
  • Strong knowledge of TCP/IP, DNS, VPNs, firewalls, certificates, and network segmentation.
  • Experience with Azure, AWS, GCP, or hybrid-cloud environments.
  • Experience with enterprise backups, disaster recovery, and production incident management.
  • Strong working knowledge of SQL and relational databases, including query troubleshooting, performance analysis, migrations, upgrades, and connectivity issues.
  • Familiarity with .NET and JavaScript-based application environments.
  • Experience with version control, build and release processes, deployment automation, and CI/CD tools such as Git, CVS, and Jenkins.
  • Experience with access controls, system hardening, firewall management, vulnerability remediation, and infrastructure security.
  • Familiarity with ISO 27001 controls and evidence requirements.
  • Familiarity with HIPAA, HITECH, ONC Health IT Certification requirements, or other U.S. healthcare privacy and security obligations is an asset.
  • Strong incident ownership, documentation, and cross-functional communication skills.
  • Strong written and verbal communication skills in English.

Nice to Have

  • Experience working in regulated healthcare environments.
  • Familiarity with laboratory information systems, electronic medical records, or clinical information systems.
  • ISO 27001 Lead Implementer or Auditor certification.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, ELK, Datadog, or SentinelOne.
  • Python or Bash scripting experience.
  • French language proficiency.
  • Familiarity with ISO 27001 controls and evidence requirements.
  • Familiarity with HIPAA, HITECH, ONC Health IT Certification requirements, or other U.S. healthcare privacy and security obligations is an asset.
  • Strong incident ownership, documentation, and cross-functional communication skills.
  • Strong written and verbal communication skills in English.


Additional Information

Joining Comtron means embarking on a dynamic environment where trust, innovation, quality, and client success guide our days. At Comtron, we promote efficiency and excellence in healthcare with a powerful LIS solution and a fully customizable web-based EHR system, with a complete suite of Revenue Cycle Management services.
  • Entrepreneurial culture;
  • Dynamic and multicultural work environment;
  • Generous health, vision, and dental group insurance coverage (after probation);
  • 2 weeks of paid time off

Our office in Great Neck (NY) is accessible by public transit and is a short walk from Great Neck station (Long Island Rail Road).

Similar Jobs

More Jobs at Medfar

More Information Technology Jobs

Find similar SRE / Infrastructure Engineer (LABGEN) jobs: