We are seeking a versatile Infrastructure Engineer to act as the guardian of our data integrity and physical backbone. This role manages the "Persistence Layer" of our global healthcare-related suite—ensuring our high-performance storage and network stacks are resilient, secure, and ready for massive scale.
You are a technical "Generalist" who is equally comfortable in a terminal as you are in a data center. You will own the health of our NetApp and Qumulo storage clusters and maintain the network security boundaries that keep our patient data safe.
Key ResponsibilitiesNetwork & Security: Maintain Fortigate firewalls and Cisco network infrastructure, ensuring high availability and secure VLAN management.
Storage Management: Administer high-performance NetApp clusters and high-capacity Qumulo systems.
Data Center Operations: Manage the physical health of our East Coast data centers, including racking, cabling, and hardware lifecycle management.
Collaboration: part of a global team, working closely with off-shore infrastructure engineers. Willing to initiate discussions with development teams early in the cycle.
Incident Response: Participate in a modernized, data-driven incident response process to ensure 24/7 stability for our production environments.
What You BringStorage Depth: 2+ years of experience with enterprise storage (specifically NetApp/ONTAP) including volume management and snapshots.
Systems Generalist: A solid foundation in Linux (RHEL, OL8) and VMware virtualization.
Networking Fundamentals: Comfortable managing firewall rules and switching in a complex, multi-site environment.
Operational Discipline: Experience working within highly regulated frameworks like HIPAA, SOC2, or ARC-AMPE.
Mobility: Ability to travel to our East Coast data center locations for hands-on hardware.
Preferred Skills & MindsetCloud Infrastructure Management: Experience with Oracle OCI is a plus.
Automation-First: Experience with (or a strong desire to master) Ansible, Terraform, and Python is a plus.
Monitoring & Logs: Experience with modern observability tools like Datadog APM and CloudStrike Falcon is a plus.
Problem Solver: You don't just fix the ticket; you find a way to automate the fix so the ticket never returns.
Read LessWe are seeking an Infrastructure Engineer who views infrastructure not just as a collection of hardware, but as a dynamic system to be engineered, automated, and optimized. We are transitioning our Tech Ops culture toward a modern Platform Engineering and SRE model.
This role is ideal for a high-velocity learner who is inherently curious, exceptionally organized, and possesses the discipline to see complex tasks through to completion.
Curiosity First: You have an insatiable desire to understand how things work under the hood. You don't just fix a problem; you investigate the root cause.
Continuous Learner: You view the rapid evolution of technology as an opportunity, not a burden. You are constantly upskilling and bringing new ideas to the table.
Owner’s Mentality: You take pride in "completing the loop." Whether it’s racking a server or writing an Ansible playbook, you ensure the job is documented, compliant, and reliable.
Collaborative Problem Solver: You thrive in a global team environment, communicating across time zones to solve production incidents and share knowledge.
Core ResponsibilitiesEngineering Reliability: Respond to and troubleshoot production incidents. Participate in deep-dive Root Cause Analysis (RCA) and contribute to "Blameless Post-Mortems" to ensure the same issue never happens twice.
Platform Growth and Maintenance: Manage the lifecycle of our Linux environment (primarily RHEL), including OS patching, system upgrades, application expansion, and the automation of these routine tasks.
Compliance & Governance: Maintain rigorous attention to detail to ensure our systems remain compliant with HIPAA, SOC2, and ARC-AMPE standards.
Documentation: Create and maintain high-quality runbooks and infrastructure documentation to empower the global team.
RequirementsEducation: Bachelor’s degree in Computer Science, or a related field, or equivalent experience.
Linux Expertise: Deep, hands-on knowledge of Linux (RHEL or variants).
Availability: Flexibility to collaborate with our offshore teams, requiring occasional off-hours communication.
Technical Versatility: You are an individual contributor with 3-6 years overall experience, with working knowledge in at least two of the following areas:
Storage Engineering: Managing a high-capacity on-prem infrastructure (NetApp or Qumulo preferred) and managing backup/recovery for Oracle databases.
Physical Infrastructure: Precision racking, stacking, and cabling with an understanding of thermal balancing and airflow management.
Network Engineering: Building and expanding multi-site networks using Cisco, F5, and Fortigate hardware, including firewall policy implementation. We’ll consider experience with alternate vendors.
Virtualization: Building and maintaining robust VM farms using VMware ESXi.
The "Modernization" Edge (Pluses)We are looking for candidates who can help us bridge the gap to a fully automated future. Experience in these areas will set you apart:
SRE & IaC: Experience creating SLOs and using Terraform for Infrastructure as Code.
Automation: Building enterprise-grade tools and scripts using Ansible, Python, Bash, and Git flows.
Containers: Knowledge of Docker and/or Kubernetes.
Observability: Moving beyond "monitoring" to true "observability" using the ELK stack, Prometheus, or Grafana to monitor applications (Weblogic, RabbitMQ, Hazelcast).
Cloud Evolution: Experience building and maintaining cloud infrastructure (Network, Compute, Storage) in Oracle OCI or other major cloud providers.
Read Less