Infrastructure Reliability Engineer
- Location
- Jersey City, NJ
- Work type
- Full Time · On-site
- Posted
- 2026-09-03
Job description
We are looking for a hands-on Infrastructure Reliability Engineer to improve reliability, automation, security, and operability across enterprise infrastructure spanning cloud and on-prem/hybrid environments. This role focuses on Terraform + Ansible automation, Windows and Unix/Linux operations, observability, incident/postmortems, and strong vulnerability remediation ownership across servers, platforms, and containerized workloads. The resources should have a minimum of 5 years experience and the business is looking for the resources to have the following skill sets:
Key Responsibilities:
Lead remediation of infrastructure vulnerabilities across Windows, Linux, middleware, and supporting Frontier AI platform components.
Drive closure of control findings, audit items, and cyber remediation commitments within agreed timelines.
Partner with Cybersecurity, Infrastructure, and Application Development teams to identify, prioritize, and remediate vulnerabilities at scale.
Establish sustainable patching, upgrade, and lifecycle management processes to reduce recurring findings.
Vulnerability management tools (Qualys, Tenable, Rapid7, etc. )
Audit and controls remediation experience - preferred
Ability to work across Cyber, Risk, Controls, Infrastructure, and Application teams.
Infrastructure Automation (Terraform + Ansible)
Build and maintain Infrastructure as Code using Terraform across cloud and/or virtualized environments.
Automate configuration, provisioning, patching, and deployments using Ansible across Linux/Unix and Windows estates.
Standardize environments (dev/test/stage/prod), build reusable modules/playbooks, and enforce configuration consistency (prevent drift).