HomeSearchAi Engineer Jobs › Infrastructure Engineer

Infrastructure Engineer

Princeton University

Princeton, New Jersey, US
More Ai Engineer jobs: Ai Engineer jobsAi Engineer salary

Job at a glance

Princeton, New Jersey, US
Location
Hybrid
Work arrangement
Princeton University
Employer
Ai Engineer jobs
Category

Job Type Full-TimeOverviewThe Princeton DMIA Integration team is seeking an experienced Infrastructure Engineer to be the backbone of our Integration Services team. This role owns the provisioning, configuration, security, and ongoing maintenance of all infrastructure that supports our integration platforms — spanning API integrations, data pipelines, and the shared services that underpin them. You will work across our current on-premises Linux environment and be instrumental in building out the cloud infrastructure as we modernize our platform stack.

This position is a hybrid role.ResponsibilitiesArchitect, Design and DevelopProvision, configure, and maintain all infrastructure supporting integration, and data and platforms including WSO2, and related middleware on on-premises Linux servers.Design and build Azure cloud infrastructure hosting platform using AKS (Azure Kubernetes Service), and other Azure services for the organization’s next-generation API, data, and applications, following infrastructure-as-code (IaC) principles.Manage and harden on-premises server environments including OS patching, package management, service configuration, and security hardening on Linux.Administer IAM integrations including LDAP directory services, Active Directory, service account management, and certificate lifecycle management.Own network configuration relevant to integration workloads: DNS, firewall rules, load balancers, VPNs, and API gateway networking.Collaborate and CoordinateCollaborate on CI/CD pipeline infrastructure, including build agents, artifact repositories, and deployment automation tooling.Support API and data engineering teams with environment provisioning, capacity planning, and performance tuning of integration runtimes.Production SupportParticipate in on-call production support rotation and serve as escalation point for infrastructure-layer incidents.Implement and maintain monitoring, alerting, and logging infrastructure (log aggregation, metrics dashboards, uptime monitoring) for all integration services.Lead disaster recovery planning, backup procedures, and environment restoration for integration platform components.QualificationsAzure cloud infrastructure — compute, networking, storage, identity, securityKubernetes and containerization — Docker, AKS or equivalent managed KubernetesInfrastructure-as-Code — Bicep, Terraform, or ARM; Bicep preferredLinux / shell scriptingPython programmingSQL — sufficient for operational queries and platform toolingNetworking — DNS, HTTP/S, TCP/IP, load balancing, proxies, firewallsIAM — Active Directory, LDAP, OAuth 2.0, SAML, certificate management7+ years provisioning and managing cloud infrastructure on Azure, including network and security configurationDemonstrated experience building or scaling a shared platform that hosts applications for multiple teams — including tenant isolation, resource governance, and self-service or templated onboardingExperience establishing CI/CD pipelines for automated deployment of containerized workloads, using Azure DevOps, Jenkins, GitHub Actions, or equivalentExperience embedding automated security controls into deployment pipelines — container image and vulnerability scanning, policy enforcement, compliance gatingLinux/Unix systems administration with RHEL/CentOS/Ubuntu — systemd, package management, performance tuningAdvanced shell and Python scripting for automation, monitoring, and operational tasksHands-on IAM experience — Active Directory, LDAP, certificate management, and SSO integrationSolid networking fundamentals — TCP/IP, DNS, HTTP/S, load balancing, firewall and proxy configurationWorking knowledge of REST APIs for infrastructure automation and tool integrationExperience with infrastructure monitoring and observability tooling — Prometheus, Grafana, ELK, Azure Monitor, or equivalentAbility to set technical direction and mentor engineers, including documenting standards and transferring knowledge to colleagues newer to modern cloud p

Search all live jobs — free, no account →