Senior AI Infrastructure Engineer
Enregistrez cette offre et organisez votre recherche
Créez un compte gratuit pour enregistrer des offres d'emploi, créer des alertes et revenir à cette liste depuis votre tableau de bord.
En continuant, vous acceptez nos Conditions d’utilisation & Politique de confidentialité.
About Vega Health
We source, implement, and monitor leading Al solutions aligned to the priorities of health systems. As part of our team,you’llplay a key role in achieving our mission: To be the objective, trusted partner equipping health systems everywhere as they scale AI solutions that improve care and operational outcomes.
At Vega Health, each team member leads with integrity, trust, and accountability and works with humble ambition, collaboration, and curiosity. Please visit vegahealth.com for more information about our story and the DNA of how we approach this work.
About the Role
This roleisresponsible for buildingand designingthe systems that will enable Vega Health to scale as an enterprise and help our customers be at their best. This role reports directly to the Director of Engineering and works with the Product team at Vega Health.
At large, this role at Vega Health will provide technical leadershipregardingour infrastructure and platform delivery both internally and toour customers. This roleis responsible fordefining architecture, guiding engineering standards, and delivering scalable softwareinfrastructure as code, and analytics capabilities acrossfully automated deployments. The successful candidate will partner with product, business, and engineering stakeholders to modernize platform capabilities, enable data-driven insights, and ensure long-term platform reliability and performance.
Relocation to Durham, North Carolina, is notrequiredfor this role; however, some travel may be expected to create structured in-person time with the team. We also expect new hires to spendone weekon-site in Durham during theinitialonboarding period.
Key Responsibilities
Design and implement scalable, secure, and cost-efficientcloud infrastructureusing automation tools.
Design, develop, and maintain production software services and tools (APIs, internal dashboards, customer onboarding/evaluation tools) alongside infrastructure work.
Develop and maintain Infrastructure as Code using tools such as Terraform, Helm, CloudFormation or equivalent technologies to ensure cloud environments are version controlled, auditable and consistently deployed.
Build, maintain and improve CI/CD pipelines to automate build, test, security validation and deployment processes for infrastructure and cloud-hosted applications, using tools such as Azure DevOps, GitHub Actions, GitLab CI or equivalent platforms.
Deploy and operate containerized workloads and cloud-native services using technologies such as Docker, Kubernetes, AKS, EKS or GKE, where appropriate to thesolution architecture.
Follow cloud architecture standards, policies, and guidelines to maintain consistency and reliability across teams and ensure compliance with data privacy and regulatory requirements (e.g., HIPAA, GDPR, FedRAMP).
Serve as senior escalation for complex cloud, infrastructure, and security issues
Implement cloud security controls including IAM, network segmentation, logging, and compliance baselines
Instrument systems for monitoringand platform observability(metrics, logging, tracing) to support proactive monitoring and incident response(e.g.Prometheus, Grafana, Datadog, etc.).
Work closely with application, infrastructure, securityand operations to translate business and technical requirements into well-architectedfully automatedsolutions.
Contribute to a platform engineering approach by developing reusable templates,self-service deployment patterns and standardized cloud services that enable application teams to deploy solutions consistently and efficiently.
Troubleshoot and resolve technical issues across platforms, deploymentpipelinesand production environments, ensuringtimely restoration of service and continuous improvement of solution reliability.
Design scalable systems, processes, and playbooks that remove operational friction internally and enable the organization to grow efficiently (e.g., developingcollateral,installation and operationaldocumentation,andsoftware to automate service and infrastructure deployment)
Interface with Vega Health internal teams or seniorcustomer ITstaff when addressing escalatedtechnical and implementationissuesrelated tomaterials designedandoperatedwithin this function.
Me