Compensation
This employer didn't list pay. Model a likely range with the calculator.
Model this offer in the calculatorJob description
Team: IT
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Principal DevOps Architect based in United States.
This role offers the opportunity to define and shape the future of cloud infrastructure for a global technology platform supporting critical healthcare and research solutions. As a senior technical leader, you will architect scalable, secure, and highly reliable cloud environments while driving modern DevOps practices across the organization. You will combine deep hands-on engineering expertise with strategic architectural vision, influencing teams through technical excellence and innovation. The position focuses on infrastructure automation, AI platform operations, reliability engineering, and compliance-driven cloud solutions. You will play a key role in advancing infrastructure as code, improving developer experiences, and ensuring systems remain secure, efficient, and audit-ready. This is an impactful opportunity for an experienced architect who thrives in complex environments and enjoys solving challenging technical problems.
Accountabilities:
The Principal DevOps Architect will serve as the senior technical authority responsible for designing, implementing, and evolving cloud platforms, DevOps practices, and AI infrastructure. This hands-on individual contributor role will influence engineering standards, improve operational excellence, and enable teams to deliver reliable and secure solutions.
- Own the cloud platform architecture, infrastructure roadmap, deployment strategies, observability practices, and AI platform direction while partnering with engineering leadership.
- Establish engineering standards, reference architectures, and best practices for infrastructure as code, CI/CD pipelines, cloud operations, and AI tooling adoption.
- Design, maintain, and optimize cloud infrastructure using Terraform as the source of truth, including reusable modules, automated deployments, and policy-driven controls.
- Build and operate scalable AWS environments across development, testing, staging, and production while ensuring performance, availability, security, and compliance.
- Develop and improve CI/CD pipelines, release processes, containerized workloads, and deployment automation using technologies such as Docker, Kubernetes, GitHub Actions, and related tools.
- Lead reliability initiatives by defining SLIs, SLOs, error budgets, monitoring strategies, and incident response processes.
- Operate and govern AI/ML platforms, including model-serving infrastructure, AI observability, security controls, cost management, and responsible AI practices.
- Implement security and compliance measures supporting healthcare data protection, audit readiness, secrets management, vulnerability management, and supply-chain security.
- Collaborate with product, engineering, QA, support, and operations teams to improve automation, documentation, and cross-functional delivery.
- Mentor engineers through technical leadership, architecture reviews, and hands-on contributions without direct management responsibility.
- Bachelor’s degree in software engineering, computer science, or equivalent technical experience.
- 10+ years of experience in DevOps, SRE, platform engineering, cloud architecture, or related disciplines, including senior individual contributor or architect-level responsibilities.
- Proven experience designing and operating large-scale cloud environments, preferably on AWS.
- Strong hands-on expertise with Terraform, infrastructure as code practices, reusable modules, automated provisioning, and CI/CD-based infrastructure management.
- Experience building and managing CI/CD pipelines using tools such as GitHub Actions, Jenkins, GitLab, or AWS-native solutions.
- Strong knowledge of Docker, Kubernetes, container orchestration, Linux administration, scripting, and automation.
- Experience with cloud observability, monitoring, troubleshooting, OpenTelemetry, and reliability engineering practices.
- Familiarity with security and compliance frameworks such as HIPAA, HITECH, HITRUST, SOC 2, PCI DSS, CIS Controls, or FedRAMP.
- Experience managing regulated data environments involving PHI, PII, or other sensitive information.
- Strong communication skills with the ability to explain complex technical concepts to both technical and business stakeholders.
- Demonstrated ability to introduce and scale new engineering practices, platforms, or operational improvements across teams.
- Experience with PHP, MySQL, SQL, or similar technologies is a plus.
- Experience operating AI/ML or generative AI platforms in production environments.
- Knowledge of LLMOps practices, including prompt management, evaluation frameworks, AI monitoring, cost attribution, and AI incident response.
- Experience supporting LLM applications, AI agents, or MCP-based services.
- Familiarity with AI security frameworks, OWASP LLM Top 10, NIST AI Risk Management Framework, and key management practices.
- Experience with policy-as-code, secrets management, SBOM generation, image scanning, and software supply-chain security.
- Experience with FinOps practices and optimizing large-scale cloud spending.
- Background in healthcare technology, clinical research, or other highly regulated industries.
- Competitive annual salary range of approximately $158,000–$190,000, depending on experience and qualifications.
- Remote work flexibility within the United States.
- Health insurance coverage.
- Long-term disability and life insurance benefits.
- Unlimited paid time off.
- Paid holidays.
- Paid parental leave.
- 401(k) retirement plan with employer matching contributions.
- Monthly connectivity stipend reimbursement.
- Employee recognition programs and anniversary rewards.
- Opportunity to work on innovative cloud, AI, and healthcare technology solutions.
- Collaborative environment with opportunities for professional growth and technical leadership.
Requirements:
The ideal candidate is a highly experienced cloud and DevOps professional with a strong background in architecture, automation, security, and large-scale platform engineering. They should be comfortable leading technical initiatives, influencing teams, and delivering solutions in regulated environments.
Preferred qualifications include:
Benefits:
Related roles