**Lead a Pod. Own the Roadmap. Keep Customers Running:** Clearscale, a leading AWS Premier Consulting Partner, empowers businesses with comprehensive cloud solutions, spanning migrations, application modernization, data and analytics, AI/ML, IoT, managed cloud, and strategic consulting. Our remote\-first team of experts has a decade\-long track record of guiding Fortune 500 enterprises, mid\-market businesses, and innovative startups across diverse industries to achieve ambitious cloud initiatives. Our Managed Services practice distinguishes itself from traditional project delivery models. We integrate specialized, fractional pods of DevOps and security engineers into long\-term customer engagements characterized by open\-ended, evolving scopes. The DevOps Team Lead serves as the architect of clarity, transforming ambiguity into actionable roadmaps, structured backlogs, and consistent, successful sprint cycles. We are seeking a technical leader who is adept at facilitating strategic sprint planning, justifying architectural decisions, and managing complex incident resolutions while leveraging AI\-driven tooling to accelerate delivery, enhance operational efficiency, and reduce manual toil. **What You'll Do:** * Lead a pod of DevOps and security engineers delivering ongoing engineering across multiple long\-term customers under a fractional delivery model, balancing engineer allocation and assignments across accounts. * Own continuous scoping, roadmap development, and technical assessments for your customers, converting open\-ended engagements into a prioritized, well\-defined backlog. * Scope and oversee delivery of both customer\-driven and Clearscale\-driven initiatives, proactively identifying and proposing improvements customers haven't asked for yet. * Facilitate weekly sprint planning and prioritization across accounts, and hold the pod accountable to sprint commitments. * Own customer incident escalations, coordinate escalated engineering response, and communicate status to the 24x7 Operations team and customers in real\-time * Partner with Service Delivery to ensure customer alignment, on\-time delivery, and sustained customer satisfaction, including participation in customer QBRs and roadmap reviews. * Build and use AI tooling to accelerate planning, delivery, and management tasks as well as hands\-on engineering work, and drive adoption of that tooling across the pod. * Establish and champion best practices for infrastructure\-as\-code, CI/CD, monitoring, logging, alerting, and cloud security across every account the pod supports. * Lead the design, implementation, and management of scalable, resilient, and secure AWS infrastructure for customers with widely varying maturity levels. * Mentor and grow DevOps and security engineers, fostering technical development and a culture of knowledge sharing across accounts. * Evaluate and introduce new tools and technologies that improve reliability, security posture, and delivery velocity. * Maintain technical documentation, runbooks, and operational procedures so knowledge lives with the pod, not with individuals. **What You'll Bring:** * 7\+ years of hands\-on DevOps engineering experience with a strong focus on automation and infrastructure\-as\-code on AWS. * 3\-5 years in a team lead or technical leadership role within a DevOps, SRE, or managed services function. * Demonstrated experience leading delivery for multiple concurrent customers or accounts, including scoping work in open\-ended or time\-and\-materials engagements. * AWS Certified Solutions Architect – Professional certification, or the ability to obtain it within 60 days of start. * Anthropic Architect Associate certification within 90 days of start (we'll support you in getting there). * Deep understanding of CI/CD pipelines and experience implementing and managing them using tools like Jenkins, GitLab CI, GitHub Actions, or AWS CodePipeline. * Extensive experience with infrastructure\-as\-code tools such as Terraform, CloudFormation, or CDK. * Strong experience with containerization and orchestration (Docker, Kubernetes, ECS/EKS). * Proven ability to design and implement monitoring, logging, and alerting solutions using tools like CloudWatch, Prometheus, Grafana, Datadog, ELK, or Splunk. * Solid grounding in cloud security best practices, networking principles, and infrastructure management. * Excellent scripting skills in Python, Bash, or Go. * Strong problem\-solving and troubleshooting skills, with the ability to lead incident response and drive root\-cause remediation. * Excellent communication and stakeholder management skills. You can hold a technical design discussion and a customer escalation call with equal confidence. * Practical experience using AI tooling in day\-to\-day engineering or delivery management work. **Ideally, You Also Have:** * Additional AWS Associate and Professional certifications * Experience in a managed services, MSP, or consulting environment with recurring