Design, build, and maintain robust, scalable, and secure infrastructure systems supporting Laurel's AI-driven platform.
Manage and optimize cloud infrastructure (AWS and Azure), Kubernetes orchestration, and CI/CD pipelines to increase deployment frequency and reliability.
Implement comprehensive observability, monitoring, and alerting to maintain system health and partner with engineering teams to optimize performance and cost-efficiency.
Laurel is an AI Time platform for professional services firms, automating work time capture and connecting time data to business outcomes for clients like EY and Crowell & Moring. The company comprises top AI, product, and engineering talent, is VC-backed by Google Ventures and IVP, and fosters an inclusive, ambitious culture.
Architect, build, and operate secure, multi-account AWS environments using modern Infrastructure as Code.
Design and optimize container orchestration with AWS ECS (Fargate) and Kubernetes (EKS) for specialized workloads.
Build unified CI/CD pipelines with GitHub Actions, embed security by design, and establish observability with Prometheus, Grafana, and CloudWatch.
Berlitz is a global language education company with a history of nearly 150 years, currently undergoing a digital transformation. The company fosters a remote-first, AI-native culture with a focus on autonomy and greenfield development, though employee count is not specified.
Lead the SRE strategy and execution for a high-growth AI company.
Build and scale a high-performing SRE team while defining reliability standards.
Architect secure, scalable cloud infrastructure and implement observability practices.
This company develops advanced AI products and agentic technology. It operates in a high-growth, international environment with a focus on operational excellence and innovation.
Build and configure GKE clusters and Google Cloud project environments using Terraform.
Implement and maintain CI/CD pipelines and environment promotion workflows.
Configure and maintain Google Cloud-native monitoring, alerting, and logging.
We design, build, and scale AI-powered solutions that create real business impact. We are building a high-performance culture grounded in five values: Empowering Excellence, Collaborative Teamwork, Unsolicited Respect, Consistent Transparency, and Efficient Communication.
Design, maintain, and support secure AWS environments across compute, storage, networking, account structure, and operational practices.
Administer Azure-hosted applications, data services, and supporting infrastructure including tenant configuration and access controls.
Implement and maintain infrastructure as code using tools such as Terraform or CloudFormation to improve consistency and reliability.
JerseySTEM is a mission-driven professional network of pro-bono contributors dedicated to improving access to STEM education and career pathways for underserved middle school girls in New Jersey. Members contribute their professional skills and leverage their networks in service of the organization's gender-equity agenda, operating remotely with a small volunteer base.
Serve as the technical backbone of cloud infrastructure operations, bridging incident detection and advanced architecture.
Build and maintain CI/CD pipelines, design IaC modules, and optimize cloud resources for performance and cost efficiency.
Lead observability initiatives, integrate DevSecOps practices, and collaborate with cross-functional teams to ensure robust cloud reliability.
CodeRoad provides end-to-end software development services, helping businesses scale with ideal infrastructure solutions. They operate with a nearshore model and focus on empowering businesses through staff augmentation, dedicated teams, and software engineering.
Design, build, and operate multi-region AWS infrastructure on Kuberneties with Terraform and Helm at 15PB+ scale.
Own high-availablity, event-driven architectures and cost optimization across the stack.
Drive developer experience, security, and incident response as the second platform team member.
ScorePlay is the AI-powered media infrastructure for sports, automating content operations for the world's biggest sports organizations. We are a 50-person remote-first team based in New York and Paris, growing 2x year over year with 98% retention.
Manage Kubernetes clusters using Rancher RKE2 and configure Calico CNI for networking.
Implement CI/CD pipelines with Jenkins, Terraform, and Ansible, integrating tools like Vault and Artifactory.
Ensure backup, recovery, and disaster recovery across hybrid cloud and on-prem infrastructure.
The company is hiring a Senior Kubernetes Engineer for an on-premises failover environment and critical application onboarding across clearing, settlement, and risk platforms. The team size and culture are not specified in the posting.
Build and maintain scalable, reliable, and secure environments on AWS using Infrastructure as Code tools.
Design and manage CI/CD pipelines, oversee Kubernetes clusters, and ensure GitOps practices.
Monitor system health with OpenTelemetry and Grafana, enforce security best practices, and mentor junior engineers.
Deutsche Telekom IT Solutions is a subsidiary of the Deutsche Telekom Group, providing IT and telecommunications services with over 5,300 employees. Recognized as Hungary's most attractive employer, it serves large corporate clients across Europe.
Design, implement, and manage secure cloud infrastructure for mission-critical applications.
Apply infrastructure-as-code practices using tools like Terraform or CloudFormation.
Collaborate with development and operations teams to optimize performance and reliability.
They provide cloud engineering and cybersecurity solutions for critical national security missions. They are a mission-focused, equal opportunity employer with a fully remote team.