Design, build, and maintain secure, scalable cloud infrastructure on Microsoft Azure.
Develop and manage Infrastructure as Code using Terraform for reusable, governed deployments.
Administer and support HashiCorp Vault and Azure DevOps CI/CD pipelines.
We are an international consulting group specializing in innovation and business transformation through technology. With over 7,200 consultants in 21 countries and a turnover of €850M, we are committed to delivering impactful, future-ready solutions.
Support deployment, operation, and reliability of production services on Kubernetes.
Monitor service health, investigate production incidents, and participate in on-call and postmortems.
Troubleshoot application runtime, networking, and service-to-service issues across Node.js and JVM.
Software Mind develops innovative solutions for global companies, partnering with tech giants and unicorns on transformative projects. They foster cross-functional engineering teams with a culture of openness, respect, and passion, combining employment with enjoyment.
Lead, mentor, and grow a team of SRE/DevOps engineers while partnering with engineering leadership to assess team needs and develop talent.
Oversee the incident management process end to end, including on-call rotations, escalation paths, incident command, postmortems, and root cause analysis.
Define and drive SRE principles like SLIs, SLOs, error budgets, capacity planning, and observability standards, championing a culture of reliability and operational excellence.
Eltropy is a rocket ship FinTech on a mission to disrupt the way people access financial services, enabling community financial institutions to digitally engage in a secure and compliant way through a world-class digital communications platform. Their platform integrates Text, Video, Secure Chat, co-browsing, screen sharing, and chatbot technology, bolstered by AI and contact center capabilities, and they value integrity, transparency, and ownership.
Own infrastructure as code and build self-service paths for product teams.
Make reliability a property of the delivery path with SLOs and alerting.
Shift Left security and manage cloud costs as an engineering responsibility.
what3words is a global addressing company that assigns unique three-word addresses to every 3m square on Earth, making locations precise and easy to share. Their technology is used by emergency services, delivery companies, and automakers across 193 countries, with a growing user base and a microservices architecture on AWS.
Build and maintain reliable, scalable, and secure infrastructure solutions for large-scale SaaS applications.
Automate infrastructure provisioning, configuration, deployment, and optimization processes.
Collaborate with R&D teams to improve production stability, reliability, and developer experience.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They use technology to ensure fair and efficient recruitment.
Design and evolve cloud infrastructure on GCP for scale and resilience.
Build internal tooling and automation that promote team autonomy and developer productivity.
Advance observability platform with metrics, logging, tracing, and alerting to reduce recovery time.
The company is a well-funded AI/ML company at the intersection of geospatial intelligence and climate technology, building products on scalable cloud infrastructure. The engineering team fosters a culture of reliability and continuous improvement, operating with a focus on SLOs, error budgets, and DORA metrics.
Design, deploy, and monitor automation services between various business systems.
Collaborate with vendors and engineering teams to deliver reliable and scalable solutions.
Establish and maintain organizational standards, policies, and best practices for DevOps environments.
We are a leading North America network-neutral interconnection and hyperscale edge data center company. Our nearly 2,000 customers are supported by our experienced leadership team, certified staff, and commitment to ESG initiatives.
Provide production support and maintenance for enterprise applications on Google App Engine, ensuring availability and performance.
Lead incident response, troubleshooting, and root-cause analysis for user-impacting issues while meeting SLAs.
Own deployments, feature enhancements, and operational improvements across microservices environments on GCP.
Innodata is a global data engineering company enabling responsible AI advancement through data, evaluation frameworks, and human expertise. With over 36 years of legacy, it delivers high-quality data solutions and services to AI builders and adopters.
Lead centralization of DevOps, SRE, database reliability, incident management, and developer experience practices.
Drive SLOs, observability, alerting, and on-call processes across teams.
Build the platform engineering function from the ground up and influence cross-cutting architecture.
First Due provides fire and EMS agencies with transformative, end-to-end software solutions to improve safety and effectiveness. The company offers a fully remote workplace with a comprehensive benefits package and opportunities for advancement.
Design and implement infrastructure using Terraform, Python, and Kubernetes on AWS.
Collaborate with engineering and data science teams to improve cloud infrastructure.
Automate CI/CD pipelines and enforce security governance and compliance.
Lyra Health is a mental health care provider serving 20 million people through employer and health plan partnerships. The company has delivered 15 million sessions and published 35 peer-reviewed studies, with a culture focused on clinical effectiveness.