Drive the performance, stability, security, and reliability of production environments with a focus on automation and proactive improvements.
Design and maintain infrastructure using Infrastructure as Code tools like Terraform, and manage Kubernetes and cloud environments.
Lead vulnerability management, incident response, and secure CI/CD practices to ensure resilience and operational excellence.
Jobgether is a platform that uses AI-powered matching to connect candidates with hiring companies. It processes applications and shares shortlists with employers, offering a remote-first and inclusive work environment.
Lead the discovery, design, and delivery of complex reliability and infrastructure initiatives, translating ambiguous problems into robust technical solutions.
Define and operate reliability practices including SLOs, SLIs, error budgets, alerting strategies, and observability standards.
Manage and scale production Kubernetes environments, build cloud infrastructure on AWS, and mentor less-senior engineers.
Jobgether is an AI-powered recruitment platform that connects candidates with partner companies through automated matching. They are a globally distributed organization focused on fair, efficient hiring processes.
Build and operate the Kubernetes platform supporting AI test and evaluation frameworks.
Design infrastructure-as-code, GitOps workflows, and automated deployment pipelines.
Own platform reliability, observability, capacity planning, and operational readiness.
OpenTeams helps enterprises and governments build AI they control, govern, and evolve themselves. Founded by the creator of NumPy and SciPy, the company is built by people with deep roots across the open-source ecosystem and maintains a remote-first culture.
Ensure reliability, scalability, and performance of cloud-based systems using Kubernetes and observability tools.
Define and monitor reliability metrics (SLIs, SLOs, MTTR) to continuously improve operational performance.
Automate operational tasks and implement Infrastructure as Code to reduce manual work and enhance efficiency.
Our partner is a technology company focused on building and maintaining reliable, scalable digital environments. They promote a culture of continuous improvement, collaboration, and proactive engineering.
Lead cloud infrastructure strategy for resilient, secure, and cost-efficient multi-account cloud environments across AWS, Azure, and GCP.
Drive Kubernetes excellence as a technical authority for production clusters including EKS and AKS.
Advance AI-enabled operations by introducing LLM-based tooling and agentic workflows to improve infrastructure development and operational efficiency.
Jobgether is a platform that uses AI-powered matching to connect candidates with hiring companies. They are a technology company focused on improving the hiring process through automation and data analysis.
Define and monitor reliability metrics such as SLI, SLO, SLA, MTTR, and MTTD.
Implement observability solutions including monitoring, alerting, dashboards, and APM.
Collaborate with multidisciplinary teams to embed reliability and observability into solutions.
The company is a technology organization focused on building and maintaining reliable digital environments. It fosters a culture of engineering excellence, collaboration, and data-driven decision-making.
Manage and troubleshoot complex distributed large-scale software systems
Build scalable, secure and reliable container-based infrastructure
Automate software delivery processes with CI/CD pipelines
Coinspaid Dev is the engineering brand behind the technology, infrastructure, and R&D expertise built within Coinspaid, focusing on advancing blockchain infrastructure engineering. With over 120 engineers and more than 11 years of industry experience, they bring together teams building distributed systems and blockchain infrastructure across 20+ blockchain networks.
Design and evolve highly available cloud architecture on Google Cloud using Terraform and GitOps.
Build and maintain secure CI/CD pipelines for Infrastructure-as-Code and develop self-service developer platforms.
Strengthen platform observability and apply SRE principles to improve reliability and operational maturity.
They are a financial technology company that provides production-critical infrastructure. They have a globally distributed team and a culture of autonomy and async-first collaboration.
Maintain and extend foundational technologies to provide highly available, distributed services to customers.
Design, implement, and operate large scale, multi-cloud Kubernetes infrastructure driven by GitOps methodology.
Embrace an empathetic, supportive, and communicative environment, pulling from one another's strengths.
InfluxData is the creator of InfluxDB, the leading time series platform for collecting, storing, and analyzing time series data at any scale. They are a remote-first company with a globally distributed workforce.
Operate and evolve AWS infrastructure for Data/AI platforms, ensuring security, scalability, and high availability.
Build CI/CD pipelines, automate provisioning with Terraform, and implement observability.
Collaborate with Data, AI, and Infrastructure teams, document standards, and drive platform improvements.
The partner company is building a modern Data Platform team focused on secure, scalable, and highly available infrastructure for Data and AI workloads. The culture emphasizes DevOps, automation, and continuous improvement, with close collaboration across Data, AI, and Infrastructure teams.
Own and drive key infrastructure modernization initiatives toward container-orchestrated infrastructure.
Design and maintain infrastructure as code across multiple cloud providers.
Provide technical leadership and mentorship across the Systems Engineering team.
Intellum is the leader in corporate education technology, powering large learning programs for brands like Google, Meta, and Amazon. We are a remote-first company with a culture that values curiosity, creativity, perseverance, and kindness, and we invest in our people through personal development budgets and annual retreats.
Serve as the technical backbone of cloud infrastructure operations, bridging incident detection and advanced architecture.
Build and maintain CI/CD pipelines, design IaC modules, and optimize cloud resources for performance and cost efficiency.
Lead observability initiatives, integrate DevSecOps practices, and collaborate with cross-functional teams to ensure robust cloud reliability.
CodeRoad provides end-to-end software development services, helping businesses scale with ideal infrastructure solutions. They operate with a nearshore model and focus on empowering businesses through staff augmentation, dedicated teams, and software engineering.
Own and evolve Quansight's cloud infrastructure across AWS, Azure, and GCP.
Lead infrastructure engagements for clients from scoping through delivery.
Contribute to open-source projects and participate in upstream communities.
Quansight is rooted in the Python data science community and helps companies build sustainable solutions on open-source software. The team is a small, collaborative, fully distributed group of open-source maintainers and engineers.
Architect, deploy, and manage highly available, fault-tolerant cloud infrastructure across Google Cloud Platform (GCP) and Google Kubernetes Engine (GKE).
Maintain and scale declarative infrastructure using Terraform across a multi-hundred-file estate, enforcing GitOps workflows with Atlantis.
Build, maintain, and optimize robust automated pipelines for continuous integration and delivery using GitHub Actions, Jenkins, and ArgoCD.
Point Wild helps customers monitor, manage, and protect against the risks associated with their identities and personal information in a digital world. Backed by WndrCo, Warburg Pincus and General Catalyst, Point Wild is a scrappy, nimble organization dedicated to creating the world’s most comprehensive portfolio of industry-leading cybersecurity solutions.
Support management and maintenance of Google Cloud Platform infrastructure.
Assist in ensuring reliability, scalability, and performance of cloud services.
Contribute to GitLab CI/CD pipelines, work with Kubernetes, and write automation scripts.
Miratech is a global IT services and consulting company that helps visionaries change the world through digital transformation. With nearly 1,000 professionals across 25 countries, the company maintains a culture of Relentless Performance with a 99% project success rate and over 30% year-over-year growth.
Architect and automate platforms to enable fast, secure code delivery with CI/CD integration.
Own the GKE/Kubernetes infrastructure on Google Cloud, ensuring scalability and reliability.
Proactively engineer security, eliminating bottlenecks and driving developer experience.
Smartclip is an ad tech company that builds a platform for developers to deploy code quickly and securely. They foster a culture of ownership, fast iterations, and minimal bureaucracy, with a remote-first approach and on-site meetings in Berlin.
Empower engineers on other teams by maintaining monitoring tooling and collaborating on observability best practices.
Enhance reliability of Kubernetes applications through resource optimization, streamlined upgrades, and scalability.
Participate in on-call and incident response processes, occasionally diving into application code to debug production issues.
Webflow is the agentic web marketing platform for modern marketing teams, helping organizations build, manage, and optimize high-performing web experiences. It serves over 2 million users worldwide across 190 countries, with tens of thousands of projects launched each month, and fosters a culture of grit, speed, and craft.
Design and advance core infrastructure for multi-cloud Kubernetes clusters and developer toolchains.
Automate operations and engineering tasks to improve productivity and reliability.
Build machine learning infrastructure to enable AI teams to train and deploy large-scale models.
Cresta provides an AI platform that transforms customer conversations into competitive advantages by combining conversational AI, real-time agent augmentation, and conversation intelligence. The company has raised over $270 million from top investors like a16z, Greylock, and Sequoia, and is led by AI industry veterans.
Design and evolve scalable cloud infrastructure on Google Cloud Platform, focusing on reliability and automation.
Strengthen observability platform with metrics, logging, and tracing to improve incident response and reduce recovery time.
Champion reliability practices like SLOs, error budgets, and DORA metrics to drive operational excellence.
They operate at the intersection of geospatial intelligence and environmental technology. They are a growing organization with a collaborative, high-impact engineering culture.