Design, develop, and maintain cloud and containerized platforms on AWS and Kubernetes.
Automate deployments, infrastructure, and monitoring using Terraform, GitHub Actions, and GitOps.
Mentor junior engineers and lead technical projects for medium to large-scale initiatives.
New Era Technology securely connects people, places, and information with end-to-end technology solutions at scale. With a global team of over 3,000 professionals, we foster a team-oriented culture that prioritizes personal and professional development.
Migrate and modernize critical infrastructure from legacy Beanstalk to EKS, managing networking, IAM, and functional parity.
Design and maintain compute and networking components such as Private Link, load balancers, and Core API used by multiple teams.
Ensure high availability and resilience of shared compute infrastructure through on-call rotations and cross-team collaboration.
VTEX is a composable and complete commerce platform that empowers brands, distributors, and retailers with flexibility and comprehensive solutions. With over 1,300 employees across 16 countries, VTEX fosters a challenge-driven environment and collaborative culture.
Lead the strategic direction, engineering, and operational management of enterprise cloud and infrastructure platforms.
Own the cloud operating model, including service catalog, SLAs/SLOs, governance, and cost optimization.
Build and develop a team of cloud platform engineers, SREs, and security engineers, driving reliability and automation.
Banner Health is one of the largest nonprofit health care systems in the country, delivering hospital services and advanced technology. They offer a comprehensive benefits package and a culture centered on making health care easier for every employee.
Own the observability, logging and alerting for Kubernetes clusters and critical workloads.
Build and maintain automation for lifecycle management of Kubernetes clusters.
Identify and root-fix reliability bottlenecks before they become incidents.
Wrapbook is an AI platform for production finance, built for feature films and TV, trusted by Netflix and Paramount. Backed by top investors, our team of over 350 employees uses AI to transform how finance teams work.
Lead Cloud Platform and SRE teams to scale securely and efficiently.
Drive infrastructure strategy, including Kubernetes (GKE) clusters and developer platform.
Champion SRE culture with SLOs, error budgets, and observability.
Prolific builds human data infrastructure for AI development. The company is a fast-growing, mission-driven organization with cross-functional teams and a strong ownership culture.
Own the technical architecture and evolution of core infrastructure.
Engineer for scale and performance through capacity modeling and bottleneck diagnosis.
Participate in on-call rotation and drive technical recovery during incidents.
Sezzle is a fintech company that revolutionizes shopping through interest-free installment plans, blending cutting-edge technology with financial empowerment. It has a dynamic and innovative team culture focused on shaping the future of fintech and retail.
Lead end-to-end technical engagements: Partner directly with engineering teams to diagnose, unblock, and resolve complex infrastructure challenges.
Execute critical migrations: Develop reference implementations, tooling, and guidance to transition teams off deprecated systems seamlessly.
Accelerate platform adoption: Act as primary technical contact for new teams onboarding to Planet's core infrastructure.
Planet designs, builds, and operates the largest constellation of imaging satellites in history, delivering unprecedented dataset via a cloud-based platform for commercial, environmental, and humanitarian sectors. A global company with offices in the US, Europe, and Slovenia, Planet values a people-centric culture and community.
Define DevOps strategy and lead infrastructure architecture across multi-environment, multi-region cloud systems.
Architect and own scalable Kubernetes platforms, infrastructure as code, and DevSecOps implementation.
Drive platform reliability, performance SLAs, cost optimization, and lead complex migrations and AI/ML platform infrastructure.
Robots & Pencils is an applied AI engineering firm that designs and ships AI co-workers for enterprise operations. Founded in 2009, the company has delivery centers across Canada, the US, Eastern Europe, and Latin America, with teams averaging over 15 years of experience.
Lead engineering teams for Porting, Hosting, and A2P Messaging, owning delivery roadmap and operational excellence.
Drive adoption of AI coding assistants and measure impact on team velocity, fostering psychological safety.
Recruit, develop, and retain a world-class engineering team, promoting a culture of candid feedback and career growth.
Twilio is shaping the future of communications, delivering innovative solutions to hundreds of thousands of businesses and empowering millions of developers worldwide. They are a remote-first company with a strong culture of connection and global inclusion, employing a diverse team making a global impact daily.
Define the strategy, roadmap, and feature priorities for k0rdent AI Kubernetes services, empowering Neocloud operators to launch managed Kubernetes on their own GPU infrastructure.
Translate requirements from NeoClouds, GPU clouds, telcos, and enterprise platform teams into clear product direction and partner with engineering to ship secure, scalable cluster lifecycle capabilities.
Manage the Kubernetes backlog, define positioning and competitive differentiation, and create field-facing assets to support strategic accounts.
Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI. They have a world-class, distributed team committed to openness and technical excellence.
Design, build, and maintain automation and tooling to reduce operational toil.
Actively participate in the incident-management lifecycle, including detection, escalation, mitigation, and post-incident review.
Provide an SRE point of view on capacity planning, resilience testing, and modernization of legacy workloads.
Seismic is the Go-To-Market Performance company, helping organizations turn strategy into revenue through an AI-powered revenue execution platform. Trusted by 2,500 organizations and over 3.5 million users globally, Seismic is headquartered in San Diego with offices across North America, Europe, and Asia-Pacific, fostering an inclusive culture.
Contribute to infrastructure automation and operational resilience across hybrid cloud and data center operations.
Implement closed-loop auto-remediation systems and SRE tooling to reduce manual intervention and incident resolution time.
Develop and maintain SLO frameworks, alerting policies, and Infrastructure-as-Code pipelines for reproducible deployments.
ServiceNow is the AI control tower for business reinvention, helping 85% of the Fortune 500 work smarter, faster, and better. They foster an AI-native culture where technology and talent are unstoppable together.
Own and drive key infrastructure modernization initiatives toward container-orchestrated infrastructure.
Design and maintain infrastructure as code across multiple cloud providers.
Provide technical leadership and mentorship across the Systems Engineering team.
Intellum is the leader in corporate education technology, powering large learning programs for brands like Google, Meta, and Amazon. We are a remote-first company with a culture that values curiosity, creativity, perseverance, and kindness, and we invest in our people through personal development budgets and annual retreats.
Design, deploy, and sustain AWS-based platform services for mission-critical Department of War applications.
Administer Kubernetes clusters including lifecycle management, security, and observability.
Build CI/CD pipelines and infrastructure-as-code using Terraform, GitOps, and automation tools.
LMI is a digital solutions provider accelerating government impact with innovation and speed. Headquartered in Tysons, Virginia, the company serves defense, space, healthcare, and energy sectors with a focus on agility and collaboration.
Develop and evolve foundational software and services enabling product and development teams.\n- Architect, design, and implement Infrastructure as Code using Terraform.\n- Deploy, manage, and optimize Kubernetes clusters on GCP (GKE) and AWS (EKS).
DoiT is a global technology company that helps cloud-driven organizations leverage cloud for business growth and innovation through data, technology, and human expertise. They work with over 4,000 customers worldwide and foster a remote-first, entrepreneurial culture.
Design and manage high-availability platforms using Kubernetes, Terraform, and Ansible with native-AI capabilities.
Develop and operate the observability stack: Grafana, Mimir, Loki, Tempo, and Prometheus on Kubernetes via GitLab CI/CD.
Build automation scripts in Python, maintain GitOps pipelines, and mentor mid-level engineers.
Flexential builds and operates critical IT platforms including observability, DevOps, and ITSM technologies. The company fosters a collaborative engineering culture and values diversity.
Own the reliability, performance, and resilience of cloud environments (AWS, Kubernetes) and define SLOs across critical services.
Lead incident response, on-call rotation, and drive root cause analysis to ensure high production quality.
Build and maintain observability systems and automate operational toil using AI tools.
Garner partners with employers to redesign healthcare by using clinical metrics to identify top doctors and incentivize members to better care. The company has helped over 2.5 million people, saved $1B in costs, and doubled annually for five years, fostering a mission-driven, high-performance culture.
Lead the design, implementation, and ongoing improvement of reliable, scalable, and secure production platforms and services.
Work closely with cross-functional teams to build and maintain resilient infrastructure and deployment patterns.
Provide technical leadership and mentorship, promoting strong engineering standards and operational best practices.
Cision is a global leader in PR, marketing and social media management technology and intelligence, helping brands connect with customers and stakeholders. They have offices in 24 countries, a network of over 1.1 billion influencers, and a culture that champions diversity, equity, and inclusion.
Design and operate scalable AWS infrastructure with containerization and orchestration tools.
Implement monitoring, logging, and infrastructure as code using Terraform.
Improve CI/CD pipelines and troubleshoot production issues in complex SDLC environments.
Sureify builds systems that support millions of users. It is a high-growth, engineering-driven SaaS company with a remote-first culture across the Americas.
Lead the strategy and execution for Upbound's control plane experiences, including Crossplane upstream, enterprise distribution, and marketplace.
Manage multiple engineering teams and managers, scaling the organization to support product velocity and ecosystem expansion.
Drive ecosystem growth by partnering with cloud providers, ISVs, and platform partners to establish Upbound as the standard for programmable infrastructure.
Upbound is the creator and primary maintainer of Crossplane, the open-source project that extends Kubernetes into an AI-native cloud control plane. As a Series B startup backed by GV, Altimeter Capital, and Intel Capital, we serve Fortune 500 companies and platform engineers across 100+ countries, with over 100 million downloads.