Manage team performance, career development, and project prioritization while driving a culture of automation.
Drive initiatives with partner teams to improve infrastructure reliability and act as crisis management.
Analyze existing processes to drive continuous improvement and efficiencies.
ServiceNow is the AI control tower for business reinvention, helping 85% of the Fortune 500 work smarter with an intelligent cloud platform. We are building an AI-native culture where technology and talent are unstoppable together, serving over 8,100 customers.
Lead cross-functional development of large-scale infrastructure programs from concept to delivery.
Align stakeholders across engineering, product, and finance to manage risks and drive program success.
Optimize engineering processes and drive cost efficiency through automated monitoring and reduction initiatives.
Lime is a global shared micromobility company on a mission to make transportation shared, affordable, and carbon-free. It has powered over one billion rides in nearly 30 countries and is a Time Magazine 100 Most Influential Company, fostering a culture of strategic thinking and operational excellence.
Extend the self-service datastore platform with provisioning automation, guardrails, and paved paths for product engineering teams.
Ship observability, alerting, and backup/disaster recovery as built-in defaults for every datastore.
Convert recurring pull-in work into platform features or AI tooling that other teams can use directly.
Greenhouse provides a hiring software platform designed to make hiring work for everyone. They have an award-winning culture recognized by Fortune and Inc., and foster inclusivity, transparency, and accountability among their teams.
Support cloud-hosted application testing, implementation, maintenance, and validation.
Monitor and troubleshoot Windows, Linux, server, network, and application environments.
Assist with incident response, documentation, and operational process improvements.
Applied Systems builds cloud software and AI-powered solutions that reinvent insurance technology for agencies and brokers worldwide. With over 40 years of experience, the company fosters a people-first culture built on trust, inclusion, and growth, supporting employees to deliver their best work.
Lead the Network Operations team through infrastructure modernization and cloud-native transition.
Drive automation using Terraform and infrastructure-as-code to reduce manual operations.
Maintain 99.999% uptime and manage corporate network infrastructure including Cisco Meraki.
Marqeta is a card issuing platform that enables companies to issue cards and manage payment operations in real time. They are a publicly-traded company powering well-known brands in the new economy, with a culture focused on customer success, innovation, and teamwork.
Define and drive the vision and multi-quarter roadmap for the infrastructure foundations team, tying investments to business outcomes.
Lead and mentor a team of infrastructure engineers, fostering ownership, collaboration, and technical excellence.
Own the reliability and safety of the foundational AWS layer, including account provisioning, core networking, and IAM access.
Affirm is reinventing credit to make it more honest and friendly, offering consumers the ability to buy now and pay later without hidden fees or compounding interest. As a publicly traded company, Affirm fosters a culture of thorough technical design review, operational excellence, and capable incident response.
Lead and develop a global team of SRE leaders, managers, and engineers, driving reliability strategy and operating model.
Own and evolve observability capabilities across metrics, logs, traces, alerting, SLI/SLOs, and service health.
Drive cloud modernization initiatives, advancing containerization and Kubernetes-based operating models.
ServiceNow is the AI control tower for business reinvention, bringing together AI, data, and workflows to help 85% of the Fortune 500 work smarter. The company fosters an AI-native culture where technology and talent are unstoppable together.
Lead technical service delivery and act as senior escalation point for Tier 1 and Tier 2 analysts, resolving complex incidents and maintaining SLA performance.
Coach, mentor, and provide real-time guidance to Service Desk Analysts, supporting team development and quality assurance.
Collaborate with Tier 3 engineers and specialist teams to drive timely resolution of incidents and service requests using Microsoft, infrastructure, and ITSM expertise.
Jobgether uses an AI-powered matching process to review applications and share top candidates with hiring companies. The partner company is a large-scale enterprise IT service delivery organization with a collaborative, technical environment.
Take an active role as co-owner of production services to ensure they are built, maintained, and operated in a reliable and scalable way.
Collaborate with Software Engineering to drive operational improvements through metric-driven analysis and help scale AWS and Kubernetes infrastructure.
Participate in a weekly on-call rotation to investigate and resolve potential system issues, and automate routine tasks in at least two programming languages.
Zerohash is the leading crypto and stablecoin infrastructure platform, powering the next generation of financial services for banks, brokerages, fintechs, and payment companies. Founded in 2017, the company has raised over $280 million from top venture firms and strategic investors, and is trusted by global brands like Morgan Stanley and Stripe, operating with a compliance-first approach.
Deploy, upgrade, and support applications, services, and operating systems.
Troubleshoot system performance and application health issues.
Kinaxis is a global leader in modern supply chain orchestration, powering complex global supply chains. With over 2000 employees and multiple Top Employer awards, they foster a culture of innovation and collaboration.
Lead enterprise-wide reliability and infrastructure projects with high autonomy, architecting scalable solutions and driving SRE best practices.
Partner cross-functionally with Engineering, Product, and Customer Success to align reliability goals with business objectives and communicate complex concepts to diverse audiences.
Provide tier 2/3 technical support to enterprise customers, conduct technical onboarding, and act as a trusted advisor for platform architecture.
Veza is the pioneer in identity security, providing an Access Graph platform that maps identity ecosystems across users, groups, roles, policies, and resources. With over 30 billion access permissions under management and now part of ServiceNow, Veza combines enterprise scale with security innovation.
Enable effective execution with Quality and Speed, in partnership with the team's Product Manager.
Ensure 3+ 9s availability of Dedicated infrastructure, ensuring security and automating for maximum scalability.
Provide clear direction, meaningful feedback and foster an environment where meaningless toil gets ruthlessly automated.
GitLab is the intelligent orchestration platform for DevSecOps, enabling organizations to increase developer productivity and improve operational efficiency. With over 50 million users and 50% of Fortune 100, GitLab fosters a high-performance culture driven by values.
Operate as the NOC’s first point of escalation for any issues raised within or out of shift that need further assistance or feedback.
Ensure NOC shift operations are executed properly and according to pre-defined SLAs.
Manage NOC’s shift schedule including PTO requests and act as focal point for administrative issues.
LivePerson is a global leader in enterprise conversations, providing a Conversational Cloud platform for brands like HSBC and Chipotle. The company powers nearly a billion interactions monthly and fosters an inclusive workplace culture that encourages collaboration and innovation.
Design and improve monitoring, logging, distributed tracing, dashboards, alerting, SLIs, and SLOs for production health.
Build and maintain automation, internal tools, and CI/CD systems to increase engineering efficiency and support reliable deployments.
Own complex production incidents from detection to resolution, turning learning into durable improvements and reducing recurring incidents.
Filevine is a Legal AI company delivering Legal Operating Intelligence for the future of legal work. It has earned recognition from Deloitte and Inc. as one of the most innovative and fastest-growing technology companies in the country.
Be the primary team to build and manage our Canadian based sovereign cloud on Google Cloud and contribute to global cloud infrastructure.
Manage enterprise infrastructure at scale with code using Terraform & Crossplane on public cloud, and develop services using Go, Python, and automation tools.
Research new technologies, collaborate with peer teams to ensure stability and security.
ServiceNow is the AI control tower for business reinvention, bringing together any AI, data, and workflow to help 85% of the Fortune 500 work smarter. They foster an AI-native culture where technology and talent are unstoppable together, offering great perks and personal growth opportunities.
Lead and coordinate teams for systems incidents and N1, N2, N3 technical support, ensuring service continuity.
Oversee SRE, reliability, monitoring, and observability practices to improve system stability and performance.
Manage capacity, availability, and business continuity initiatives, anticipating risks and maintaining service levels.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It operates remotely and focuses on using technology to streamline recruitment processes.
Lead a distributed infrastructure operations team across US and APAC time zones, overseeing bare-metal environments and network engineering.
Own incident response, capacity planning, hardware lifecycle, and operational reliability across four global data centers.
Manage vendor relationships, budgets, and infrastructure strategy while collaborating with security and engineering leadership.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It uses technology to ensure applications are reviewed quickly and objectively, and the platform is designed to streamline the hiring process.
Monitor and troubleshoot systems, networks, and applications to ensure optimal performance and availability.
Perform system administration tasks including server configuration, maintenance, and upgrades.
Collaborate with vendors and recommend new technologies to improve system performance and scalability.
Ocrolus builds an AI workflow and analytics platform for lenders, processing nearly one million credit applications monthly with over 99% accuracy. Trusted by over 400 customers, the company is a fast-growing, remote-first team valuing empathy, curiosity, humility, and ownership.
Own end-to-end delivery of AWS cloud infrastructure programs, managing scope, schedules, and dependencies across engineering teams.
Coordinate cloud cost-optimization initiatives, tracking progress and risks to meet savings targets.
Drive security and governance initiatives, ensuring alignment and on-track delivery.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It operates as a technology-driven organization with a global, collaborative team focused on efficient recruitment.