Source Job

US

  • Lead escalation triage and incident response for complex customer support cases, coordinating with engineering and product teams to drive resolution.
  • Build and maintain application observability using Datadog monitors, dashboards, and alerts to ensure platform health and proactive response.
  • Translate field experience into documentation, authoring troubleshooting guides and runbooks, and mentor Tier 1/2 support staff.

Linux Kubernetes Datadog Elasticsearch Cybersecurity

20 jobs similar to Senior Engineering Support Engineer

Jobs ranked by similarity.

Global Unlimited PTO

  • Reproduce, debug, and escalate technical issues in partnership with Engineering and Product.
  • Act as a technical expert and initial escalation point for our SDK and Cloud products.
  • Work directly with developers to provide guidance and resolve complex issues.

Ditto is a peer-to-peer sync engine that enables developers to build real-time applications that stay connected without internet. With over $145 million in funding and a globally distributed team, we are a fast-growing startup committed to diversity and inclusion.

$122,400–$177,500/yr
US Unlimited PTO 24w maternity 24w paternity

  • Serve as the primary technical escalation point for Customer Success, TAM, and Support teams, resolving complex security and technology issues.
  • Identify knowledge gaps and improve documentation, processes, and self-service resources to reduce recurring technical questions.
  • Mentor team members and collaborate across Security, Engineering, and Customer Success to scale solutions and knowledge sharing.

Global 7w PTO

  • Lead the design and operation of LivePerson's observability platforms across logs, metrics, traces, alerting, and synthetic monitoring.
  • Own large-scale observability pipelines using technologies like Elastic Cloud, Grafana, Prometheus, and Kafka.
  • Provide technical leadership and mentorship while driving best practices in DevOps, cloud engineering, and observability.

LivePerson is a leader in trusted enterprise conversational AI and digital transformation, powering nearly a billion conversational interactions every month. The company is recognized as the #1 Most Innovative AI Company by Fast Company and fosters a diverse, inclusive culture that empowers employees globally.

US

  • Engage directly with customers through support cases, remote sessions, and live conversations to understand technical challenges and provide effective solutions.
  • Troubleshoot complex Linux, cloud, networking, storage, virtualization, and container-related issues across customer environments.
  • Analyze logs, stack traces, system behavior, and application issues to identify potential operating system or application-level bugs and determine appropriate next steps.

The company supports customers operating critical workloads across Linux, cloud, and open source environments. It is a fully distributed organization with a global team and a culture of continuous learning and collaboration.

South Korea

  • Act as an escalation point for difficult customer interactions and lead the resolution of complex issues through Zendesk tickets, merge requests, email, and video calls.
  • Mentor and guide teammates on complex troubleshooting and customer communication through pairing sessions and cross-team collaboration.
  • Advocate for and drive improvements to support practices, product quality, and reusable documentation and support content that directly influence team objectives.

GitLab is the intelligent orchestration platform for DevSecOps, enabling organizations to increase developer productivity and accelerate digital transformation. With more than 50 million registered users and trusted by over 50% of the Fortune 100, GitLab fosters a high-performance culture driven by values and continuous knowledge exchange, where all team members are expected to embrace AI as a core productivity multiplier.

$126,290–$190,000/yr
United States 18w maternity 12w paternity

  • Empower engineers on other teams by maintaining monitoring tooling and collaborating on observability best practices.
  • Enhance reliability of Kubernetes applications through resource optimization, streamlined upgrades, and scalability.
  • Participate in on-call and incident response processes, occasionally diving into application code to debug production issues.

Webflow is the agentic web marketing platform for modern marketing teams, helping organizations build, manage, and optimize high-performing web experiences. It serves over 2 million users worldwide across 190 countries, with tens of thousands of projects launched each month, and fosters a culture of grit, speed, and craft.

$107,000–$128,000/yr
US 3w PTO

  • Provide senior-level technical engineering and support for enterprise clients on hybrid legacy and Kubernetes infrastructure.
  • Design and maintain CI/CD pipelines, manage containerized workloads on AWS EKS, and administer legacy server systems.
  • Collaborate with cross-functional teams and provide mentorship on automation and containerization practices.

Cotiviti is a healthcare analytics company specializing in payment accuracy and revenue integrity solutions for enterprise clients. The company fosters a collaborative culture and offers competitive benefits to support its team members.

US

  • Build and operate monitoring, tracing, alerting, and observability infrastructure for system reliability.
  • Drive platform security initiatives with preventative controls and resilient architecture.
  • Lead incident response and recovery, including root-cause analysis and preventative measures.

This role is with a partner company managing AI-powered products. They are a growing technology organization with a fully distributed US-based team and a collaborative culture focused on large-scale infrastructure and AI technology.

US Unlimited PTO 18w maternity 12w paternity

  • Lead technical implementations and drive adoption of Chainguard images for Public Sector customers.
  • Partner with account teams to understand customer vision and drive account level strategies.
  • Ensure high quality of Chainguard Images by acting as customer zero and collaborating with Product and Engineering teams.

Chainguard provides hardened, secure, and production-ready builds of open source software. They serve Fortune 500 enterprises and global industry leaders, and are venture-backed by top investors.

$241,000–$270,000/yr
US Unlimited PTO

  • Architect the end-to-end reliability, performance, and resilience of cloud environments, including the SLO framework for critical services.
  • Lead incident response, on-call rotation, root cause analysis, and build a culture of corrective actions.
  • Build observability platforms to detect issues proactively and mentor engineers on reliability standards.

Garner is on a mission to transform the U.S. healthcare system by partnering with employers to steer members to better-performing doctors, resulting in better care and lower costs. With 550+ proprietary clinical metrics, they have helped over 2.5 million people and saved $1B in healthcare costs, recently raising a Series E and doubling five years running.

Ireland

  • Investigate and resolve customer technical issues across cloud security posture management, vulnerability scanning, threat detection, and container/Kubernetes security.
  • Troubleshoot cloud connector and integration failures across AWS, Azure, GCP, OCI, and SaaS platforms.
  • Design and implement automation leveraging AI agents and tooling to improve triage accuracy and resolution efficiency.

Wiz is a cloud security platform that enables teams to secure cloud and AI applications by connecting code, cloud, and runtime into a single shared context. As one of the fastest-growing startups, powered by Google, the company is trusted by over 65% of the Fortune 100 and scans over 230 billion files daily.

$145,000–$177,000/yr
US

  • Build platform capabilities that enable engineering teams to deliver reliable software safely and efficiently.
  • Lead complex technical initiatives spanning cloud infrastructure, Kubernetes, observability, automation, and networking.
  • Design and implement solutions that improve availability, scalability, performance, and resilience of the platform.

Everbridge empowers enterprises and government organizations to anticipate, mitigate, respond to, and recover from critical events. The company focuses on building resilient systems and fostering a culture of ownership, continuous improvement, and operational excellence.

US

  • Design, build, and maintain highly reliable backend services and distributed systems for RapidFort's security platform.
  • Build systems for processing and analyzing large volumes of security, vulnerability, container, and runtime data.
  • Solve complex problems involving concurrency, performance, scalability, and distributed processing across Linux, Kubernetes, and cloud environments.

RapidFort is a cybersecurity company focused on securing and optimizing modern software supply chains and cloud-native environments. The company works with enterprise and U.S. public-sector customers in security-sensitive environments, fostering a culture of technical ownership and collaboration.

US Unlimited PTO

  • Design, develop, and maintain secure cloud environments for distributed LogRhythm deployments.
  • Act as a technical escalation point for SIEM engineers, solving complex security and infrastructure issues.
  • Build automation and operational tooling using PowerShell, SQL, Bash, or Python.

The company specializes in cybersecurity and managed SIEM services, focusing on protecting and scaling distributed LogRhythm environments. They offer a remote work environment with a collaborative culture and opportunities for professional growth.

APAC

  • Resolve technical cases for customers, troubleshooting unexpected behaviors and answering questions about the ServiceNow platform.
  • Use various diagnostic tools and technologies (web, chat, email, phone) to provide amazing customer support with empathy and excellent communication.
  • Manage challenging issues, coordinate with additional teams for complex cases, and provide input on process and product improvements.

Armis from ServiceNow protects the entire attack surface and manages an organization's cyber risk exposure in real time, combining with ServiceNow's security capabilities to build an autonomous defense platform. ServiceNow is the AI control tower for business reinvention, helping 85% of the Fortune 500 work smarter, faster, and better, with an AI-native culture.

United States

  • Lead a backend engineering team focused on building and scaling microservices, data pipelines, and APIs for a critical cybersecurity platform.
  • Mentor engineers and champion technical vision, ensuring reliability, scalability, and production-grade development.
  • Collaborate cross-functionally with product and engineering leadership to translate objectives into scalable solutions.

This position is listed on behalf of a partner company, which manages all applications and next steps. The environment is remote-first, mission-driven, and highly collaborative, with teams working across multiple regions.

Global 6w PTO

  • Own and improve production infrastructure reliability and stability.
  • Prepare, execute, and support deployments and infrastructure changes.
  • Build and maintain Infrastructure-as-Code solutions using Ansible and Terraform.

Social Discovery Group (SDG) is a group of social discovery companies that solve problems of loneliness, isolation, and disconnection by transforming virtual intimacy into the new normal. Our international team of digital nomads works remotely from all over the world and we are proud to be a two-time 'Great Place to Work' winner (USA & Japan, 2024–2025) and a Top-5 Company for Work-From-Anywhere Jobs (FlexJobs, 2025).

$107,100–$154,000/yr
US

  • Lead complex customer escalations from assessment through resolution, restoring trust with enterprise clients.
  • Coordinate cross-functional investigations across R&D, Engineering, Product, and Support teams.
  • Translate technical findings into clear communication for operational and executive stakeholders.

Our partner is a company in enterprise cybersecurity, focused on restoring customer trust through high-impact technical support. They are a fast-moving organization that values empathy and strong cross-functional collaboration, seeking a Critical Situation Manager to lead complex escalations.

$200,000–$240,000/yr
US Unlimited PTO 16w maternity 16w paternity

  • You define architecture and operational standards for managed services across AWS. - You serve as the final technical escalation point for complex customer situations. - You shape Honeycomb's open source strategy in the OpenTelemetry ecosystem and mentor engineers.

Honeycomb is a service for observability, defining developer tools for the near and present future. We are a fully distributed company of over 200 talented and inclusive bees, named to Forbes' America's Best Startups of 2022 and 2023.

US

  • Apply SRE principles to improve reliability, scalability, and performance of production systems.
  • Design and implement automation to reduce operational toil and improve engineering efficiency.
  • Lead incident response and develop sustainable solutions for complex production issues.

The hiring company is a technology organization focused on reliability and operational excellence. They offer a fully remote, collaborative environment with opportunities for technical leadership and career growth.