Own the reliability, performance, and availability of all production databases across EHR, Data, and AI teams.
Manage database alerting, observability, and on-call rotation, and educate engineering teams on database hygiene.
Dive deep into query optimization, cost efficiency, and architecture to ensure clean platform scaling.
Prompt Health revolutionizes healthcare by providing automated B2B enterprise software for rehab therapy businesses. As a rapidly growing market leader, the company fosters a talented and collaborative culture focused on smart work and positive impact.
Ensure reliability, scalability, and operational excellence of analytics and data systems.
Provide technical leadership and direction to an offshore contract team.
Drive incident response, automation, and data governance initiatives.
Workiva provides an AI-powered platform that unifies finance, risk, and sustainability for complex organizations. It is a large enterprise with a collaborative and innovative culture centered on data integrity and trust.
Own the reliability, performance, and availability of production databases powering a healthcare technology platform.
Manage database observability, automation, incident response, and cloud cost optimization.
Educate engineering teams on query optimization and database best practices to prevent issues.
Jobgether is a job matching platform that uses AI to connect candidates with hiring companies. They focus on efficient, objective candidate evaluation and are part of a growing tech ecosystem.
Be on an on-call rotation responding to production incidents and support service engineers.
Run infrastructure with Ansible, Puppet, Terraform, and Kubernetes, making monitoring alert on symptoms.
Design and maintain core infrastructure scaling to hundreds of thousands of concurrent users.
Our client's Cloud Operations team is expanding its SRE function, keeping user-facing services and production systems running smoothly. The team specializes in systems like networking, Linux kernel, and distributed systems, blending pragmatic operations with software engineering.
Extend the self-service datastore platform with provisioning automation, guardrails, and paved paths for product engineering teams.
Ship observability, alerting, and backup/disaster recovery as built-in defaults for every datastore.
Convert recurring pull-in work into platform features or AI tooling that other teams can use directly.
Greenhouse provides a hiring software platform designed to make hiring work for everyone. They have an award-winning culture recognized by Fortune and Inc., and foster inclusivity, transparency, and accountability among their teams.
Lead and own core database systems to ensure high performance, availability, and security.
Work with engineering and data teams to shepherd schema changes from design to deployment.
Build and maintain automated test and deployment data pipelines with SRE and DevOps teams.
Branch is a fintech company that helps companies accelerate payments and provides working Americans with accessible financial services. They are a remote-first company with employees throughout the U.S., emphasizing transparency, accountability, and trust.
Own database reliability and drive high availability and disaster recovery strategies.
Architect for scale, including partitioning, sharding, and capacity forecasting.
Optimize performance through deep-dive query optimization and proactive bottleneck remediation.
Finom is a European tech startup developing an all-in-one financial B2B platform integrating banking, accounting, and invoicing. With over $346 million in total funding, the company nurtures an innovative and inspiring work environment where bold ideas thrive.
Design and maintain highly available, scalable systems to ensure exceptional customer experiences.
Drive automation and eliminate operational toil through self-service tooling and process improvements.
Lead incident response and mentor engineers to improve reliability practices.
Redzone provides a connected workforce solution for manufacturers to improve plant efficiency and worker productivity. The company is part of QAD Inc. and fosters a collaborative, customer-focused culture with a strong technology team.
Design and implement reliability strategies for distributed systems across AWS and GCP, defining SLIs and SLOs.
Build and enhance observability solutions using monitoring, logging, tracing, and alerting platforms.
Lead incident response, root cause analysis, and postmortem processes to improve system reliability.
We specialize in creating high-performing nearshore IT teams to help North American clients innovate faster and more efficiently. We are a people-first, purpose-driven company with a growing team, offering an inclusive culture and real growth opportunities.
Design and maintain infrastructure-as-code patterns using Terraform and Kubernetes for scalable deployments.
Build monitoring, logging, and alerting systems, lead incident response, and drive continuous reliability improvements.
Embed security into infrastructure and optimize performance, costs, and automation across the platform.
Remote enables global employment compliantly, allowing businesses to recruit, pay, and manage international teams. With a future-focused culture and fully remote team across six continents, it builds an innovative HR platform with automation and AI.
Administer and maintain the Snowflake platform including databases, roles, and performance optimization.
Implement security, automation, and governance standards across AWS cloud environments.
Collaborate with data engineering and security teams to deliver a scalable cloud data platform.
Empower is a financial services company focused on transforming financial lives. They offer a flexible work environment and career paths, with associates volunteering thousands of hours for community causes.
Design and implement scalable, reliable data store solutions for online and analytical workloads.
Manage and optimize SQL Server and MySQL databases, including schema design, query tuning, replication, backup/recovery, and performance monitoring.
Collaborate with data engineering and data science teams to deliver scalable, flexible solutions and contribute to data governance and security initiatives.
SurveyMonkey is the world’s most popular platform for surveys and forms, combining powerful capabilities with intuitive design. Trusted by millions from startups to Fortune 500 companies, it is a global company with over 25 years of history and a culture that champions inclusion and curiosity.
Build automation and Infrastructure as Code for scalable, self-service database provisioning and operations.
Ensure reliability and performance of relational databases like PostgreSQL and Aurora through deep expertise.
Collaborate with engineering teams to integrate data pipelines and reduce operational toil across diverse datastores.
Gemini is a global crypto and Web3 platform founded in 2014, offering crypto products in over 70 countries. As a publicly traded company, it focuses on bridging traditional finance with the cryptoeconomy and scaling engineering teams.
You will own and deliver quarterly goals for your team, leading engineers through ambiguity to solve open-ended problems.
You will proactively identify technical solutions and operational processes that strengthen incident readiness and response.
You will foster a culture of quality and ownership by setting or improving code review and design standards.
Affirm is reinventing credit to make it more honest and friendly, offering consumers the flexibility to buy now and pay later. The company has a strong engineering culture focused on reliability and ownership.
Lead technical and managerial direction for the SRE team, defining reliability, observability, and operational excellence strategy.
Coordinate critical incident responses and root cause analysis, collaborating with architecture, development, security, and product teams.
Drive automation, continuous improvement, and adoption of SRE, DevOps, and Platform Engineering best practices.
Experian is a global data and technology company that drives opportunities for people and businesses worldwide. With 25,200 employees in 32 countries, it has a people-centric, inclusive culture recognized by awards such as World's Best Workplaces™ 2025.
Lead and coordinate teams for systems incidents and N1, N2, N3 technical support, ensuring service continuity.
Oversee SRE, reliability, monitoring, and observability practices to improve system stability and performance.
Manage capacity, availability, and business continuity initiatives, anticipating risks and maintaining service levels.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It operates remotely and focuses on using technology to streamline recruitment processes.
Define and drive reliability of systems at the scale of millions of clients, strengthening SRE practices. - Develop observability platforms and serve as a strategic partner to product engineering teams. - Enhance proactive resilience through early-warning systems, AI/ML, and incident management.
XTB is a global FinTech company specializing in online trading of financial instruments. As the largest FinTech in Poland and a leader in Central and Eastern Europe, we operate across multiple continents and are a certified Great Place to Work, focusing on employee development and training.
Improve system availability, scalability, and resilience across Flowcode's platforms.
Manage and scale core AWS infrastructure through Infrastructure as Code (Terraform) and enhance disaster recovery.
Oversee monitoring, logging, and alerting infrastructure, and develop high-signal metrics and dashboards.
Flowcode is a technology company specializing in QR code and smart link solutions for offline-to-online engagement. The company is a growth-stage startup seeking high-performing individuals who thrive in a fast-paced, demanding environment.
Lead enterprise-wide reliability and infrastructure projects with high autonomy, architecting scalable solutions and driving SRE best practices.
Partner cross-functionally with Engineering, Product, and Customer Success to align reliability goals with business objectives and communicate complex concepts to diverse audiences.
Provide tier 2/3 technical support to enterprise customers, conduct technical onboarding, and act as a trusted advisor for platform architecture.
Veza is the pioneer in identity security, providing an Access Graph platform that maps identity ecosystems across users, groups, roles, policies, and resources. With over 30 billion access permissions under management and now part of ServiceNow, Veza combines enterprise scale with security innovation.