Take an active role as co-owner of production services to ensure they are built, maintained, and operated in a reliable and scalable way.
Collaborate with Software Engineering to drive operational improvements through metric-driven analysis and help scale AWS and Kubernetes infrastructure.
Participate in a weekly on-call rotation to investigate and resolve potential system issues, and automate routine tasks in at least two programming languages.
Zerohash is the leading crypto and stablecoin infrastructure platform, powering the next generation of financial services for banks, brokerages, fintechs, and payment companies. Founded in 2017, the company has raised over $280 million from top venture firms and strategic investors, and is trusted by global brands like Morgan Stanley and Stripe, operating with a compliance-first approach.
Be on an on-call rotation responding to production incidents and support service engineers.
Run infrastructure with Ansible, Puppet, Terraform, and Kubernetes, making monitoring alert on symptoms.
Design and maintain core infrastructure scaling to hundreds of thousands of concurrent users.
Our client's Cloud Operations team is expanding its SRE function, keeping user-facing services and production systems running smoothly. The team specializes in systems like networking, Linux kernel, and distributed systems, blending pragmatic operations with software engineering.
Define and implement reliability strategy including SLOs, SLIs, error budgets, and incident practices.
Manage cloud infrastructure on AWS using Infrastructure as Code and ensure Kubernetes scalability.
Lead incident response and establish chaos engineering practices to strengthen platform resilience.
This partner company builds a globally scaled, AI-native platform with a focus on reliability and event-driven systems. They offer a collaborative international culture with significant technical ownership and continuous improvement.
Design and evolve scalable cloud infrastructure on Google Cloud Platform, focusing on reliability and automation.
Strengthen observability platform with metrics, logging, and tracing to improve incident response and reduce recovery time.
Champion reliability practices like SLOs, error budgets, and DORA metrics to drive operational excellence.
They operate at the intersection of geospatial intelligence and environmental technology. They are a growing organization with a collaborative, high-impact engineering culture.
Owns coordination and execution of minor releases, patch deployments, and production emergency deployments.
Ensures release artifacts, versioning, and documentation are complete and traceable.
Manages deployment timing and sequencing with development, testing, and implementation teams.
VXcis Development organization coordinates software releases and deployments to minimize production risk. They are a growing team seeking a dedicated specialist to own release readiness and execution.
Design and scale reliable AWS infrastructure across multiple accounts, ensuring security and compliance.
Collaborate with software engineers to advise on infrastructure best practices and optimize CI/CD pipelines.
Automate critical infrastructure updates and enhance observability with monitoring tools like Prometheus.
Constructor is an AI-first ecommerce search and discovery platform that helps shoppers find the right products and enables leading global e-commerce brands to drive revenue and conversion gains. They are a fully remote, diverse team offering unlimited vacation, training budgets, and the chance to work with smart colleagues.
Lead end-to-end deployment of North in private cloud and on-premises environments, including planning, configuration, testing, and rollout.
Partner with enterprise IT teams to assess infrastructure, security requirements, and data management practices.
Design and implement deployment strategies tailored to client needs, ensuring compliance with data privacy and security standards.
Cohere is a security-first enterprise AI company that builds cutting-edge foundation models and end-to-end products. It is a global team of researchers, engineers, and designers passionate about their craft.
Be the primary engineer responsible for infrastructure as code, CI/CD, mobile release automation, data pipelines, and monitoring.
Help set the direction for AI use across engineering and build shared tooling, agent workflows, and automation.
Scale yourself across engineering through documentation, demos, and tooling other engineers can use on their own.
The Mind Company develops award-winning, science-backed mental fitness apps like Elevate, Balance, and Spark to help millions improve cognitive and wellness skills. They are a fully remote team of passionate learners driven by a mission to bring mental fitness to every mind.
Serve as Scrum Master for Agile software development teams, facilitating ceremonies and removing impediments.
Partner with cross-functional teams to analyze requirements and coordinate delivery across multiple teams.
Foster a collaborative Agile culture, mentor junior team members, and promote continuous improvement.
Capital Technology Group provides expert consulting services in software development, digital transformation, human-centered design, data analytics, and cybersecurity. We have been trusted by federal and commercial clients for over a decade, and are recognized as a 2025 Top Workplace by The Washington Post, fostering a culture rooted in our core values.
Develops GitLab CI/CD Pipelines and automates configurations within Kubernetes.
Designs and maintains modular, reusable IaC configurations using Terraform across cloud platforms like AWS.
Provides automated and manual validations of Information Assurance Controls in accordance with DoD Guidelines.
Entarian is a premier provider of mission-critical engineering and technology solutions, delivering secure digital solutions for defense and civilian missions. Founded in 1993, Entarian is a diversified engineering and federal technology leader with a culture of driving national resilience and operational effectiveness.