Own and evolve enterprise Linux platforms across on-premises and cloud environments, driving platform strategy, architecture, automation, security, and reliability.
Lead complex initiatives end to end, acting as a senior escalation point during incidents and driving root-cause analysis and permanent corrective actions.
Embed security-by-design principles, partner with security teams on hardening and compliance, and mentor engineers to advance platform engineering maturity.
Architect and automate business systems ecosystem integrating SaaS, cloud, and identity platforms.
Build APIs, integrations, and internal tools to improve global team workflows.
Implement security and compliance measures including identity lifecycle and Zero Trust.
Our partner is a technology organization focused on transforming internal IT operations through automation and platform engineering. They operate with a globally distributed, remote-first team that values autonomy, innovation, and continuous improvement.
Execute end-to-end server onboarding, including imaging and firmware updates, to support rapid global expansion.
Build and improve automation scripts and provisioning pipelines to handle high-volume server activation.
Collaborate with cross-functional teams to track onboarding metrics and drive process improvements.
Vultr provides high-performance cloud infrastructure solutions including Cloud Compute, GPU, Bare Metal, and Cloud Storage. With over 30 global data centers and hundreds of thousands of customers, it is the world's largest privately-held cloud infrastructure company, known for its self-funded growth and flexible culture.
Build reusable deployment automation using Ansible across Linux and Windows environments.
Develop scalable and portable solutions using virtualization and container technologies.
Support critical customer deployments and participate in a shared out-of-hours on-call rotation.
Techwan, part of the Everbridge Group, develops mission-critical command and control systems for emergency services and public safety agencies. The company offers a flexible, remote-first working environment with significant autonomy and ownership on international projects.
Operate the Monad node fleet, including health, sync, upgrades, and incident response for validators, full nodes, and archive nodes.
Own infrastructure-as-code with Ansible, Terraform, and Kubernetes, and build observability with Prometheus, Grafana, and Loki.
Design and build AI agent tooling for automated operations, including runbooks-as-code and deterministic guardrails.
Category Labs designs and builds decentralized technology, including the Monad blockchain, a high-performance EVM-compatible Layer 1. The team raised $225M in series A funding and is a lean, collaborative group of engineers and researchers with a culture of low ego and high-quality output.
Lead and execute vulnerability remediation initiatives across Windows and Linux servers.
Plan, coordinate, and support automated OS patching with minimal disruption.
Collaborate with application teams and IT stakeholders to ensure compliance.
A partner company is hiring an Infrastructure Engineer to strengthen security, reliability, and automation of large-scale enterprise infrastructure. The company fosters a collaborative culture focused on DevOps and automation.
Provide advanced technical support with a focus on root-cause analysis, security, and employee experience.
The company is a global, technology-driven organization that redefines internal IT as an engineering discipline. It fosters a culture of innovation, automation, and continuous improvement with a distributed workforce.
Design, build, and optimize internal platforms, automation workflows, and business system integrations.
Develop automation scripts, internal tools, and API-driven integrations using modern languages.
Implement security controls, identity management, and compliance processes across business systems.
This company is a globally distributed organization that builds and optimizes internal technology ecosystems. They have a collaborative international team and an engineering-driven, remote-first culture.
Own and improve production infrastructure reliability and stability.
Prepare, execute, and support deployments and infrastructure changes.
Build and maintain Infrastructure-as-Code solutions using Ansible and Terraform.
Social Discovery Group (SDG) is a group of social discovery companies that solve problems of loneliness, isolation, and disconnection by transforming virtual intimacy into the new normal. Our international team of digital nomads works remotely from all over the world and we are proud to be a two-time 'Great Place to Work' winner (USA & Japan, 2024–2025) and a Top-5 Company for Work-From-Anywhere Jobs (FlexJobs, 2025).
Design, develop, and maintain automated diagnostic, validation, and remediation frameworks for production GPU hardware using Python and infrastructure automation tools.
Engineer and support Python-based agents, APIs, and Ansible automation for hardware provisioning, telemetry, and health monitoring.
Analyze workload performance, thermals, and diagnostic output to identify hardware issues and improve validation methodologies.
Vultr is on a mission to make high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators worldwide. We are the world's largest privately-held cloud infrastructure company, self-funded for over a decade, with 33 global data centers and hundreds of thousands of active customers.
Configure and manage Linux security policies, including access controls, firewalls, and privileged-user management.
Monitor SIEM alerts and logs to identify threats, apply patches, and conduct vulnerability assessments.
Develop security automation scripts and incident response strategies to strengthen infrastructure.
Jobgether is an AI-powered recruitment platform that matches candidates with hiring companies. It operates with a remote-first culture and emphasizes fast, objective, and fair candidate evaluation through AI matching.
Design and implement production-quality infrastructure and internal platforms with end-to-end ownership.
Build and expand automated testing infrastructure for software deployed across diverse edge hardware.
Improve CI/CD pipelines, build and release systems, and overall platform reliability.
The company is a technology organization that builds infrastructure for reliable software delivery across complex edge environments. It is a small, fully remote engineering team with a culture of ownership, technical judgment, and fast execution.
Automate standard operating procedures and optimize daily efficiency of systems and cloud management.
Work with a portfolio of products across different technologies, data sets, and problems.
Provide automation runbooks to the cloud engineering team for core systems management and SaaS application platforms.
The client's Cloud Operations team is expanding and focuses on building automation tooling for cloud management. The team size and culture are not specified in the posting.
Be on an on-call rotation responding to production incidents and support service engineers.
Run infrastructure with Ansible, Puppet, Terraform, and Kubernetes, making monitoring alert on symptoms.
Design and maintain core infrastructure scaling to hundreds of thousands of concurrent users.
Our client's Cloud Operations team is expanding its SRE function, keeping user-facing services and production systems running smoothly. The team specializes in systems like networking, Linux kernel, and distributed systems, blending pragmatic operations with software engineering.
Creating a golden ISO for imaging servers prior to shipping.
Performance testing Apache Traffic Server.
Building automation for server provisioning, patching, and monitoring/alerting.
ELEVI provides services to Federal and Commercial clients. They foster a culture of trust, empowerment, and diversity, offering flexible work arrangements and competitive benefits.
Lead the expansion of our Terraform module library and refine existing Infrastructure-as-Code frameworks for Azure.
Build an AI-enabled self-service layer that generates compliant infrastructure automatically from developer requests.
Automate governance and developer workflows using ServiceNow, Port.io, and CI/CD pipelines.
Valce Talent Solutions enhances talent attraction capacities for technology companies, specializing in IT, software development, cybersecurity, and project management. Founded in 2016, we have 24 employees and foster a culture of innovation and employee satisfaction.
Design, manage, and optimize AWS infrastructure across production and non-production environments to ensure scalability, reliability, and performance.
Architect highly available cloud solutions that support API-driven and data-intensive applications while implementing infrastructure-as-code using Terraform or CloudFormation.
Establish and maintain operational best practices, including monitoring, alerting, incident response, root cause analysis, and disaster recovery planning.
The company is a technology-driven organization focused on building scalable cloud infrastructure. It values innovation, collaboration, and operational excellence, with a remote-first culture.
Implement assigned security controls across Kubernetes, Linux, database, and model-serving environments.
Build monitoring, vulnerability scanning, and automation to keep security controls current.
Contribute to AI tooling, documentation, and evidence collection for federal reviews.
Axle is a bioscience and IT company offering advancements in translational research, biomedical informatics, and data science applications to research organizations. Their team includes biomedical science, software engineering, and program management experts who support top research centers, including NIH.
Develop and maintain infrastructure automation solutions using Ansible.
Design, implement, and enhance CI/CD pipelines and operational tooling.
Troubleshoot complex Linux-based infrastructure and distributed systems issues to maintain high availability.
itD is a consulting and software development company that blends diversity, innovation, and integrity with real business results. They are a woman- and minority-led firm offering a dynamic culture of respect, empowerment, and recognition.
Design, deploy, and maintain resilient infrastructure for Oracle Utilities environments.
Perform Linux systems administration, networking, scripting, and SQL optimization.
Troubleshoot incidents and provide technical support to clients across North America.
The company specializes in supporting business-critical Oracle Utilities environments across North America. It operates as a remote-first, collaborative team with a focus on reliability and client success.
Drive the performance, stability, security, and reliability of production environments with a focus on automation and proactive improvements.
Design and maintain infrastructure using Infrastructure as Code tools like Terraform, and manage Kubernetes and cloud environments.
Lead vulnerability management, incident response, and secure CI/CD practices to ensure resilience and operational excellence.
Jobgether is a platform that uses AI-powered matching to connect candidates with hiring companies. It processes applications and shares shortlists with employers, offering a remote-first and inclusive work environment.