Design, build, and maintain scalable data pipelines and ETL/ELT processes using Databricks and Spark.
Architect and optimize data models and storage solutions for analytics and operational use.
Implement observability, alerting, and data quality monitoring for critical pipelines.
Sonatype provides end-to-end software supply chain security solutions, protecting against malicious open source and managing SBOMs. With over 2,000 organizations and 15 million developers using its platform, it focuses on innovation and security in software development.
Own the identity data model and coordinate cross-ecosystem mappings for billions of package observations.
Architect streaming and batch pipelines using Java, Spark, HBase, Databricks, and AWS primitives.
Lead design of data contracts for downstream teams, ensuring schema stability and versioning.
Sonatype is the software supply chain security company, providing end-to-end security solutions including proactive protection against malicious open source and enterprise-grade SBOM management. With over 2,000 organizations including 70% of the Fortune 100 and 15 million developers relying on its platform, Sonatype is a market leader in software supply chain security.
Design, develop, and maintain ETL/ELT data pipelines from multiple sources.
Integrate and optimize data storage and processing for high performance and scalability.
Collaborate with data scientists and analysts to ensure clean, structured datasets for analysis.
iCodde is a technology recruitment and staffing company. They focus on building data infrastructure and offer referral bonuses to leverage personal networks.
Design, build, and maintain business-critical data pipelines, including data ingestion and AWS Glue/Spark jobs feeding the warehouse and Tableau.
Build and operate core data infrastructure: the data warehouse, orchestration, and BI/reporting platform, deployed across multiple regions.
Drive work across producer and consumer teams, communicate clearly, and own outcomes end to end.
Horizon3 is a fast-growing cybersecurity company that provides autonomous pentesting and assessment operations via its NodeZero platform. They are a team of former special ops and engineers, committed to a culture of respect, collaboration, ownership, and results.
Design and maintain data pipelines on AWS using Spark, Python, and SQL.
Orchestrate workflows with Airflow, containerize with Docker and ECS.
Collaborate within agile teams to deliver high-quality data solutions.
Devoteam is a leading European consulting firm focused on digital strategy, technology platforms, cybersecurity, and business transformation through technology. With over 10,000 employees across 20 countries in Europe, the Middle East, and Africa, they foster a close-knit professional culture.
Own the Airflow codebase end-to-end, building reusable templates and enforcing standards.
Build and extend large-scale Spark pipelines on AWS Glue and support migration to Databricks.
Drive data model improvements around commercial pharma data with focus on structure and lineage.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, it surpassed $3B in revenue and values doing the right thing, customer success, employee success, and speed.
Support new and existing customers in their data engineering needs, guiding them to make optimal technical decisions.
Build and operationalize complex data solutions, including data governance, security, and quality.
Collaborate across teams to deliver data products and educate end users on analytic environments.
SunnyData is a high-growth consulting company specializing in data engineering and AI, dedicated to the Databricks platform. They foster a collaborative and innovative culture with a focus on customer impact and career growth.
Develop and maintain scalable, resilient data pipelines using PySpark and distributed processing frameworks.
Modernize legacy processes and design data solutions on AWS, building data products and complex transformations.
Support data quality and optimization initiatives, collaborating with stakeholders to translate requirements into technical solutions.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It focuses on efficiency and fairness in recruitment, promoting diversity and inclusion in the workplace.
Design, build, and maintain scalable data pipelines and platform capabilities to power analytics, AI/ML, and healthcare products.
Implement data quality, governance, and security controls, ensuring healthcare data is handled securely and in compliance.
Provide technical leadership, mentor engineers, and collaborate across teams to translate requirements into scalable data solutions.
Experity is a mission-driven team transforming on-demand healthcare across the U.S., empowering urgent care clinics with industry-leading software. They foster a culture of care, growth, and celebration with day-one benefits, career development, and a supportive team environment.
Build and maintain data pipelines using Databricks, dbt, Airbyte, and Airflow.
Contribute to data platform architecture and modeling, delivering trusted datasets.
3+ years experience in Data Engineering with strong SQL, Python, and cloud skills.
Ayming is an international consulting firm that supports businesses in digital and technological transformation. The company is expanding its technology, data, and transformation teams in Portugal and fosters a collaborative, innovative work environment.
Design and implement scalable data architectures to meet business needs.
Develop and optimize data pipelines for ingesting and processing large volumes of data.
Provide technical leadership and mentorship to junior engineers.
Oportun is a mission-driven financial services company that provides affordable credit and financial tools. Since inception, it has provided over $21.3 billion in credit and saved members $2.5 billion, and fosters a diverse, inclusive culture.
Design, build, and maintain scalable data pipelines using Python and AWS DynamoDB.
Develop and optimize DynamoDB data models and indexing strategies for high throughput.
Establish data validation, monitoring, and alerting mechanisms to ensure pipeline reliability.
Kunai builds full-stack technology solutions for banks, credit and payment networks, infrastructure providers, and their customers. With a team that thrives in a culture of collaboration and creativity, the company offers competitive compensation and professional development opportunities.
Guide strategic enterprise customers through cloud data engineering transformations, including performance testing and production-ready pipeline architecture.
Prove platform value by architecting solutions for big data, data warehousing, and lakehouse use cases.
Support technical sales through custom proofs of concept, workload sizing, and community workshops.
Databricks is the Data and AI company, providing a unified platform for data, analytics, and AI to over 20,000 organizations worldwide, including 70% of the Fortune 500. Headquartered in San Francisco with 30+ offices globally, it fosters a diverse and inclusive culture and offers comprehensive benefits.
Design and deliver complex data platform components or migration solutions on AWS using Databricks.
Troubleshoot performance, scalability, and reliability issues in cloud-native data environments.
Communicate technical topics clearly to stakeholders and produce high-quality documentation.
Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. As a fully remote global company with employees in Canada, the United States, and Latin America, they celebrate diverse cultures and foster a community of technological curiosity.
Design and build scalable data pipelines using Apache Spark and cloud technologies.
Develop and maintain data integrations and transformation processes from multiple sources.
Collaborate with technical teams to ensure data quality and optimize data architecture.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They focus on using technology to streamline the hiring process and offer a remote work environment.
Build and maintain ETL/ELT pipelines to onboard new data sources and ensure reliable data delivery to the platform.
Collaborate with cross-functional teams to support data models, infrastructure, and data governance standards.
Explore and pilot new tools to improve pipeline reliability and onboarding speed.
B Lab is the nonprofit behind Certified B Corporations, a community of businesses that meet verified social, environmental, and governance standards. They are a global nonprofit with a mission-driven team working towards economic systems change.
Become a trusted data and AI advisor, translating business questions into AI-ready data architectures.
Design and implement AI-optimized data platforms, including cloud data warehouses, lakehouses, and ETL/ELT pipelines.
Engineer modern ELT/ETL pipelines and data models using SQL, Python, and tools like Snowflake, Databricks, and dbt.
Aimpoint Digital is a fully remote data and analytics consultancy that partners with innovative software providers to solve complex business problems. The team is dynamic and collaborative, working independently on client engagements across industries.
Build and maintain data ingestion and transformation pipelines using Dagster, dbt, and Snowflake through PR-driven development.
Manage production monitoring, backfills, and Airbyte connections to ensure reliable daily data flows.
Collaborate with analytics and product teams on dimensional modeling and data quality, and participate in code review and runbook writing.
PadSplit is at the forefront of solving the affordable housing crisis through its tech-driven platform. They have a growing remote-first team and offer a competitive benefits package including unlimited PTO and paid parental leave.
Identify, design, and implement internal process improvements to automate manual processes and optimize data delivery.
Collaborate with data and analytics experts to enhance functionality in data systems.
Develop ETL jobs using MSSQL, Snowflake, Talend, and other tools to move data from OLTP databases to the Data Lake.
GreenSky is a financial technology company that provides innovative lending solutions. The company fosters a culture of collaboration and data-driven decision-making.