You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Design, develop, and maintain scalable, production-ready data pipelines and data products using Spark (Python/SQL) in a Databricks environment.
Lead the integration and transformation of complex data from diverse DoD and federal health systems into reliable, reusable data products.
Provide technical guidance and mentorship to other engineers, helping teams navigate complex technical challenges.
540 is a forward-thinking company that delivers innovative technology solutions for government missions. The team has a culture of breaking down barriers and solving mission-critical problems.
Design, build, and maintain scalable data pipelines and ETL/ELT processes using Databricks and Spark.
Architect and optimize data models and storage solutions for analytics and operational use.
Implement observability, alerting, and data quality monitoring for critical pipelines.
Sonatype provides end-to-end software supply chain security solutions, protecting against malicious open source and managing SBOMs. With over 2,000 organizations and 15 million developers using its platform, it focuses on innovation and security in software development.
Guide strategic enterprise customers through cloud data engineering transformations, including performance testing and production-ready pipeline architecture.
Prove platform value by architecting solutions for big data, data warehousing, and lakehouse use cases.
Support technical sales through custom proofs of concept, workload sizing, and community workshops.
Databricks is the Data and AI company, providing a unified platform for data, analytics, and AI to over 20,000 organizations worldwide, including 70% of the Fortune 500. Headquartered in San Francisco with 30+ offices globally, it fosters a diverse and inclusive culture and offers comprehensive benefits.
Build, maintain, and optimize scalable data infrastructure for AI/ML features.
Partner with ML Engineers, Data Scientists, and Software Engineers to transform healthcare data into high-quality datasets.
Contribute to data pipelines, improve data quality, and ensure reliable, performant, and well-governed data for AI systems.
Tebra is the only all-in-one EHR+ platform built exclusively for independent healthcare practices. More than 42,000 private practices trust Tebra to streamline operations, increase revenue, and reduce burnout, with a culture that values starting with the customer, keeping it simple, staying entrepreneurial, being better together, and celebrating success.
Lead technical discovery and design solutions for customer workloads in data engineering, analytics, and ML.
Build and deliver compelling proofs-of-concept and live demos on the Databricks Platform.
Own frontline technical relationships with customer engineers and data teams to drive platform adoption.
Databricks is the Data and AI company. Over 20,000 organizations worldwide, including 70% of the Fortune 500, rely on its unified platform for data, analytics, and AI, with a culture of proactiveness and customer-centricity.
Become a trusted data and AI advisor, translating business questions into AI-ready data architectures.
Design and implement AI-optimized data platforms, including cloud data warehouses, lakehouses, and ETL/ELT pipelines.
Engineer modern ELT/ETL pipelines and data models using SQL, Python, and tools like Snowflake, Databricks, and dbt.
Aimpoint Digital is a fully remote data and analytics consultancy that partners with innovative software providers to solve complex business problems. The team is dynamic and collaborative, working independently on client engagements across industries.
Build and maintain data pipelines using Databricks, dbt, Airbyte, and Airflow.
Contribute to data platform architecture and modeling, delivering trusted datasets.
3+ years experience in Data Engineering with strong SQL, Python, and cloud skills.
Ayming is an international consulting firm that supports businesses in digital and technological transformation. The company is expanding its technology, data, and transformation teams in Portugal and fosters a collaborative, innovative work environment.
Design and deploy highly performant end-to-end data architectures using Databricks and Apache Spark.
Develop and manage scalable batch and streaming solutions in cloud ecosystems like AWS, Azure, or GCP.
Lead large-scale data engineering projects, ensuring technical delivery and client management.
Derex Technologies Inc provides IT consulting, staffing solutions, and software services to global clients. With over two decades of experience, the company delivers high-quality technology professionals across various industries.
Design, build, and maintain scalable data pipelines and platform capabilities to power analytics, AI/ML, and healthcare products.
Implement data quality, governance, and security controls, ensuring healthcare data is handled securely and in compliance.
Provide technical leadership, mentor engineers, and collaborate across teams to translate requirements into scalable data solutions.
Experity is a mission-driven team transforming on-demand healthcare across the U.S., empowering urgent care clinics with industry-leading software. They foster a culture of care, growth, and celebration with day-one benefits, career development, and a supportive team environment.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Design, build, and maintain scalable data pipelines and systems using Python, PySpark, and AWS-native services.
Develop and optimize data platforms on AWS, leveraging technologies like Redshift, Glue, and Athena.
Mentor junior engineers and collaborate with cross-functional teams in an Agile environment.
Capital Technology Group provides expert consulting services in software development, digital transformation, and data analytics. Recognized as a Top Workplace by The Washington Post, they foster a culture where employees feel trusted and empowered.
Design and deliver complex data platform components or migration solutions on AWS using Databricks.
Troubleshoot performance, scalability, and reliability issues in cloud-native data environments.
Communicate technical topics clearly to stakeholders and produce high-quality documentation.
Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. As a fully remote global company with employees in Canada, the United States, and Latin America, they celebrate diverse cultures and foster a community of technological curiosity.
Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.
CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.
Own architecture and delivery for modern data platforms end-to-end.
Shape client data strategy and design scalable solutions.
Mentor internal team members and ensure high-quality delivery.
8th Light is a technology solutions consultancy that partners with organizations from early-stage startups to the world's largest companies. Founded in 2006, we foster an open, collaborative culture grounded in honesty and continuous education.
Maintain and monitor Kafka, Hadoop, Presto, and RDBMS systems.
Ingest, validate, and process internal and third-party data flows.
Build Kafka consumers using Spark Streaming for near-real-time aggregation.
PulsePoint sits at the intersection of healthcare and adtech, helping brands interpret health journey signals. We are 300+ employees, growing, and a leading player in the US healthcare ad market.
Design, architect, and develop robust GCP data pipelines for healthcare data solutions.
Maintain a holistic view of information assets through documentation and create technical design specs.
Collaborate with cross-functional teams to drive ML and data engineering initiatives that improve patient outcomes.
Egen is a data-first consulting company leveraging Google Cloud and Salesforce to help clients drive impact through data. It is a fast-growing, entrepreneurial firm with a culture focused on learning, innovation, and solving tough problems.
Design and build scalable data pipelines and schemas for client engagements.
Lead data architecture and modeling discussions, weighing tradeoffs.
Mentor teammates and contribute to 8th Light's culture and values.
8th Light is a technology solutions consultancy that partners with organizations to build software, platforms, and agentic products. Founded in 2006, the company fosters an open, collaborative culture grounded in honesty and continuous learning.
Design and build ingestion pipelines from enterprise source systems into the Databricks lakehouse using Delta Lake, owning Bronze-layer ingestion and building Silver-layer pipelines for cleansing and standardization.
Implement and maintain Databricks platform constructs for secure delivery, including catalogs, schemas, service principals, and job orchestration, and build CI/CD pipelines for data platform assets.
Apply data classification and segregation requirements within pipeline design, build data quality controls reflecting business meaning, and partner with teams to ensure reliable, governed data for downstream use.
Shield AI is a venture-backed defense-tech company founded in 2015 with the mission of protecting service members and civilians with intelligent systems, developing products like Hivemind autonomy software and V-BAT and X-BAT aircraft. The company has offices and facilities across the U.S., Europe, the Middle East, and Asia-Pacific, and its technology actively supports operations worldwide.
Own the Airflow codebase end-to-end, building reusable templates and enforcing standards.
Build and extend large-scale Spark pipelines on AWS Glue and support migration to Databricks.
Drive data model improvements around commercial pharma data with focus on structure and lineage.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, it surpassed $3B in revenue and values doing the right thing, customer success, employee success, and speed.