Evaluate AWS Serverless and IaC tasks for technical accuracy and reliability.
Provide clear, actionable feedback on deployment pipeline errors and architecture issues.
Ensure tasks are realistic, reproducible, and supported by robust tests.
We source experienced technical specialists to audit tasks used to train AI systems. Our project ensures that AI training workflows are technically rigorous and accurate, with a focus on high-quality deliverables.
Assess technical accuracy and reproducibility of AWS Trainium/NKI tasks.
Provide actionable feedback on kernel execution bugs and logic errors.
Evaluate hardware acceleration inefficiencies and compilation issues.
We source experienced technical specialists to audit AI training tasks and evaluation workflows. The project is remote and freelance, with a focus on technical accuracy and efficiency.
Assess technical accuracy, realism, and reproducibility of AI-assisted developer workflow tasks.
Provide actionable feedback on IDE integration faults, AI-assisted coding inefficiencies, and logic errors.
Apply deep knowledge of modern developer tooling, including AI coding assistants and telemetry/trace analysis.
We are sourcing experienced technical specialists to audit tasks used in training and evaluating AI systems. We focus on ensuring technical rigor and accuracy in developer workflow tasks, operating as a remote freelance project.
Design and build complex, hands-on lab scenarios that showcase Sysdig across containers, Kubernetes, Linux and public cloud.
Own the automation and reliability of the training platform, including provisioning and scripting.
Design and deliver engaging technical training, workshops and demos for diverse audiences.
Sysdig stops attacks in real-time by detecting changes in cloud security risk with runtime insights and open source Falco. It is a well-funded startup with a large enterprise customer base, recognized as a best place to work.
Assess software engineering tasks for technical accuracy, realism, and reproducibility.
Provide actionable feedback on codebase integration issues and logic errors.
Ensure AI training workflows are rigorous and practically applicable.
Project World Wide sources experienced technical specialists for AI training task auditing. This freelance contract opportunity focuses on ensuring technical rigor and accuracy in AI workflows.
Audit multiple-choice and coding questions for technical accuracy in a FastAPI microservices exam.
Verify code prompts, evaluate grading rubrics, and add edge test cases.
Submit a corrected JSON file with your expert modifications.
Terac builds the world's largest pool of vetted human experts for AI. Researchers and AI labs use Terac to recruit, screen, and pay study participants across many industries and languages.
Own the infrastructure layer for AI workloads including inference serving, Kubernetes, and agent-sandboxing platforms.
Manage the serving tier for open-weight models, Kubernetes operators, and stateful data planes.
Oversee the sandbox runtime, control-plane services, and observability tooling.
AZX accelerates positive impact in critical industries through AI transformation, specializing in physics-informed ML and enterprise AI solutions for climate and sustainability. Founded in 2024, the company is a profitable public benefit corporation with a growing team working with category leaders in real estate, energy, logistics, and utilities.
Build and deploy production code to support customer AI inference workloads on Tenstorrent's hardware and software stack.
Debug and optimize across the full inference stack, from serving layer to kernel dispatch, and translate customer issues into actionable requirements.
Operate Kubernetes and observability tools to manage multi-node AI clusters and ensure reliability.
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. Their diverse team of technologists has developed a high-performance RISC-V CPU from scratch, and they value collaboration, curiosity, and a commitment to solving hard problems.
Design, build, and operate Kubernetes infrastructure for AI workloads using Terraform and GitOps.
Define SLOs, run incident response, and create runbooks for reliable AI platform operations.
Drive AI-specific observability, FinOps, and security practices across the platform.
We are an AI-native consulting partner working with clients like PayPal, adidas, and NatWest to build digital products and services. Our team of over 600 has scaled quickly, earning Great Place to Work-Certified status multiple years in a row.
Evaluate AI-generated coding interactions end to end for usefulness, accuracy, and consistency with strong engineering practices.
Assess whether coding agents demonstrate sound technical reasoning and practical engineering judgment rather than just producing working-looking code.
Provide actionable feedback, distinguishing between adequate and exceptional AI response quality to shape evaluation standards.
Jobgether uses an AI-powered matching process to ensure your application is reviewed quickly and fairly against the role's core requirements. They are a third-party recruitment platform that partners with companies to fill positions, with a streamlined selection process.
Audit multiple-choice options and correct answers for technical accuracy, eliminating ambiguous distractors.
Verify coding question prompts and grading rubrics, and write additional edge test cases.
Format and return the final corrected exam in a valid JSON structure.
Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.
Analyze repositories for maintainability, dependency risks, insecure patterns, and build fragility.
Use AI coding agents to accelerate code review, vulnerability explanation, and remediation proposals.
Run quality and security scanning tools, support reproducible builds, and create evidence packs.
Deutsche Telekom IT Solutions provides IT and telecom services for large clients in Germany and Europe. With 5300+ employees, it is Hungary's most attractive employer, with sites in Budapest, Debrecen, Pécs, and Szeged.
Architect and deliver scalable microservices for cybersecurity range simulations, driving technical standards across systems.
Design and implement AI-enabled platform capabilities to support intelligent emulation behaviors and autonomous workflows.
Mentor engineers, lead complex initiatives, and establish best practices in system reliability, API design, and security-first engineering.
SimSpace is an AI Proving Ground that provides a realistic, live-fire cyber range for training, testing, and validation. Founded in 2015 by experts from U.S. Cyber Command and MIT Lincoln Laboratory, the company fosters a human-centered culture with continuous learning and professional growth, outperforming industry benchmarks in mobility and rewards.
Extend the self-service datastore platform with provisioning automation, guardrails, and paved paths for product engineering teams.
Ship observability, alerting, and backup/disaster recovery as built-in defaults for every datastore.
Convert recurring pull-in work into platform features or AI tooling that other teams can use directly.
Greenhouse provides a hiring software platform designed to make hiring work for everyone. They have an award-winning culture recognized by Fortune and Inc., and foster inclusivity, transparency, and accountability among their teams.
Build out the Agentic Platform as a product for internal teams, creating paved paths with sane defaults.
Design, build, and operate internal MCP servers, plus the standards, scaffolding, and registry for them.
Serve as technical lead for retrieval-augmented generation, including ingestion, chunking, embedding, and retrieval quality measurement.
Porch Group is a leading vertical software and insurance platform that helps homebuyers move, maintain, and protect their homes. The company has approximately 30 thousand business relationships, went public in 2020, and is building a global team across the US, Mexico, and India.
Provide Apache Airflow expertise directly to customers, solving challenging problems and optimizing configurations.
Learn and build expertise across software engineering disciplines including Airflow, Kubernetes, and Cloud Engineering.
Own the customer experience, working directly with customers to prioritize and resolve issues, and provide guidance on the path to production.
Astronomer empowers data teams to bring mission-critical software, analytics, and AI to life and is the company behind Astro, the industry-leading unified DataOps platform powered by Apache Airflow. Trusted by more than 800 of the world's leading enterprises, Astronomer lets businesses do more with their data.
Design and deliver significant components and core subsystems of our Kubernetes platform, such as secrets management, workload identity, storage, or cluster networking, from design through production operation.
Contribute to the architecture of distributed workloads, working with dependent teams to get runtime and isolation models right, while spending most time hands-on in code.
Own operability of built systems including SLOs, failure modes, upgrades, migrations, and on-call, and mentor earlier-career engineers.
ServiceNow is the AI control tower for business reinvention, bringing together any AI, any data, and any workflow to help 85% of the Fortune 500 work smarter, faster, and better. The company fosters an AI-native culture where technology and talent are unstoppable together, with a focus on freeing people from busywork.
Define and lead platform engineering strategy across complex, multi-environment cloud systems.
Architect scalable Kubernetes platforms, own IaC standards, and drive DevSecOps implementation.
Mentor engineers, partner with leadership on infrastructure direction, and lead complex migrations.
Robots & Pencils is an applied AI engineering firm that designs and ships AI co-workers for enterprise operations. Founded in 2009, with delivery centers in Canada, the US, Eastern Europe, and Latin America, we are a nimble team of senior engineers averaging 15+ years of experience.
Develop and evaluate AI training data for LLM and AI agent platforms.
Create coding tasks and write reference-quality solutions for evaluation.
Critically assess AI-generated code for correctness, security, and maintainability.
Toloka is a leading expert human data platform for AI agents and LLMs, providing high-quality training data. The company focuses on improving AI models through human feedback and structured evaluation.
Architect and maintain critical cloud platform components on AWS EKS with high availability and automated resilience.
Establish SRE standards including SLO/SLI tracking, error budget frameworks, and automated operational tooling.
Design and implement OpenTelemetry capture pipelines for telemetry data feeding downstream platforms.
Inflect is a US-based advisory and marketplace that revolutionizes how companies buy and sell digital infrastructure services. They operate with a focus on high-impact consulting and autonomous work.