About Mercor Mercors mission is to organize human intelligence to power the AI economy. Were a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid
Evaluate the quality, correctness, and reproducibility of software-engineering benchmark tasks used to train and evaluate a frontier AI labs models. Youll assess repository-level tasks, reference patches, test harnesses, and grading integrity — and provide clear, rubric-based
Evaluate the quality and correctness of AI-assisted software-development traces used to train and evaluate a frontier AI labs models. Youll assess end-to-end coding sessions produced with AI-assisted developer tools — judging correctness, workflow soundness, and reasoning
Evaluate the quality, correctness, and cloud-architecture soundness of AWS serverless and infrastructure-as-code tasks used to train and evaluate a frontier AI labs models. Youll assess multi-service serverless designs, IaC fidelity, and cross-service integration correctness — and
Evaluate the quality, correctness, and methodological rigor of applied machine-learning tasks used to train and evaluate a frontier AI labs models. Youll assess experiment design, model-selection reasoning, and evaluation methodology — and provide clear, rubric-based written
Evaluate the quality, correctness, and production-readiness of Kubernetes tasks used to train and evaluate a frontier AI labs models. Youll assess cluster-operations scenarios, manifest correctness, and failure-mode troubleshooting — and provide clear, rubric-based written feedback.Basic Qualifications
Join a leading AI labs cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced Large Language Models.1. Overview Join a leading AI labs
Evaluate the quality, fidelity, and completeness of vulnerability-reproduction and remediation tasks used to train and evaluate a frontier AI labs models. Youll assess whether CVE reproductions are faithful, fixes are sound, verification logic is rigorous, and
Own Your Intelligence Prime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team. Our platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training,
Own Your Intelligence Prime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team. Our platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training,
Own Your Intelligence Prime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team. Our platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training,
(Remote - North America, LATAM or Europe) Nango (YC W23) is a developer infrastructure company and the leading provider of API access for agents and apps. It enables AI applications to connect to the real world
(Remote - North America, LATAM or Europe) Nango (YC W23) is a developer infrastructure company and the leading provider of API access for agents and apps. It enables AI applications to connect to the real world
(Remote — North America, LATAM or Europe) Nango (YC W23) is a developer infrastructure company and the leading provider of API access for agents and apps. It enables AI applications to connect to the real world
What We Do Rillet serves accounting and finance teams, the financial brains of their companies. Our job is to help them run the numbers with impossible speed, accuracy, and insight. Rillet is the AI-native ERP built
What We Do Rillet serves accounting and finance teams, the financial brains of their companies. Our job is to help them run the numbers with impossible speed, accuracy, and insight. Rillet is the AI-native ERP built
Own Your Intelligence Prime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team. Our platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training,
Business Domain Expert Position: Business Domain Expert Type: Hourly Contract Compensation: $60–$80/hour Location: Remote About the Opportunity Mercor is conducting a structured evaluation pilot to measure how effectively frontier AI models perform real-world business tasks and how
Applied Health & Medicine Benchmark Specialist Position: Applied Health & Medicine Benchmark Specialist Type: Hourly Contract Compensation: $94–$119/hour Location: Remote About the Opportunity This opportunity is for experienced medical and health science professionals to contribute their
Applied Mathematics Benchmark Specialist Position: Applied Mathematics Benchmark Specialist Type: Hourly Contract Compensation: $61–$77/hour Location: Remote About the Opportunity This opportunity is for expert mathematicians and applied mathematics researchers to contribute to an advanced AI research