Careers

Build the marketplace that pays experts what their judgment is worth.

12 open staff roles across engineering, operations, trust & safety, expert supply, go-to-market, finance and design. Every band on this page is published, and every applicant gets a reply.

Looking for AI training work instead?
These are salaried roles at Cortext Labs, Inc. itself. If you came here to earn money evaluating AI output, writing reference answers or grading against rubrics, you want the expert marketplace — that is contract work, paid weekly, and it is a completely separate application.
94
Full-time employees, up from 11 in January 2025
4
Countries where Cortext Labs is the direct employer
41%
Of staff based outside the United States
12
Open roles on this page
The work

Four problems that consume most of our engineering time

Not a mission statement. These are the things that are actually hard here, and what you would be joining to work on.

01

Supply is perishable

A partner opens a programme on Monday needing 40 Portuguese-speaking clinicians for a six-week window. We have 42,000+ active experts and no reliable way to know which of them will be free on Thursday. Matching has to weigh a skills graph, credential state, availability that decays by the hour and prior quality — and be wrong in a way that is recoverable.
02

Friday has to work

Payouts run weekly across 130 countries through Stripe and Wise. Per-country minimums, tax-form gating, stale beneficiary details and rails that confirm asynchronously all conspire against a clean batch. A run that half-lands is worse than one that fails and retries, so the ledger, not the rail, has to be the system of record.
03

Grading the graders

There is no ground truth for whether a written rationale is good. We approximate it with gold-standard tasks seeded at about 4% of volume, inter-rater agreement thresholds and senior adjudication — then watch all of it drift the moment a lab revises its rubric mid-programme. Detecting that drift in days rather than weeks is an open problem here.
04

Telling people from models

The most damaging fraud on an AI-training marketplace is work generated by a language model and sold as expert judgment. Detection has to survive an adversary who adapts, while holding a precision bar high enough that honest contributors are not treated as suspects. Every flag routes to a human with the evidence attached, and every decision can be appealed.
Open roles

12 roles hiring now

Salary bands are published because withholding them wastes everyone's time. Every band is the real range we would make an offer in, not a headline number.

Team
Marketplace (7 engineers, 1 designer, 1 PM)
Reports to
Engineering Manager, Marketplace
Working mode
Hybrid — three days a week at 1 Sansome Street
Equity
0.05% – 0.12%, four-year vest, one-year cliff, ten-year exercise window
About the role

Marketplace owns everything an expert touches between finding a project and being matched to one: the Explore board, search and filtering, the shared application checklist, the screening flow, and the state machine that moves an application from submitted to matched to active. Roughly 3,400 applications a week pass through it, and about a third stall somewhere in the middle — usually at an identity check or a skills assessment that timed out.

The immediate work is unglamorous and high-leverage. The application checklist is currently one component holding eleven pieces of state, and every new lab program wants a variant of it. We want it decomposed into a declarative step definition so that adding a program-specific step — a language screen, an NDA, a licensure upload — is a data change rather than a release.

You would be the fifth engineer on the team and the second senior one. There is no separate front-end or back-end here: you will write the React, the Postgres queries behind it, and the background jobs that reconcile the two when they disagree.

What you'll do
  • Own the application state machine end to end — schema, transitions, background reconciliation, and the UI that renders it
  • Rebuild the application checklist as a declarative step definition so new lab programs ship without a code change
  • Cut the stall rate on in-progress applications, currently around 34%, by instrumenting each step and fixing the three worst drop-offs
  • Design and run the search and filtering layer on the Explore board as the catalogue grows past a few hundred concurrent programs
  • Write the migrations, run them yourself, and keep them reversible — we do not have a DBA
  • Carry a one-week primary on-call rotation roughly every seven weeks, covering marketplace services only
  • Review other engineers' pull requests with actual comments, and take the same in return
  • Write a one-page design memo before any project expected to take more than two weeks
What we're looking for
  • Five or more years building and operating production web applications, with real ownership of at least one system that other teams depended on
  • Strong TypeScript and React — you are comfortable reasoning about rendering, state colocation and why a component re-rendered
  • Practical relational database skills: you can read a query plan, add the right index, and explain why a migration is or is not safe under load
  • Experience with long-running, multi-step user flows where partial state and retries are the norm rather than the exception
  • You have debugged a production incident to root cause and written the follow-up yourself
  • Clear written English — most decisions here are made in a document, not a meeting
Nice to have
  • Next.js App Router in production, including the parts that are still awkward
  • Experience with marketplaces, two-sided supply, or anything where the hard problem was matching rather than throughput
  • Background jobs at scale — queues, idempotency, exactly-once illusions and their failure modes
  • You have worked somewhere between 20 and 150 people and can say what broke as it grew
The interview loop
01
Recruiter screen30 minutes

Scope of past ownership, motivation, location and compensation fit.

02
Hiring manager conversation45 minutes

One system you built, in depth — the constraints, what you rejected and what you got wrong.

03
Take-home work sample2 – 3 hours, paid at $150/hour

A small multi-step form flow with a resumable server-side state machine. We read the commit history and the README, not just the diff.

04
Work-sample review60 minutes

Extending your own submission live under a changed requirement, plus the tradeoffs you deferred.

05
Systems and data modelling60 minutes

Schema design and failure modes for a payout ledger. No whiteboard algorithms.

06
Operating principles45 minutes with an engineer outside the team

How you handle disagreement, unclear ownership and being wrong in public.

Apply for CX-ENG-114Send a CV and a short note. No cover letter, no portfolio site required.
Team
Platform (6 engineers)
Reports to
Head of Engineering
Working mode
Hybrid — three days a week at 1 Sansome Street
Equity
0.12% – 0.22%, four-year vest, one-year cliff, ten-year exercise window
About the role

Every Friday, Cortext runs a payout batch across 130 countries through Stripe and Wise. The batch is the single most consequential piece of software in the company: an expert who is not paid on time does not come back, and a payout run that half-lands is materially worse than one that fails outright and retries cleanly.

The platform team owns that batch, the double-entry ledger behind it, the tax-form collection that gates it (W-9 for US contributors, W-8BEN for everyone else), and the identity verification that decides whether a payout account may exist at all. Today the ledger and the payout executor share a database and too many assumptions. We want them separated, with the ledger as the system of record and the executor as a dumb, aggressively idempotent consumer.

This is a staff role because the hard part is not writing the code — it is sequencing a migration of live money movement without a maintenance window, and being the person who can say with evidence that Friday will be fine.

What you'll do
  • Own the payout pipeline: ledger, batch executor, rail adapters for Stripe and Wise, retry and reversal handling
  • Separate the ledger from the executor and migrate live balances without pausing a payout week
  • Drive failed-payout rate down from the current 1.9% of transfers, most of which are stale or mistyped beneficiary details
  • Build the reconciliation that proves, every Monday, that the ledger and both rails agree to the cent — and alert when they do not
  • Own the tax-document gate: collection, validation, expiry and the blocking rules that keep an unverified account from being paid
  • Set the engineering bar for the platform group: design review, incident review, and what we will not build
  • Partner with Finance on 1099-NEC generation and with Trust & Safety on account-takeover response
  • Carry primary on-call for money movement, including the Friday batch window
What we're looking for
  • Eight or more years in backend engineering, with at least three on payments, ledgers, billing or another domain where being wrong costs money
  • You have built or substantially rewritten a double-entry ledger and can explain why the naive schema fails
  • Deep experience with idempotency, exactly-once semantics and the compensating actions that replace them in practice
  • You have migrated a production system carrying real financial state, and can describe the rollback plan you actually needed
  • Comfort with third-party payment rails and their asymmetries — settlement timing, partial failures, and webhooks that arrive twice or never
  • Willingness to be on call for the payout window, and the judgment to know when to stop a batch
  • Written communication strong enough that Finance and Trust & Safety can act on your documents without a meeting
Nice to have
  • Cross-border contractor payouts, or anything involving KYC and sanctions screening at volume
  • Experience with Wise or Stripe Connect specifically, including their less-documented edges
  • Familiarity with US contractor tax reporting (1099-NEC, W-8BEN, backup withholding)
  • You have been the person who wrote the postmortem for a payment incident that reached customers
The interview loop
01
Recruiter screen30 minutes

Depth of payments experience, scope expectations, on-call appetite.

02
Head of Engineering conversation45 minutes

Where you have set technical direction, and where it was overruled.

03
Ledger design deep dive75 minutes

Model a payout ledger with holds, reversals, multi-currency and partial rail failure. Whiteboard-equivalent, no coding.

04
Migration planning exercise60 minutes

Sequence a live migration of balances with no downtime: cutover, dual-write, verification, rollback.

05
Incident retrospective45 minutes

Walk us through a real money-movement incident you owned, including your part in causing it.

06
Cross-functional interview45 minutes with Finance and Trust & Safety

Whether non-engineers can work with you when the stakes are high and the facts are incomplete.

Apply for CX-ENG-121Send a CV and a short note. No cover letter, no portfolio site required.
Team
Data (4 engineers, embedded with Research Operations)
Reports to
Engineering Manager, Data
Working mode
Hybrid — two days a week in the New York research operations office
Equity
0.02% – 0.06%, four-year vest, one-year cliff, ten-year exercise window
About the role

A lab partner sends us a task specification; we turn it into batches an expert can actually complete, route them, collect the submissions, run quality checks, and deliver a dataset back under a contractual schema. That round trip is the product, and it currently runs on a mix of Airflow DAGs, a Postgres staging schema and three scripts nobody wants to own.

Volume is the forcing function: submissions crossed 210,000 in the last quarter across 40+ lab partners, and every partner wants a slightly different delivery format, redaction rule and acceptance threshold. The team's job is to make partner-specific delivery a configuration rather than a fork.

You will spend as much time with Research Operations as with engineers. The people who know why a batch failed are usually programme managers, not services.

What you'll do
  • Own the ingestion-to-delivery pipeline for lab task data: batching, routing, submission capture, QA sampling, packaging and hand-off
  • Replace per-partner delivery scripts with a declared schema plus validation, so a new partner is onboarded in days rather than weeks
  • Build the QA sampling layer — gold-standard task injection at roughly 4% of volume, inter-rater agreement scoring, and the alerts when agreement drops
  • Instrument batch health so Research Operations can see, without asking, which batches are behind and why
  • Enforce redaction and retention rules on submission data, including partner-specific deletion windows
  • Backfill and reprocess historical batches when a guideline changes mid-programme, without corrupting delivered datasets
  • Keep pipeline cost visible and roughly flat as volume grows
What we're looking for
  • Four or more years building production data pipelines that other people depended on being correct
  • Strong Python and SQL; you write pipelines that are testable rather than pipelines that happen to have worked so far
  • Experience with a workflow orchestrator (Airflow, Dagster, Prefect or equivalent) including its operational failure modes
  • Data modelling for evolving schemas — you have handled a breaking upstream change without a full rewrite
  • Comfort owning data quality as a first-class product concern, not a downstream complaint
  • You can work directly with non-technical operators and translate their descriptions into checks
Nice to have
  • Human annotation, labelling or evaluation data specifically — inter-rater agreement, adjudication, rubric drift
  • dbt, or another opinionated transformation layer, in a team of more than two
  • Experience with data deletion and retention obligations under a customer contract
The interview loop
01
Recruiter screen30 minutes

Pipeline ownership, scale you have handled, New York hybrid fit.

02
Hiring manager conversation45 minutes

A pipeline you owned end to end and what its worst week looked like.

03
Pipeline design exercise60 minutes

Design ingestion-to-delivery for a new partner with an incompatible schema and a 48-hour SLA.

04
SQL and data modelling60 minutes

Real queries against a messy submissions schema, plus how you would restructure it.

05
Research Operations partner interview45 minutes

Whether a programme manager could get a straight answer from you at 6pm on a delivery day.

Apply for CX-ENG-108Send a CV and a short note. No cover letter, no portfolio site required.
Team
Quality Signals (3 engineers, 1 research scientist)
Reports to
Head of Engineering
Working mode
Fully remote within the US, with two company onsites a year
Equity
0.04% – 0.10%, four-year vest, one-year cliff, ten-year exercise window
About the role

Cortext is paid for expert judgment, so the question we cannot avoid is: how good is this submission, and how do we know? There is no ground truth for whether a written rationale is a good rationale. We approximate it with seeded gold-standard tasks, inter-rater agreement, adjudication by senior reviewers, and a set of learned signals that flag work for human review.

This role owns those learned signals. Concretely: scoring models that rank submissions for QA sampling, drift detection when a guideline changes mid-programme, and detection of submissions that were themselves generated by a language model — which, on an AI-training marketplace, is the failure mode that most directly damages the product we sell.

The constraint that shapes everything: a false positive costs a real person their income. Every flag this team produces routes to a human reviewer with the evidence attached, and we hold ourselves to a precision target rather than a recall one.

What you'll do
  • Build and maintain the submission scoring models that decide which of roughly 8% of sampled work gets human QA
  • Own LLM-generated-submission detection, including the evaluation set, the precision target and the review path for every flag
  • Detect calibration drift across reviewers and programmes, and surface it before a delivery misses its acceptance threshold
  • Design labelling and adjudication processes with Research Operations to generate the training data you need
  • Ship models as services the marketplace can call synchronously, with latency budgets and sane fallbacks
  • Run offline and online evaluation honestly, including publishing the cases where a model made things worse
  • Write the appeal path documentation with Trust & Safety, so an expert can contest a flag with a human
What we're looking for
  • Five or more years in applied ML, with models you shipped and then maintained through their decay
  • Solid grounding in evaluation: precision/recall tradeoffs, threshold selection under class imbalance, and why offline metrics mislead
  • Practical NLP experience, including fine-tuning or adapting language models for classification or scoring
  • You have built a system where a false positive harmed a person, and can describe the guardrails you added
  • Strong Python engineering — your models run in production without a separate team to rewrite them
  • Comfort designing annotation processes rather than assuming labelled data appears
Nice to have
  • Experience with RLHF, preference data, or model evaluation pipelines at a lab or data vendor
  • Anomaly and fraud detection where adversaries adapt to your detector
  • Published or internal work on inter-rater agreement and rubric calibration
The interview loop
01
Recruiter screen30 minutes

Applied ML background, shipping history, remote-US eligibility.

02
Technical screen60 minutes

A model you shipped: data, evaluation, thresholds, and how it degraded.

03
Applied ML case75 minutes

Design detection for LLM-generated submissions given an adaptive adversary and a hard precision floor.

04
Engineering interview60 minutes

Serving, latency budgets, retraining cadence and what you monitor.

05
Harms and appeals review45 minutes with Trust & Safety

How you reason about the cost of being wrong about a person.

Apply for CX-ENG-133Send a CV and a short note. No cover letter, no portfolio site required.
Team
Research Operations (11 people across New York and London)
Reports to
Director of Research Operations
Working mode
Hybrid — three days a week in the New York research operations office
Equity
0.02% – 0.05%, four-year vest, one-year cliff, ten-year exercise window
About the role

A lab programme is a contract with a schema, an acceptance threshold, a delivery calendar and a cohort of experts who have to be recruited, briefed, calibrated and kept. Delivery managers own that from kickoff to final hand-off. You would run three to five concurrent programmes, typically 40 to 300 experts each, worth $200,000 to $2M in delivered work over six to twelve weeks.

The recurring failure is not effort, it is calibration. A lab changes its rubric in week three; agreement between reviewers collapses; the delivery misses its acceptance threshold; and nobody notices until the weekly report. The job is to notice in day two, re-brief the cohort, and renegotiate the schedule with the lab before it becomes a credit note.

This is an operating role, not a coordination role. You will be reading submissions yourself, sitting in on adjudication calls, and telling a partner that their specification is ambiguous.

What you'll do
  • Own three to five concurrent lab programmes end to end: kickoff, cohort build, calibration, delivery, retrospective
  • Translate a partner's task specification into briefs, rubrics and worked examples an expert can act on without a call
  • Run calibration: seed gold-standard tasks, monitor inter-rater agreement weekly, and re-brief cohorts before agreement drops below the acceptance threshold
  • Hold the delivery schedule, including the uncomfortable conversation when a partner's mid-programme change makes it unachievable
  • Work with Expert Supply on cohort composition, attrition and backfill — cohorts typically lose 12–18% over a ten-week programme
  • Escalate quality disputes and pay disputes to the right owner within one business day, and close the loop with the expert
  • Run a written retrospective on every programme, with the numbers, and feed the fixes into the next one
  • Keep the programme margin honest: flag scope creep in writing before absorbing it
What we're looking for
  • Four or more years running operational programmes with external stakeholders and a real delivery deadline
  • Comfort with data: you can build the agreement report yourself in SQL or a spreadsheet rather than filing a ticket for it
  • Demonstrated ability to write instructions that non-specialists follow correctly the first time
  • Experience managing a distributed contributor or contractor population, including attrition and quality variance
  • Directness with clients — you have told a paying customer that their request is not possible, and kept the account
  • Availability that overlaps the US East Coast working day, with occasional flexibility for London and Bengaluru handoffs
Nice to have
  • Prior work in data labelling, annotation, market research fieldwork, or clinical trial operations
  • Exposure to AI or ML programmes, including familiarity with evaluation rubrics
  • A second language relevant to our expert base — Spanish, Portuguese, French, Hindi or Swahili
The interview loop
01
Recruiter screen30 minutes

Programme scale you have run, stakeholder exposure, New York hybrid fit.

02
Hiring manager conversation45 minutes

A programme that went wrong and what you did in the first 48 hours.

03
Written exercise90 minutes, take-home

Turn a deliberately ambiguous lab specification into a brief, a rubric and three worked examples.

04
Operating review60 minutes

Live: agreement has dropped to 0.61 in week three. Diagnose, decide, and write the partner email.

05
Cross-functional interview45 minutes with Expert Supply and Trust & Safety

How you balance delivery pressure against the experts doing the work.

Apply for CX-OPS-042Send a CV and a short note. No cover letter, no portfolio site required.
Team
Africa expert network (6 people, growing to 10 this year)
Reports to
VP of Operations
Working mode
Hybrid — three days a week in the Nairobi office
Equity
0.02% – 0.05%, four-year vest, one-year cliff, ten-year exercise window
About the role

The Nairobi office runs the Africa expert network: onboarding, verification support, community, quality coaching and the day-to-day of keeping several thousand contributors working well across the continent. It is the fastest-growing part of the expert base and the one where the standard playbook fits worst — payment rails behave differently, connectivity is uneven, and a mobile-first contributor experience is the default rather than an accommodation.

You would lead the team and own the outcomes: time from application to first paid task, first-90-day retention, and quality scores relative to the global cohort. Today the median African contributor waits longer than the global median at the identity-verification step, and closing that gap without weakening the check is the first order of business.

This role has real authority to change process, and a corresponding obligation to argue for it with evidence at the global operations review.

What you'll do
  • Lead and grow the Nairobi operations team, including hiring, performance and career development for six to ten people
  • Own onboarding throughput for the Africa network: application to first paid task, currently a median of 14 days against a global median of 9
  • Run quality coaching for contributors flagged by QA, with a documented path back to good standing rather than silent deactivation
  • Work with the Platform team on payout reliability across African corridors, and represent the corridor-specific failures that get lost in the global average
  • Build community: local expert meetups, a moderated forum, and a first-response channel with a same-business-day target
  • Partner with Trust & Safety on identity fraud patterns specific to the region, without letting suspicion default onto honest contributors
  • Report monthly on network health with numbers the rest of the company can act on
What we're looking for
  • Six or more years in operations, community or customer teams, including at least two managing people directly
  • You have run a distributed contributor, agent or gig population of at least 1,000 people
  • Track record of process design that survived contact with volume — documented, measured and revised
  • Fluent written English and the ability to hold your position in a room where the loudest voices are in San Francisco
  • Comfort with data: you build your own reports and defend the definitions in them
  • Existing right to work in Kenya
Nice to have
  • Swahili, French, Amharic, Yoruba or Hausa in addition to English
  • Experience with mobile money rails and their reconciliation quirks
  • Prior work with an international company where the head office was in a different timezone and a different mood
The interview loop
01
Recruiter screen30 minutes

Team scale managed, network size handled, right-to-work confirmation.

02
VP of Operations conversation45 minutes

A process you rebuilt and the metric that proved it worked.

03
Written exercise90 minutes, take-home

A plan to cut time-to-first-paid-task from 14 days to 9 without loosening identity verification.

04
Leadership interview60 minutes

Managing performance, developing people, and a decision you made that your team disagreed with.

05
Cross-functional interview45 minutes with Trust & Safety and Platform

Whether you can escalate a regional problem into a global roadmap.

Apply for CX-OPS-057Send a CV and a short note. No cover letter, no portfolio site required.
Team
Trust & Safety (5 investigators, 1 policy lead)
Reports to
Head of Trust & Safety
Working mode
Fully remote within EMEA, with quarterly team weeks in London
Equity
0.01% – 0.03%, four-year vest, one-year cliff, ten-year exercise window
About the role

Trust & Safety decides who is allowed to earn on Cortext. The caseload splits roughly into three: identity fraud (one person operating several accounts, or an account operated by someone other than the verified holder), work fraud (submissions produced by a language model and presented as expert judgment), and account takeover.

Every case ends in a decision that affects someone's income, so the standard is evidentiary rather than intuitive: a written case file, the specific signals relied on, the contributor's response where policy allows one, and a decision an appeals reviewer can audit six months later. We publish our reversal rate internally and treat a high one as a signal about our own thresholds.

This is deliberately not a queue-clearing role. Investigators are expected to spot the pattern behind a cluster of cases and write the policy or detection change that removes it, and roughly a fifth of your time is reserved for exactly that.

What you'll do
  • Investigate escalated fraud, identity and account-integrity cases to a written, auditable standard
  • Decide enforcement — warning, restriction, suspension or permanent removal — and communicate it in plain language the contributor can act on
  • Handle appeals on cases you did not decide, and reverse them when the evidence does not hold
  • Identify clusters and repeat patterns, and convert them into detection rules with the Quality Signals team
  • Maintain the case-file standard and keep decision quality consistent across the team through weekly calibration sessions
  • Contribute to policy: propose the wording change when a rule proves unenforceable or unfair in practice
  • Meet the same-day first-response target on safety reports raised through the Trust & Safety channel
What we're looking for
  • Three or more years in trust and safety, fraud investigation, financial crime, content moderation policy or a comparable investigative role
  • Demonstrated ability to write a case file that a stranger could audit without asking you questions
  • Sound judgment under ambiguity, and the discipline to record uncertainty rather than resolve it with confidence you do not have
  • Comfort querying data — you can pull the account graph yourself rather than requesting it
  • Resilience for the material: this role involves reviewing deception, occasional abusive communication, and, on red-team programmes, distressing model output
  • Working hours overlapping European business hours, and existing right to work in an EMEA country where we can employ
Nice to have
  • Experience with KYC, sanctions screening or document verification vendors
  • Familiarity with LLM-generated text detection and its limits
  • A second European language for contributor correspondence
  • Prior work on an appeals or adjudication function
The interview loop
01
Recruiter screen30 minutes

Investigative background, exposure tolerance, location and employment eligibility.

02
Hiring manager conversation45 minutes

A case you got wrong, and how you found out.

03
Case exercise2 hours, take-home, paid at $150/hour

Three anonymised case packets: reach a decision on each, write the file, and state what evidence would change your mind.

04
Case review60 minutes

Defending your three decisions against new evidence introduced live.

05
Policy and ethics interview45 minutes

Where enforcement and fairness conflict, and how you decide when both answers are defensible.

Apply for CX-TAS-019Send a CV and a short note. No cover letter, no portfolio site required.
Team
Expert Supply (8 people across San Francisco, London and Bengaluru)
Reports to
Head of Expert Supply
Working mode
Hybrid — three days a week in the Bengaluru office
Equity
0.02% – 0.04%, four-year vest, one-year cliff, ten-year exercise window
About the role

Expert Supply is recruiting, except the candidates are contributors, the roles are lab programmes, and the cycle time is days. When a partner opens a programme needing 60 licensed pharmacists across three timezones inside two weeks, this team either finds them or the programme does not run.

The APAC role owns supply across India, Southeast Asia, Japan, Korea and Australia — the fastest-growing region by applications and the one where credential verification is hardest, because degree and licence registries differ by country and several are not queryable at all.

The work is half sourcing channel management (universities, professional bodies, alumni networks, targeted paid acquisition) and half conversion: the drop between application and first paid task is where most of the value is lost, and it is measurable to the day.

What you'll do
  • Own APAC supply targets by speciality: medicine, law, engineering, quantitative finance and languages
  • Build sourcing channels beyond paid acquisition — professional bodies, university networks, specialist communities — and report cost per activated expert per channel
  • Run rapid cohort builds against programme deadlines, typically 40 to 200 verified experts in ten working days
  • Own credential verification workflow for APAC qualifications, including the countries where no central registry exists
  • Reduce application-to-first-task drop-off, currently 41% in the region, with interventions you can prove worked
  • Manage two supply associates, and grow the team as regional demand grows
  • Feed realistic supply forecasts into Go-to-market so we do not sell a cohort we cannot staff
What we're looking for
  • Five or more years in recruiting, talent operations, marketplace supply or high-volume sourcing
  • You have hit a hard numeric hiring or activation target repeatedly, and can describe the funnel that got you there
  • Experience owning a funnel metric end to end, with your own reporting
  • Credential or background verification experience in at least one regulated profession
  • Excellent written English and the judgment to write outreach that does not read as spam
  • Existing right to work in India
Nice to have
  • Experience recruiting clinicians, lawyers or other licensed professionals
  • Prior work at a marketplace, staffing platform or gig platform
  • Working knowledge of a second APAC language
  • Familiarity with paid acquisition mechanics and their measurement
The interview loop
01
Recruiter screen30 minutes

Funnel ownership, target attainment, right-to-work confirmation.

02
Hiring manager conversation45 minutes

A cohort build under deadline: channels, conversion and what you cut.

03
Sourcing plan exercise90 minutes, take-home

A plan to verify and activate 60 licensed pharmacists across three APAC markets in ten working days.

04
Plan review60 minutes

Defending your channel mix, cost assumptions and verification approach.

05
Cross-functional interview45 minutes with Research Operations

Whether delivery can trust your forecasts.

Apply for CX-SUP-028Send a CV and a short note. No cover letter, no portfolio site required.
Team
Go-to-market (4 people)
Reports to
VP of Go-to-market
Working mode
Hybrid — three days a week at 1 Sansome Street, plus partner travel
Equity
0.03% – 0.07%, four-year vest, one-year cliff, ten-year exercise window
About the role

Cortext sells to a small, technically sophisticated set of buyers: research and data leads at frontier AI labs and at the largest applied-AI teams inside enterprises. There are perhaps 120 accounts on earth worth calling, we already work with 40+, and the buyer can tell within ten minutes whether you understand evaluation.

Deals are $250,000 to $4M in annual committed spend, run three to six months, and are won on delivered quality rather than on a deck. Expect a paid pilot in almost every cycle: a small cohort, a real specification and an acceptance threshold you will be measured against.

The quota for this role is $4.5M in new committed spend in year one, with a ramp of two quarters. We publish the quota, the accelerators and the commission plan before you sign, and we do not change them mid-year.

What you'll do
  • Own a named list of 25 to 40 frontier lab and enterprise AI accounts end to end, from first contact to signed agreement
  • Run technical discovery deep enough to write the pilot specification yourself with Research Operations
  • Design and close paid pilots that convert — the current pilot-to-contract rate is 46%, and it is the number this role moves
  • Negotiate commercial terms with procurement and legal, including data handling, retention and indemnity provisions
  • Forecast honestly, in writing, weekly — an inflated forecast costs Expert Supply a cohort they cannot unbuild
  • Bring competitive and product feedback back with enough specificity to be actionable
  • Support renewal and expansion with the delivery team through the first year of every account you close
What we're looking for
  • Six or more years closing enterprise deals above $250,000 annual value, with quota attainment you can evidence
  • Genuine technical fluency in AI and ML — you can hold a conversation about evaluation, fine-tuning and data quality without a solutions engineer
  • Experience selling a service or data product where delivery quality, not features, was the deciding factor
  • You have negotiated with enterprise procurement and legal, and know which terms are worth escalating
  • Disciplined written forecasting and pipeline hygiene
  • Willingness to travel roughly 25% of the time, mostly domestic US
Nice to have
  • An existing network at AI labs or applied research teams
  • Prior experience at a data, annotation or evaluation vendor
  • Comfort with usage-based or committed-spend commercial models
The interview loop
01
Recruiter screen30 minutes

Deal sizes, quota history, territory and compensation fit.

02
VP of Go-to-market conversation45 minutes

Your largest closed deal, dissected: entry point, blockers, and what nearly killed it.

03
Discovery role play60 minutes

A live discovery call with a research lead who is sceptical of every data vendor they have used.

04
Account plan presentation60 minutes, prepared in advance

A written plan for one real account in our territory, presented to three interviewers.

05
Cross-functional interview45 minutes with Research Operations

Whether delivery believes what you would promise a customer.

Apply for CX-GTM-011Send a CV and a short note. No cover letter, no portfolio site required.
Team
Go-to-market (4 people)
Reports to
VP of Go-to-market
Working mode
Hybrid — three days a week in the London office
Equity
0.04% – 0.09%, four-year vest, one-year cliff, ten-year exercise window
About the role

Partnerships is the other side of supply. The fastest route to 200 verified radiologists is not paid acquisition — it is an agreement with a professional body, a university department or a specialist association that already has them and is willing to say we are legitimate.

This role owns those agreements across Europe, the Middle East and Africa: institutional partnerships that bring credentialled experts into the network, plus the co-marketing and revenue-share terms that make them worth signing for both sides. The London office already anchors the EMEA expert network, so you would be building on an existing base rather than from zero.

It also owns the harder conversation: what we will not agree to. Exclusive supply arrangements, data-sharing terms that conflict with our privacy commitments, and referral structures that read as payment for credentials are all off the table, and part of the job is saying so early.

What you'll do
  • Source, negotiate and close institutional partnerships across EMEA — professional bodies, universities, specialist associations and alumni networks
  • Own a target of 12 signed partnerships in the first year, measured on activated experts rather than logos
  • Structure commercial terms including referral fees, co-marketing and data handling, with Legal and Finance
  • Build the partner onboarding path with Expert Supply so a signed agreement produces verified contributors within 30 days
  • Represent Cortext at professional and academic conferences, and run the follow-up that makes the trip worth it
  • Maintain relationships after signature — the second-year renewal is where most partnerships quietly fail
  • Say no clearly and early to terms that conflict with our privacy, exclusivity or fairness commitments
What we're looking for
  • Six or more years in partnerships, business development or institutional sales, with agreements you owned from sourcing to signature
  • Experience negotiating with institutions rather than only companies — universities, regulators, professional bodies and their committee timelines
  • Ability to translate a commercial structure into terms Legal can draft without a translation layer
  • Strong understanding of EMEA data protection expectations, particularly around candidate and professional data
  • Excellent written English; a second European language is a real advantage
  • Existing right to work in the United Kingdom, and willingness to travel roughly 30% of the time across EMEA
Nice to have
  • An existing network among European professional bodies or university faculties
  • Experience in regulated professional markets — medicine, law, accountancy or engineering
  • Prior work at a marketplace or credentialling platform
The interview loop
01
Recruiter screen30 minutes

Partnership scope, institutional experience, right-to-work confirmation.

02
VP of Go-to-market conversation45 minutes

A partnership you signed that did not deliver, and what you learned about the diligence you skipped.

03
Partnership plan exercise2 hours, take-home, paid at $150/hour

A prioritised EMEA partnership map with targets, structures and a first-90-days plan.

04
Negotiation interview60 minutes

A live negotiation where the counterpart wants exclusivity and a data-sharing clause we cannot give.

05
Cross-functional interview45 minutes with Legal and Expert Supply

Whether your agreements are operable by the people who inherit them.

Apply for CX-GTM-016Send a CV and a short note. No cover letter, no portfolio site required.
Team
Finance (4 people)
Reports to
Controller
Working mode
Fully remote within the US, with two company onsites a year
Equity
0.01% – 0.03%, four-year vest, one-year cliff, ten-year exercise window
About the role

Cortext paid $18.4M to experts in 2025 across 130 countries, in weekly batches, to people who are contractors rather than employees. That produces a month-end close with an unusual shape: high transaction volume, many currencies, two payment rails, and a tax-reporting obligation that lands as a single hard deadline every January.

This role owns contractor payments accounting: accruals for work completed but not yet paid, reconciliation between the internal ledger and both rails, foreign exchange treatment, and the 1099-NEC cycle. It also owns the boring, load-bearing work of keeping W-9 and W-8BEN records complete enough that January is a process rather than an emergency.

Close is currently eight business days. The first objective of this role is six, without loosening a single control.

What you'll do
  • Own the contractor payments cycle in the close: accruals, cut-off, reconciliation to Stripe and Wise, and unresolved-item follow-up
  • Reconcile the internal ledger to both payment rails weekly and investigate every variance, not only material ones
  • Own the annual 1099-NEC cycle end to end, including TIN matching, corrections and contributor queries
  • Maintain W-9 and W-8BEN completeness, and work with Platform on the gating rules that keep an unverified payee unpaid
  • Handle multi-currency accounting and FX revaluation for non-USD payouts
  • Cut close from eight business days to six by automating reconciliation rather than by skipping steps
  • Support the first external audit, including documenting controls that currently exist only in practice
  • Answer contributor payment queries escalated by Support with an accurate answer within one business day
What we're looking for
  • Five or more years in accounting, including US GAAP month-end close ownership
  • Direct experience with contractor or marketplace payouts at volume, or with a high-transaction payments environment
  • Working knowledge of US contractor tax reporting: 1099-NEC, W-9, W-8BEN, backup withholding and TIN matching
  • Strong reconciliation discipline and genuine comfort with large datasets — advanced spreadsheet skills, and SQL is a real plus
  • Experience with multi-currency transactions and FX revaluation
  • Clear written communication with people who are not accountants, including contributors asking why a payment differs from their expectation
Nice to have
  • CPA, or an equivalent qualification in progress
  • Experience preparing for a first external audit at a venture-backed company
  • Familiarity with Stripe Connect or Wise reporting exports
  • Exposure to non-US contractor reporting obligations
The interview loop
01
Recruiter screen30 minutes

Close ownership, payouts experience, remote-US eligibility.

02
Controller conversation45 minutes

How you have shortened a close, and which controls you refused to remove.

03
Technical exercise90 minutes, take-home

Reconcile a deliberately broken week of payout data across two rails and explain each variance.

04
Exercise review60 minutes

Your reconciliation logic, materiality judgment and what you would automate first.

05
Cross-functional interview45 minutes with Platform Engineering

Whether you can specify a financial control precisely enough for an engineer to build it.

Apply for CX-FIN-007Send a CV and a short note. No cover letter, no portfolio site required.
Team
Design (2 designers)
Reports to
Head of Product
Working mode
Hybrid — three days a week at 1 Sansome Street
Equity
0.04% – 0.09%, four-year vest, one-year cliff, ten-year exercise window
About the role

You would be the second designer, working on everything an expert sees: the Explore board, the application checklist, the screening flow, the task workspace where the actual work happens, and the earnings and payout screens where trust is either built or lost.

The interesting constraint is that our users are experts at something other than software. A consultant radiologist reading a rubric at 11pm on a phone is not going to explore the interface; the design has to be unambiguous on first read, in a second language, on a mid-range Android device, on an uneven connection.

The second constraint is money. Screens that explain what someone earned, why a submission was rejected, or when a payout will land are the highest-stakes surfaces in the product. They are also, today, the least designed. Fixing that is the first project.

What you'll do
  • Own design for the expert-facing product end to end: research, interaction, visual, and the specifics engineering needs to build it
  • Redesign the earnings, submission-feedback and payout surfaces so an expert can answer what was paid, what was rejected and why, without contacting support
  • Run lightweight research with real contributors — six to eight sessions per project, including non-native English speakers and mobile-only users
  • Design for a mobile-first, low-bandwidth reality, and pressure-test on a mid-range Android device rather than a simulator
  • Extend the design system with the second designer, and keep it honest about what is actually implemented
  • Bring accessibility into the default: contrast, focus order, screen-reader labelling and keyboard paths, tested rather than assumed
  • Write the copy in your flows first, then refine it with the team — the words are the interface here
What we're looking for
  • Four or more years designing production software, with shipped work you can walk through in detail
  • A portfolio showing complex, multi-step flows rather than marketing pages — we care about the third and fourth screen
  • Genuine interaction design skill, plus enough visual craft to make it feel finished
  • You have run your own user research and changed your design because of it
  • Ability to write clear product copy, including error and rejection messaging
  • Comfort working closely with engineers, including reviewing implementation and pushing back on drift
Nice to have
  • Experience designing for users outside your own language and market
  • Accessibility work you can evidence, ideally to a WCAG standard
  • Design systems work in a small team where you had to keep it small
  • You can prototype in code well enough to test an interaction
The interview loop
01
Recruiter screen30 minutes

Scope of shipped work, portfolio depth, San Francisco hybrid fit.

02
Portfolio review60 minutes

Two projects in depth: the constraints, the rejected directions and the outcome after launch.

03
Design exercise3 hours, take-home, paid at $150/hour

Redesign the earnings screen for a mobile-only contributor whose latest submission was partially rejected.

04
Exercise critique60 minutes

How you take and use critique, and how you would test the design you produced.

05
Engineering collaboration interview45 minutes

Whether an engineer can build your work without inventing the missing states.

Apply for CX-DES-005Send a CV and a short note. No cover letter, no portfolio site required.
Nothing that fits?
Send a CV and two paragraphs on what you would want to own to careers@cortext-ai.uk. We keep speculative applications on file for six months and genuinely do go back to them — three of the roles above were written around someone who wrote in first.
How we operate

Four principles, and what each one costs

Values are only informative if they trade something away. Each of these names the thing it gives up.

The expert is the customer, even when the lab pays the bill

Labs pay the invoices; contributors do the work and can leave at any time. When those interests conflict — a rate that is too low for the difficulty, a rejection rate a partner wants raised, a deadline that only works if people work through a weekend — we resolve it toward the contributor and tell the partner why.

What it costsThis costs revenue in a countable way. We have declined programmes over pay floors and lost at least one renewal over a rejection-threshold argument. Go-to-market carries that cost in their quota, and nobody is expected to pretend it is painless.

Write it down before you build it

Anything expected to take more than two weeks starts as a one-page memo: the problem, the options, the choice and what would prove it wrong. It circulates for two days of comments before work begins, and the decision is recorded in the same document rather than in a meeting nobody minuted.

What it costsIt is slower to start and genuinely uncomfortable for people who think by prototyping. We accept losing the first week to gain the ability to reconstruct, a year later, why a system is the way it is — and to onboard people from documents instead of interruptions.

Measure quality, not throughput

No operations role at Cortext is compensated on volume of tasks delivered, and no engineering team is measured on tickets closed. The metrics that carry weight are acceptance rate, inter-rater agreement, dispute reversal rate and contributor retention at 90 days.

What it costsWe ship less per quarter than we could and our delivery timelines are longer than some competitors quote. It also means a team can be doing excellent work in a quarter where the output chart is flat, and managers have to be able to argue that in a review.

Small teams, long ownership

Teams cap at around eight people and own a problem for years rather than a project for a quarter. Whoever builds a system carries its pager, answers its support escalations and writes its postmortems. Six people own payouts end to end, including the Friday batch.

What it costsInternal transfers are rarer than at a company that rotates people, and you will spend real time on the unglamorous maintenance of things you shipped 18 months ago. If your preference is to build and hand off, this will grate.

Say the number

Salary bands are published on every opening and inside the company. Programme margins, attrition, dispute reversal rates and missed deliveries are shared at the monthly all-hands with the actual figures. Performance feedback names the specific behaviour and the specific consequence rather than gesturing at a theme.

What it costsTransparency removes the comfortable ambiguity people sometimes rely on. Knowing what your colleagues are paid is not always pleasant, and a review that names a behaviour is harder to hear than one that does not. We think the alternative is worse, but we do not pretend it is free.
Benefits

What you get, with the numbers attached

Written out in full because 'competitive benefits' tells a candidate nothing they can compare.

Health
100% of employee premiums, 80% for dependants

Medical, dental and vision. In the US that is a choice of three plans including a $0-deductible PPO; outside the US it is private cover at an equivalent level through a local provider, plus a top-up where the national system already covers the basics. Coverage starts on day one, not after 30 days.

Equity
Options with a ten-year exercise window

Four-year vesting, one-year cliff, monthly thereafter. If you leave, you keep ten years from grant to exercise rather than the standard 90 days — so you are never forced to choose between changing job and abandoning vested equity. Early exercise is permitted, and we tell you the strike price, the preferred price and the outstanding share count before you sign.

Retirement
401(k) with a 4% match, vested immediately

Dollar-for-dollar on the first 4% of salary, with no vesting schedule on the match. UK staff get 6% employer pension contribution; India is EPF plus a 4% supplementary contribution; Kenya is NSSF plus a matched private scheme.

Time off
25 days, with a 15-day minimum you actually have to take

25 days of paid leave plus local public holidays, plus a company-wide shutdown between 24 December and 1 January that does not come out of your allowance. Sick leave is separate and uncapped within reason. Managers are measured on whether their team takes at least 15 days; unlimited-PTO policies reliably produce less time off, so we do not run one.

Family
16 weeks fully paid parental leave, from day one

The same 16 weeks for every new parent regardless of gender or route to parenthood, including adoption and surrogacy, with no tenure requirement. Return is phased: four weeks at 80% hours on 100% pay. We also cover $10,000 per year toward fertility treatment, adoption or surrogacy costs.

Learning
$2,000 a year, rolling over to $4,000

Courses, books, conferences and certifications, approved by your manager rather than by Finance. Unspent budget rolls forward one year to a $4,000 ceiling, so a conference in your second year does not require skipping learning in your first.

Workspace
$1,500 to set up, $500 every two years after

Yours to spend on desk, chair, monitor and peripherals; the hardware itself (laptop, external display) is provided separately and is not counted against it. Fully remote staff also get $80 a month toward connectivity, and a co-working desk where working from home is not practical.

Together
Two company onsites a year, five days each

One in San Francisco, one rotating across the hub cities, with travel, accommodation and time zone recovery paid. Distributed teams meet for a further team week each quarter. Remote staff are not expected to fund their own visibility.

Wellbeing
12 fully covered therapy sessions a year

Through the employer plan, with no referral needed and nothing reported back to Cortext beyond aggregate usage. Trust & Safety staff additionally get monthly clinical supervision, a capped daily exposure limit on distressing material, and a rotation off high-exposure queues every six weeks.

Moving
Up to $15,000 relocation, and we pay the immigration bills

For hires relocating to a hub city. We sponsor work visas where the role and market allow — H-1B transfers and cap-subject filings, O-1, UK Skilled Worker, and Indian and Kenyan equivalents — and we pay the legal fees and government charges for you and your dependants rather than reimbursing them later.

Hiring process

What each stage tests, and how long it takes

Role-specific loops are listed on each opening. This is the shape every one of them follows.

01
Application reviewDecision within 5 business days

Whether your experience matches the specific problems in the role description. Every application is read by a person on the hiring team — there is no automated resume filter, and no ranking model deciding who is seen.

We reply to every applicant, including the ones we decline. If you have not heard back in 5 business days, the email address on the role is the right place to chase us.

02
Recruiter screen30 minutes, video

Scope of your past work, what you are looking for, and the practical constraints — location, right to work, notice period and compensation expectations. We state the band on the call if you have not already read it here.

03
Hiring manager conversation45 minutes, video

One piece of work you owned, in depth: the constraints you faced, the options you rejected, and what you would do differently. This is the round where we decide whether the level is right, and we will tell you if we think it is not.

04
Work sample90 minutes to 3 hours, on your own time

The actual work of the role, on a problem drawn from something we have really faced. Any exercise expected to take more than two hours is paid at $150 an hour, invoiced or paid through the same rails our contributors use.

You may use AI tools on the work sample. We ask you to say where you did, because the follow-up round assumes you can defend every line of what you submitted.

05
Interview loop3 to 4 hours, split across two days if you prefer

Role-specific depth — the exact rounds are listed on each opening, so you know what is coming before you agree to it. Interviewers get your work sample in advance and are expected to have read it.

06
Operating principles interview45 minutes with someone outside the hiring team

How you behave when you disagree, when ownership is unclear, and when you turn out to be wrong. It carries a real veto, and it is the round most likely to change a decision.

07
References and offerOffer within 3 business days of the final round

Two references of your choosing, taken after the loop and never before you have told your current employer. The offer is the band published on the role, positioned by level and interview evidence.

We do not ask for salary history, we do not counter competing offers, and we do not reward negotiation with a higher number — the same evidence produces the same offer for two candidates at the same level.

Application to offer
19 calendar days at the median, 31 at the 90th percentile. The slowest part is usually scheduling the loop around your working week, not our decision-making.
Reply after each stage
Within 3 business days, including a decline. If a decision is genuinely delayed, we tell you why and give a new date rather than going quiet.
Feedback on request
Available after the work sample and after the loop. It is specific and written; in some jurisdictions we limit what we can say, and we will tell you when that is the case.
Work-sample pay
$150 an hour for any exercise expected to exceed two hours, paid whether or not you are hired, and whether or not you complete it.
Interview recording
None. We do not record interviews, do not use automated video scoring, and do not run AI assessments of candidates.
Reapplying
Any time for a different role. For the same role, we ask for six months unless the hiring manager invites you back sooner, which happens more often than people expect.
Access

Adjustments we will make, without asking why

You do not need a diagnosis, a letter, or a reason. Ask the recruiter at any point, including after the loop has started.

  • Extra time on any work sample, or an untimed version of it
  • Interview questions and the exercise brief sent 24 hours in advance
  • A written or asynchronous alternative to a live exercise
  • Audio-only interviews instead of video, with no explanation required
  • Live captions, or an interpreter — including sign language — arranged and paid for by us
  • Screen-reader-compatible materials, and exercises that do not depend on a specific IDE or design tool
  • Scheduled breaks, or a loop split across three or four shorter days
  • Scheduling around religious observance, caring responsibilities, medical appointments or a current job
  • A step-free interview location and accessible facilities for any on-site round
Equal opportunity

Cortext Labs, Inc. hires on ability to do the job. We do not screen on race, colour, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, veteran status, marital status, pregnancy, caring responsibilities, or genetic information — and where local law lets an employer ask about any of those, we do not ask.

We do not ask for salary history in any jurisdiction, including the ones where it is still legal. Offers are made against the published band and the level we assessed you at.

Where we can employ you

Cortext Labs is the direct employer in the United States, the United Kingdom, Kenya and India — San Francisco, New York, London, Nairobi, Bengaluru. Remote roles are open across those countries and, for some positions, the wider region named on the listing.

We do not currently sponsor visas for roles outside the US, and we say so on the listing rather than at the offer stage.

About these listings
Cortext Labs, Inc. is a demonstration build and the openings above are illustrative — the requisition references, salary bands and interview loops were written to show what a real careers page looks like, not to fill real jobs. Please do not treat any of them as a live vacancy.

Read the loop before you apply.

Every role lists exactly what each round tests. If a stage looks like a waste of your time, tell the recruiter — we have cut rounds because a candidate was right.

See open roles →About Cortext