Select your area of expertise below to browse active hiring waves and start your application.
Upload your resume to instantly see which open roles match your expertise.
Audit LLM hallucinations and enforce strict safety constraints within senior python / ml evaluator domains.
Audit LLM hallucinations and enforce strict safety constraints within c++ systems & performance domains.
Audit LLM hallucinations and enforce strict safety constraints within frontend / react architecture domains.
Design complex prompting scenarios and analyze model outputs for accuracy in data engineering / sql architect.
Curate high-quality, expert-level training datasets specifically for security & cryptography.
Curate high-quality, expert-level training datasets specifically for general swe / leetcode master.
Curate high-quality, expert-level training datasets specifically for rust backend engineer.
Evaluate reasoning chains and logic paths for frontier models using go cloud infrastructure expertise.
Review AI-generated code/text against domain-specific rubrics for ai research scientist tasks.
Audit LLM hallucinations and enforce strict safety constraints within devops / sre architect domains.
Evaluate reasoning chains and logic paths for frontier models using blockchain / web3 developer expertise.
Review AI-generated code/text against domain-specific rubrics for ios / swift engineer tasks.
Audit LLM hallucinations and enforce strict safety constraints within android / kotlin expert domains.
Evaluate reasoning chains and logic paths for frontier models using embedded systems c/c++ expertise.
Curate high-quality, expert-level training datasets specifically for oncology ai reviewer.
Design complex prompting scenarios and analyze model outputs for accuracy in neurology / bci annotator.
Evaluate reasoning chains and logic paths for frontier models using general practitioner evaluator expertise.
Evaluate reasoning chains and logic paths for frontier models using pharmacology expert expertise.
Audit LLM hallucinations and enforce strict safety constraints within surgical methodology domains.
Design complex prompting scenarios and analyze model outputs for accuracy in radiology / mri annotator.
Curate high-quality, expert-level training datasets specifically for cardiology diagnostic reviewer.
Design complex prompting scenarios and analyze model outputs for accuracy in pediatrics case specialist.
Curate high-quality, expert-level training datasets specifically for psychiatry / therapy ai analyst.
Design complex prompting scenarios and analyze model outputs for accuracy in dermatology vision model expert.
Evaluate reasoning chains and logic paths for frontier models using pathology slide reviewer expertise.
Design complex prompting scenarios and analyze model outputs for accuracy in quant dev / algo trader.
Review AI-generated code/text against domain-specific rubrics for actuarial science reviewer tasks.
Curate high-quality, expert-level training datasets specifically for corporate finance / m&a.
Design complex prompting scenarios and analyze model outputs for accuracy in phd level mathematician.
Review AI-generated code/text against domain-specific rubrics for tax strategy & compliance tasks.
Curate high-quality, expert-level training datasets specifically for hedge fund risk analyst.
Audit LLM hallucinations and enforce strict safety constraints within derivatives pricing expert domains.
Evaluate reasoning chains and logic paths for frontier models using crypto economics researcher expertise.
Review AI-generated code/text against domain-specific rubrics for venture capital analyst tasks.
Review AI-generated code/text against domain-specific rubrics for macroeconomics forecaster tasks.
Design complex prompting scenarios and analyze model outputs for accuracy in legal document reviewer (bilingual).
Review AI-generated code/text against domain-specific rubrics for japanese/english cultural context tasks.
Curate high-quality, expert-level training datasets specifically for arabic natural language processing.
Audit LLM hallucinations and enforce strict safety constraints within spanish / english localization domains.
Review AI-generated code/text against domain-specific rubrics for mandarin tone & nuance evaluator tasks.
Curate high-quality, expert-level training datasets specifically for french medical translation.
Review AI-generated code/text against domain-specific rubrics for german financial documentation tasks.
Design complex prompting scenarios and analyze model outputs for accuracy in hindi speech-to-text annotator.
Curate high-quality, expert-level training datasets specifically for korean sentiment analysis.
Evaluate reasoning chains and logic paths for frontier models using russian political context reviewer expertise.
Design complex prompting scenarios and analyze model outputs for accuracy in creative writer / fiction author.
Review AI-generated code/text against domain-specific rubrics for journalism fact checker tasks.
Audit LLM hallucinations and enforce strict safety constraints within history / humanities expert domains.
Evaluate reasoning chains and logic paths for frontier models using instruction following evaluator expertise.
Design complex prompting scenarios and analyze model outputs for accuracy in prompt engineering specialist.