- Remoto
- Híbrido
- Tiempo completo
€28,000 - €42,000 a year
Sobre el rol Buscamos un profesional experimentado en salud mental con amplia experiencia en ensayos clínicos y expertise en evaluaciones psiquiátricas. Te incorporarás a nuestra red global de Evaluadores Independientes Centrales (EIC) para apoyar múltiples estudios de depresión y otros trastornos del estado de ánimo. En ...
YO AI Labs is seeking a Hebrew bilingual expert to remotely evaluate Hebrew audio content for linguistic quality and authenticity. You will provide detailed English feedback and justify decisions, adhering to established guidelines and deadlines. The role requires native Hebrew, at least B2+ English, and strong written ...
Spain, Córdoba
Ver vacante18808
Volga Partners is seeking an AI Language Quality Evaluator fluent in Greek and English for evaluating localization and UI bug detection tasks. This is a remote, freelance, task-based engagement with flexible hours and on-demand work suitable for supplemental income. Candidates should have a Bachelor’s degree in linguistics, ...
- Remoto
- Tiempo parcial
- Contrato
Up to US$40 an hour
... tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated ...
- Remoto
- Tiempo parcial
- Contrato
US$12 an hour
... provided text captions and task requirements. - Voice Quality: Evaluate whether generated voices sound natural and, where applicable, match the reference audio. - Evaluation & Feedback: Assign accurate scores and identify issues according to the provided evaluation guidelines. What we're looking for - Native-level fluency in ...
Have the right to work in Spain
... the processes. - You will prepare system review reports based on the collected and analyzed data. - You will coordinate follow-up groups to communicate process evaluations and trends. - You will review documentation that describes production operations and the functioning of equipment and facilities. - You will review qualification ...
Fully remote
... translate quantitative evaluation results into decisions for researchers, product teams, and senior leadership. Nice to have - Deep familiarity with open-source evaluation frameworks - Experience designing evaluations for agentic systems: tool use, multi-turn reasoning, environment-based benchmarks (SWE-bench, GAIA, WebArena-style), ...
Up to US$30,625 a year
About the role Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. About the Role You’ll design coding tasks that challenge frontier AI coding agents. Each task is ...
Spain, Barcelona
Ver vacanteUp to US$30,625 a year
... (pytest-cov, https://jobeax.com/link/khUWQLUnflUiPBzz, gcov, llvm-cov, kcov). - Fuzzing or property-based testing (Hypothesis). - Prior contribution to agent-evaluation benchmarks or related frameworks. What you'll get - Paid contributions, rates up to $35/hour *. - Task-based compensation equivalent to hourly rate, depending ...
Spain, Barcelona
Ver vacanteUp to US$70,000 a year
... testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated ...