DameTrabajoEntrar

Research Scientist (Remote/US/LATAM)

Anyone Ai · Fully Remote; Uruguay; Chile; United Kingdom; Ecuador - Fully Remote; Mexico - Fully Remote; Colombia - Fully Remote; France; Peru - Fully Remote; Brazil; Argentina - Fully Remote · Remoto

Es un puesto remoto: contratan desde México, Colombia, Perú, Argentina y 4 países más.

El aviso no publica el sueldo.

Lo publica Anyone Ai y está vigente desde el 6 de julio de 2026.

Toca Postularme y entra con tu cuenta de Google: te llevamos al aviso en Ashby y te ayudamos a armar el CV para este puesto.

Postularme

Descripción del puesto

Research Scientist, LLM Evaluations & Benchmarking

Anyone AI Labs

Reports to: CEO

The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or measuring the wrong thing. You'll own that problem at Anyone AI, measuring frontier model capability.

This is a research role at heart: you decide what a good evaluation is , design the benchmarks that prove it, and defend the methodology under lab scrutiny. You'll build frontier-grade evaluation packages across reasoning, coding, agents, tool use, and multi-modal — grounded in expert-verified truth, validated against multiple models, and QC'd to survive buyer-side review.

Responsibilities

What we're looking for

Toca Postularme y entra con tu cuenta de Google: te llevamos al aviso en Ashby y te ayudamos a armar el CV para este puesto.

Postularme

Avisos parecidos

Preguntas frecuentes

¿Es remoto el puesto de Research Scientist (Remote/US/LATAM)?

Es un puesto remoto: contratan desde México, Colombia, Perú, Argentina y 4 países más.

¿Cuánto paga?

El aviso no publica el sueldo.

¿Dónde se publicó este aviso?

En Ashby. DameTrabajo lo encontró ahí y te lleva a postularte en el aviso original.