ironquill.tech/board

$ cat jobs/research-scientist-remote-us-latam-see-posting-5b8e264da9ff.json

Research Scientist (Remote/US/LATAM)

See posting·EU·Europe (remote)·mid
llm
Apply on euremotejobs → Get AI match score →
Research Scientist, LLM Evaluations & Benchmarking Anyone AI LabsReports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don’t have reliable ways to measure what frontier models can and can’t do, and the field mostly runs on benchmarks that are saturated, contaminated,

Similar remote roles

AI Interaction Designer
bright vision technologies · US · mid
Chatbot Developer (WhatsApp, Telegram, Discord) - Freelance
mindrift · UK · mid
Stage de fin d'études / Ingénieur·e IA / Data
Talan · Worldwide · mid
Italian Audio QA Annotation Specialist
matchatalent · Worldwide · mid
Chatbot Developer (WhatsApp, Telegram, Discord) - Freelance
mindrift · APAC · mid
Senior Software Engineer, Backend - Platform (Core AI Automation)
Coinbase · Worldwide · senior
Staff Product Manager, ML Foundations and GenAI
Stripe · Worldwide · senior
Security Engineer
Stripe · Worldwide · mid