evaluation
1,000 ofertas de empleo de evaluation en México. Encuentra ofertas actualizadas diariamente de los principales portales de empleo.
-
AI Evals Engineer — Evaluation Datasets
Hace 1 semana
distrito federal, distrito federal, México Prophetic Technologies Inc Jornada completaProphetic Software is not able to sponsor employment visas now or in the future. Candidates must be authorized to work in the United States without current or future sponsorship to be considered for this role.About Prophetic:Real estate development is a multi-billion-dollar industry that has run on fragmented data, manual processes, and gut instinct for...
-
New Business Evaluation Manager
Hace 1 semana
Nuevo León, Nuevo León, México Metalsa Jornada completaAbout the CompanyWe are a global company with 65+ years of experience in the automotive industry. We manufacture safe and sustainable products for people around the world. We are working for a better future where we enrich communities every day by being committed to people, innovation and our planet. If you have what it takes to accelerate Metalsa's vision...
-
New Business Evaluation Manager
Hace 1 semana
Estado de Nuevo León, Estado de Nuevo León, México Metalsa Jornada completaAbout the Company We are a global company with 65+ years of experience in the automotive industry. We manufacture safe and sustainable products for people around the world. We are working for a better future where we enrich communities every day by being committed to people, innovation and our planet. If you have what it takes to accelerate Metalsa's...
-
Frontend Code Evaluation Specialist
Hace 1 semana
tlalnepantla de baz, estado de méxico Mercor Jornada completaAbout the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey. Position: Frontend Engineer Type: Contract Compensation: $90/hour Location: Remote Role Responsibilities - Render a...
-
AI Image Evaluation Analyst
Hace 1 semana
México, México iMerit Jornada completaThe work iMerit is looking for detail oriented analysts to evaluate and rank AI generated responses to image based prompts. You will judge answers on accuracy, relevance, clarity, conciseness, safety, localization, and how well they follow the user's instructions, then explain your reasoning in writing. Much of the job comes down to this: look at the...
-
Head of Search Evaluation
Hace 1 semana
Ciudad de México, Ciudad de México Straive Jornada completa $600,000 - $900,000 Por obraStraive is a global leader in enterprise-grade data analytics and AI solutions, actively seeking a L2 Team Lead for the Search Evaluation/Content Moderation project. You will ensure overall quality and operational success, master complex guidelines, and bridge policy owners with the onsite execution team.As a leader, you will train and transfer knowledge,...
-
Team Lead
Hace 1 semana
Ciudad de México, Ciudad de México Straive Jornada completa $600,000 - $900,000 Por obraStraive is a global leader in enterprise-grade data analytics and AI solutions, committed to empowering businesses across various industries with cutting-edge technology and expert insights. Backed by EQT, a top private equity firm, we are uniquely positioned to drive innovation through significant investments and an entrepreneurial spirit. Our core focus is...
-
Speech AI Evaluation Specialist
Hace 1 semana
santiago de querétaro, querétaro, México Rws Trainai Jornada completaWe are looking for Speech AI Evaluation Specialist to support the improvement of AI-generated content in Spanish (Mexico). Job Type: Freelance Location: Mexico (work from home) Work Schedule: Part-time - 10+ hours per week. Flexible - work whenever you want. Start Date: Immediately Duration: TBC Rate: 7 USD per hour Help Shape the Future of AI Are you a...
-
AI Evals Engineer: Ground Truth
Hace 1 semana
Ciudad de México, Ciudad de México Prophetic Technologies Inc Jornada completa $2 - $3 Por obraProphetic Software is seeking a full-time ML evaluation engineer to design and own ground-truth datasets for our AI stack. You will define correctness, create label schemas, and ensure data quality across modules.You will work with product and engineering to source data, calibrate automated judges, and report metrics to accelerate product iterations. Strong...
-
AI Evals Engineer — Evaluation Datasets
Hace 1 semana
Ciudad de México, Ciudad de México Prophetic Technologies Inc Jornada completa $2 - $3 Por obraProphetic Software is not able to sponsor employment visas now or in the future. Candidates must be authorized to work in the United States without current or future sponsorship to be considered for this role.About Prophetic:Real estate development is a multi-billion-dollar industry that has run on fragmented data, manual processes, and gut instinct for...
-
Generative Audio Evaluation Premium
Hace 2 semanas
Ciudad de México RWS Trabajo remoto Jornada completa USD 8 IndefinidoWe are looking for Generative Audio Evaluation Specialists! A great entry point into ongoing work within one of our most active AI markets! Job Type: Freelance Location: Remote (Mexico) Work Schedule: Part-time - 10+ hours per week. Flexible - work whenever you want! Rate: 8 USD/hour Help Shape the Future of AI We’re currently hiring Generative Audio...
-
Generative Audio Evaluation
Hace 2 semanas
City, CMX, México RWS Trabajo remoto Jornada completaWe are looking for **Generative Audio Evaluation** **Specialists** ! A great entry point into ongoing work within one of our most active AI markets! **Job Type:** Freelance **Location:** Remote (Mexico) **Work Schedule:** Part-time - 10+ hours per week. Flexible - work whenever you want! **Rate:** 8 USD/hour **Help Shape the...
-
Remote AI Evaluation
Hace 2 semanas
Mexico, chihuahua AI Chopping Block Jornada completaMindrift connects specialists with project-based AI opportunities for leading tech companies, focusing on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.We are building a dataset to evaluate AI coding agents and design tasks from intermediate states of simulated environments. The agent writes most of...
-
AI Image Evaluation Analyst
Hace 2 semanas
, México iMerit Technology Jornada completaThe work iMerit is looking for detail oriented analysts to evaluate and rank AI generated responses to image based prompts. You will judge answers on accuracy, relevance, clarity, conciseness, safety, localization, and how well they follow the user's instructions, then explain your reasoning in writing. Much of the job comes down to this: look at the...
-
AI Response Evaluation Analyst
Hace 2 semanas
, México iMerit Technology Jornada completaiMerit is seeking AI Response Evaluation Analyst to participate in a project where you will review AI generated responses and decide which ones actually hold up: accurate, relevant, concise, safe, well localized, and true to what the user asked for. Every ranking comes with a short written rationale, and those comments feed directly into how the models...
-
AI Response Evaluation Analyst
Hace 2 semanas
México, México iMerit Jornada completaiMerit is seeking AI Response Evaluation Analyst to participate in a project where you will review AI generated responses and decide which ones actually hold up: accurate, relevant, concise, safe, well localized, and true to what the user asked for. Every ranking comes with a short written rationale, and those comments feed directly into how the models...
-
Research Scientist
Hace 2 días
Ciudad de México Anyone Ai Jornada completaResearch Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...
-
Research Scientist
Hace 2 días
mérida, yucatán, México Anyone Ai Jornada completaResearch Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...
-
Research Scientist
Hace 2 días
celaya, guanajuato, México Anyone Ai Jornada completaResearch Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...
-
Research Scientist
Hace 2 días
iztacalco, distrito federal, México Anyone Ai Jornada completaResearch Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...