evaluation
10,000 ofertas de empleo de evaluation en México. Encuentra ofertas actualizadas diariamente de los principales portales de empleo.
-
Junior Software Engineer — Product QA
Hace 1 semana
Monterrey, Chiapas, México Cultured Supply Jornada completaRole Description Cultured Supply is hiring a junior software engineer to work closely with our lead engineer, designing and implementing feedback loops and revision mechanisms for our platform, Mission Control, and its AI‑powered workflows. This is a generalist software engineering role with an emphasis on product quality, manual and automated testing,...
-
Elite Research Scientist
Hace 1 semana
México Perle Jornada completaAbout Perle Perle powers the most ambitious AI initiatives in the world with human intelligence at scale. We work with the world’s leading model builders and enterprises to deliver expert-in-the-loop data, model evaluation, and trust systems that make AI safe, responsible, efficient and exceptionally high-performing. Our clients don’t buy software,...
-
Strategy Consultant
Hace 7 días
Mexico, chihuahua Mindrift Jornada completaDescriptionToloka AI supports frontier model post-training by building domain-specific reinforcement learning environments, tasks, and evaluation frameworks designed by real practitioners. Mindrift, powered by Toloka — a leading enterprise AI and machine learning data partner since 2014 — connects top domain experts with cutting-edge AI initiatives....
-
Freelance Agent Evaluation Engineer
Hace 2 semanas
Mexico Mindrift Trabajo remoto Jornada completa $40 IndefinidoDescriptionPlease submit your CV in English and indicate your level of English proficiency.Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.We're building a dataset to evaluate AI coding agents -...
-
AI Response Evaluation Analyst
Hace 2 semanas
México iMerit Jornada completaiMerit is seeking AI Response Evaluation Analyst to participate in a project where you will review AI generated responses and decide which ones actually hold up: accurate, relevant, concise, safe, well localized, and true to what the user asked for. Every ranking comes with a short written rationale, and those comments feed directly into how the models...
-
GenAI/ML Evaluation Engineer
Hace 2 semanas
México Centraprise Jornada completaJob Title: GenAI/ML Evaluation EngineerExperience: 3–8 YearsOpenings: 1–2Focus: GenAI/ML Evaluation, Data Pipelines, Test Infrastructure, NLPLocation: Mexico, RemoteRole OverviewWe are looking for an engineer to build an automated evaluation and regression framework for a PII rewriting/sanitization service. The role involves building Golden Datasets,...
-
Research Scientist
Hace 2 días
San Miguel Topilejo, México Anyone Ai Jornada completaResearch Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...
-
Lighting and Cinematography Consultant
Hace 17 horas
Mexico, chihuahua Weekday Ai Jornada completaDescription This role is for one of our clients Compensation: $60-$90 per hour Join a leading AI lab at the forefront of generative AI innovation and help shape the future of advanced visual intelligence systems. We are seeking experienced professionals in lighting, cinematography, visual effects, and digital content creation to contribute their expertise...
-
Research Scientist
Hace 3 días
Ciudad de México Anyone Ai Jornada completaResearch Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...
-
Research Scientist
Hace 3 días
San Antonio Tecómitl, México Anyone Ai Jornada completaResearch Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...
-
Software Engineer
Hace 15 horas
México Alignerr Trabajo remoto Jornada completaSoftware Engineer (C#) — Internal Tooling (AI Infrastructure) About The Role What if your C# expertise could directly shape the infrastructure powering the next generation of AI? We're looking for a senior full-stack C# engineer to build the data pipelines, annotation systems, and evaluation tooling that leading AI labs depend on every day. This isn't...
-
Remote Sensory
Hace 2 días
Tlalnepantla, Estado de México Mondelez España Galletas Production SLU Jornada completaMondelez México seeks a consumer science professional to drive sensory evaluation and consumer research, aiming to independently lead innovation projects and ensure products deliver the intended consumer experience across categories.The role trains to progressively own sensorial evaluation and projects, with a focus on collaboration with R&D, Marketing, and...
-
Senior AI Engineer — Enterprise AI Architect
Hace 17 horas
Veracruz, México BairesDev Jornada completa EUR 2 - EUR 3 Por obraBairesDev is hiring a Senior AI Engineer to influence the architecture of an enterprise AI assistant used across multiple functions. You will stay hands-on in code, shaping orchestration, retrieval strategies, and evaluation methods for an agentic AI platform. You’ll work with distributed teams, contribute to multi-agent design, and help implement robust...
-
Senior AI Engineer — Enterprise AI Architect
Hace 9 horas
Heroica Veracruz, Veracruz, México BairesDev Jornada completaBairesDev is hiring a Senior AI Engineer to influence the architecture of an enterprise AI assistant used across multiple functions. You will stay hands-on in code, shaping orchestration, retrieval strategies, and evaluation methods for an agentic AI platform.You’ll work with distributed teams, contribute to multi-agent design, and help implement robust...
-
Remote Civil Engineer for AI Training Tasks
Hace 2 días
mexico RemoteJobsOne Jornada completaRemoteJobsOne is seeking a Civil Engineer contractor to contribute expert evaluation tasks for a high-impact AI training project. You will design authentic civil engineering challenges reflecting real-world complexity and ensure they align with professional standards.Candidates should have 5+ years of civil engineering experience, PE licensure or active...
-
Research Scientist
Hace 2 días
Mexico City Anyone AI Trabajo remoto Jornada completaResearch Scientist, LLM Evaluations & Benchmarking Anyone AI LabsReports to: CEO · Remote / LatAm / USThe role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or measuring the wrong thing. You'll...
-
Cosmetic R&D Testing Engineer — Stability
Hace 4 horas
Tijuana, Baja California, México Markwins Beauty Brands Jornada completaMarkwins Beauty Brands in Mexico - Tijuana is seeking an R&D Testing Engineer to support the development, testing, evaluation and validation of cosmetic and personal care products. The role involves laboratory testing, material and product evaluation, and collaboration with R&D, Quality, Regulatory and Production teams to ensure quality, safety and...
-
mexico Alignerr Corp. Jornada completaAlignerr is seeking a Senior C# Infrastructure Engineer to design and build high-performance data pipelines and evaluation systems that support AI training. This fully remote, hourly contract role offers 20–40 hours per week and flexibility to work from anywhere.You will join a team delivering scalable data tooling and backend services for annotation,...
-
México Alignerr Corp. Jornada completaAlignerr is seeking a Senior C# Infrastructure Engineer to design and build high-performance data pipelines and evaluation systems that support AI training. This fully remote, hourly contract role offers 20–40 hours per week and flexibility to work from anywhere.You will join a team delivering scalable data tooling and backend services for annotation,...
-
Purchasing Agent
Hace 22 horas
mexico RemoteJobsOne Jornada completaThis is a fully remote position, open to candidates based in Mexico. Pay: $30–$65/hr Role Title: Purchasing Agent Role Type: Contractor Location: Remote We are engaging Purchasing Agents to contribute to a customer's project focused on advancing AI systems in the procurement domain. In this role, you'll apply your expertise to help train next-generation AI...