evaluation

1,000 ofertas de empleo de evaluation en México. Encuentra ofertas actualizadas diariamente de los principales portales de empleo.


  • distrito federal, distrito federal, México Prophetic Technologies Inc Jornada completa

    Prophetic Software is not able to sponsor employment visas now or in the future. Candidates must be authorized to work in the United States without current or future sponsorship to be considered for this role.About Prophetic:Real estate development is a multi-billion-dollar industry that has run on fragmented data, manual processes, and gut instinct for...


  • Nuevo León, Nuevo León, México Metalsa Jornada completa

    About the CompanyWe are a global company with 65+ years of experience in the automotive industry. We manufacture safe and sustainable products for people around the world. We are working for a better future where we enrich communities every day by being committed to people, innovation and our planet. If you have what it takes to accelerate Metalsa's vision...


  • Estado de Nuevo León, Estado de Nuevo León, México Metalsa Jornada completa

    About the Company We are a global company with 65+ years of experience in the automotive industry. We manufacture safe and sustainable products for people around the world. We are working for a better future where we enrich communities every day by being committed to people, innovation and our planet. If you have what it takes to accelerate Metalsa's...


  • tlalnepantla de baz, estado de méxico Mercor Jornada completa

    About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey. Position: Frontend Engineer Type: Contract Compensation: $90/hour Location: Remote Role Responsibilities - Render a...


  • México, México iMerit Jornada completa

    The work iMerit is looking for detail oriented analysts to evaluate and rank AI generated responses to image based prompts. You will judge answers on accuracy, relevance, clarity, conciseness, safety, localization, and how well they follow the user's instructions, then explain your reasoning in writing. Much of the job comes down to this: look at the...


  • Ciudad de México, Ciudad de México Straive Jornada completa $600,000 - $900,000 Por obra

    Straive is a global leader in enterprise-grade data analytics and AI solutions, actively seeking a L2 Team Lead for the Search Evaluation/Content Moderation project. You will ensure overall quality and operational success, master complex guidelines, and bridge policy owners with the onsite execution team.As a leader, you will train and transfer knowledge,...

  • Team Lead

    Hace 1 semana


    Ciudad de México, Ciudad de México Straive Jornada completa $600,000 - $900,000 Por obra

    Straive is a global leader in enterprise-grade data analytics and AI solutions, committed to empowering businesses across various industries with cutting-edge technology and expert insights. Backed by EQT, a top private equity firm, we are uniquely positioned to drive innovation through significant investments and an entrepreneurial spirit. Our core focus is...


  • santiago de querétaro, querétaro, México Rws Trainai Jornada completa

    We are looking for Speech AI Evaluation Specialist to support the improvement of AI-generated content in Spanish (Mexico). Job Type: Freelance Location: Mexico (work from home) Work Schedule: Part-time - 10+ hours per week. Flexible - work whenever you want. Start Date: Immediately Duration: TBC Rate: 7 USD per hour Help Shape the Future of AI Are you a...


  • Ciudad de México, Ciudad de México Prophetic Technologies Inc Jornada completa $2 - $3 Por obra

    Prophetic Software is seeking a full-time ML evaluation engineer to design and own ground-truth datasets for our AI stack. You will define correctness, create label schemas, and ensure data quality across modules.You will work with product and engineering to source data, calibrate automated judges, and report metrics to accelerate product iterations. Strong...


  • Ciudad de México, Ciudad de México Prophetic Technologies Inc Jornada completa $2 - $3 Por obra

    Prophetic Software is not able to sponsor employment visas now or in the future. Candidates must be authorized to work in the United States without current or future sponsorship to be considered for this role.About Prophetic:Real estate development is a multi-billion-dollar industry that has run on fragmented data, manual processes, and gut instinct for...

  • Generative Audio Evaluation Premium

    Hace 2 semanas


    Ciudad de México RWS Trabajo remoto Jornada completa USD 8 Indefinido

    We are looking for Generative Audio Evaluation Specialists! A great entry point into ongoing work within one of our most active AI markets! Job Type: Freelance Location: Remote (Mexico) Work Schedule: Part-time - 10+ hours per week. Flexible - work whenever you want! Rate: 8 USD/hour Help Shape the Future of AI We’re currently hiring Generative Audio...


  • City, CMX, México RWS Trabajo remoto Jornada completa

    We are looking for **Generative Audio Evaluation** **Specialists** ! A great entry point into ongoing work within one of our most active AI markets! **Job Type:** Freelance **Location:** Remote (Mexico) **Work Schedule:** Part-time - 10+ hours per week. Flexible - work whenever you want! **Rate:** 8 USD/hour **Help Shape the...

  • Remote AI Evaluation

    Hace 2 semanas


    Mexico, chihuahua AI Chopping Block Jornada completa

    Mindrift connects specialists with project-based AI opportunities for leading tech companies, focusing on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.We are building a dataset to evaluate AI coding agents and design tasks from intermediate states of simulated environments. The agent writes most of...


  • , México iMerit Technology Jornada completa

    The work iMerit is looking for detail oriented analysts to evaluate and rank AI generated responses to image based prompts. You will judge answers on accuracy, relevance, clarity, conciseness, safety, localization, and how well they follow the user's instructions, then explain your reasoning in writing. Much of the job comes down to this: look at the...


  • , México iMerit Technology Jornada completa

    iMerit is seeking AI Response Evaluation Analyst to participate in a project where you will review AI generated responses and decide which ones actually hold up: accurate, relevant, concise, safe, well localized, and true to what the user asked for. Every ranking comes with a short written rationale, and those comments feed directly into how the models...


  • México, México iMerit Jornada completa

    iMerit is seeking AI Response Evaluation Analyst to participate in a project where you will review AI generated responses and decide which ones actually hold up: accurate, relevant, concise, safe, well localized, and true to what the user asked for. Every ranking comes with a short written rationale, and those comments feed directly into how the models...

  • Research Scientist

    Hace 2 días


    Ciudad de México Anyone Ai Jornada completa

    Research Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...

  • Research Scientist

    Hace 2 días


    mérida, yucatán, México Anyone Ai Jornada completa

    Research Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...

  • Research Scientist

    Hace 2 días


    celaya, guanajuato, México Anyone Ai Jornada completa

    Research Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...

  • Research Scientist

    Hace 2 días


    iztacalco, distrito federal, México Anyone Ai Jornada completa

    Research Scientist, LLM Evaluations & Benchmarking Anyone AI Labs — Human Data Division Reports to: CEO · Remote / LatAm / US The role Evaluation is one of the hardest open problems in AI: we still don't have reliable ways to measure what frontier models can and can't do, and the field mostly runs on benchmarks that are saturated, contaminated, or...