benchmark
9,631 ofertas de empleo de benchmark en México. Encuentra ofertas actualizadas diariamente de los principales portales de empleo.
-
AI Benchmark Engineer | Native Language Specialist Premium
Hace 8 horas
Mexico (Remote), Mexico City LILT Trabajo remoto Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
metepec, estado de méxico Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
xochimilco, distrito federal, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
xico, veracruz, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
guadalajara, jalisco, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
monterrey, nuevo león, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
tijuana, baja california, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
naucalpan, estado de méxico Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
miguel hidalgo, distrito federal, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
mexicali, baja california, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
hermosillo, sonora, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
aguascalientes, aguascalientes, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
culiacán, sinaloa, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
azcapotzalco, distrito federal, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
guadalupe, nuevo león, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
santa catarina, nuevo león, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
morelia, michoacán, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
lerdo, durango, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
puerto vallarta, jalisco, México Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...
-
AI Benchmark Engineer | Native Language Specialist
Hace 22 horas
tlalnepantla, estado de méxico Lilt Inc. Jornada completaAbout The OpportunityWe are building a rigorous, verifiable evaluation suite of Terminal-Bench tasks designed to test the limits of large language models on multilingual software challenges. Our goal is to measure multilingual robustness across prompt language effects, non-English data processing, and complex locale/encoding edge cases in terminal...