Benchmarking & Evaluation Researcher
TAIRC · Or HaNer
תיאור המשרה
About the role
The Benchmarking & Evaluation Researcher will help build standardized tests and benchmarks to assess AI models. This position works within the Core AI Research team to ensure model performance is measured accurately, fairly, and safely.
Key responsibilities
- Develop benchmark datasets that evaluate accuracy, fairness, robustness, and safety of AI models.
- Design evaluation protocols, including metrics, conditions, and processes for model assessment.
- Analyze evaluation results, interpret outcomes, and provide feedback for model improvement.
- Collaborate with external consortiums and organizations to create open benchmarks that promote transparency.
Required profile
- Master's or doctoral degree in statistics, data science, or a related field.
- Proven experience in dataset creation, statistical analysis, and evaluation framework design.
Required skills
- Dataset creation
- Statistical analysis
- Evaluation framework design
What we offer
- Opportunity to contribute to open AI benchmarking initiatives.
- Collaboration with external research consortiums and internal AI teams.
- Volunteer position with potential for future paid appointment pending funding availability.
Questions fréquentes
מדוע אתם מדווחים על ההצעה הזו?
Explore further
Salaries, guides and searches for ישראל.
Salaries by job title
הגש בקשה ב-30 שניות
הזינו את המייל שלכם כדי להגיש בקשה. חשבון יווצר אוטומטית.
בהמשך, אתם מסכימים לתנאי השימוש שלנו.
כבר יש לכם חשבון? התחברות
מתפרסם לפני חודש
תפוגה בעוד שבוע מעכשיו
40 צפיות · 0 interested
הגדל את סיכוייך
העלה את קורות החיים שלך: אנו מציעים לך מודעות תואמות לפרופיל שלך.
מנתח את קורות החיים שלך...
TAIRC
Or HaNer