📍 المملكة العربية السعودية تحديث مستمر على مدار الساعة

وظيفة باحث ذكاء اصطناعي - بيانات متعددة اللغات لدى Jobgether في السعودية

AI Researcher - Multilingual Data
🏢 Jobgether
🕒 نُشرت: (اليوم) 📍 السعودية وظائف الهندسة والتقنية
التقديم على الوظيفة من المصدر الرسمي ↗

تفاصيل الوظيفة

انضم إلى بيئة بحثية متطورة حيث يلتقي ابتكار الذكاء الاصطناعي متعدد اللغات بأثر واقعي. تبحث شركتنا الشريكة عن باحث في الذكاء الاصطناعي - بيانات متعددة اللغات (AI Researcher - Multilingual Data) في السعودية.

المهام والمسؤوليات

  • تصميم وإجراء أبحاث تركز على مجموعات البيانات متعددة اللغات، بما في ذلك جمع البيانات وتصفيتها وإزالة التكرار وتقييم الجودة وتحسينها.
  • تطوير استراتيجيات مبتكرة للغات منخفضة الموارد واللغات طويلة الذيل عبر تقنيات العينات المتقدمة، وزيادة البيانات، وتعلم المناهج الدراسية.
  • البحث وتحسين نماذج اللغات الكبيرة متعددة اللغات من خلال تعزيز النقل عبر اللغات، والمحاذاة، والمتانة، وتعلم التمثيل.
  • بناء وصيانة وتحسين معايير التقييم متعددة اللغات لقياس جودة النموذج وأدائه عبر اللغات.
  • التعاون الوثيق مع مهندسي التعلم الآلي والباحثين للتأثير على خطوط التدريب، وهياكل النماذج، واستراتيجيات النشر في الإنتاج.
  • نشر نتائج الأبحاث في مؤتمرات AI و NLP الرائدة مع المساهمة في مبادرات المصادر المفتوحة عند الاقتضاء.
  • ترجمة نتائج البحوث إلى تحسينات عملية تعزز أنظمة الذكاء الاصطناعي الجاهزة للإنتاج.

الشروط والمتطلبات

  • خلفية متقدمة في معالجة اللغة الطبيعية (NLP)، أو التعلم الآلي (Machine Learning)، أو الذكاء الاصطناعي (Artificial Intelligence)، أو مجال وثيق الصلة.
  • خبرة بحثية مثبتة في النمذجة اللغوية متعددة اللغات أو عبر اللغات مع منشورات في مؤتمرات أو مجلات مرموقة مثل ACL، EMNLP، NeurIPS، ICML، أو ICLR.
  • خبرة عملية في العمل مع مجموعات بيانات نصية واسعة النطاق متعددة اللغات وسير عمل التعلم الآلي الحديث.
  • فهم قوي لترميز النصوص متعددة اللغات (multilingual tokenization)، وتصميم المفردات، والتعلم بالنقل، وتعلم التمثيل متعدد اللغات، وتقييم جودة مجموعة البيانات، وتقنيات التصفية، وتخفيف التحيز.
  • إتقان لغة بايثون (Python) وأطر التعلم العميق الحديثة مثل PyTorch أو JAX.
  • القدرة على العمل بشكل مستقل، وإدارة المبادرات البحثية، وتقديم نتائج عالية الجودة في بيئة شركة ناشئة سريعة الحركة.
  • تعتبر الخبرة في اللغات منخفضة الموارد، والكتابات غير اللاتينية، ومعايير التقييم متعددة اللغات (XTREME، FLORES، TyDi QA)، ومشاريع NLP مفتوحة المصدر، أو تدريب نماذج اللغات الكبيرة ميزة قوية.

المزايا

  • حزمة تعويضات تنافسية.
  • فرصة أسهم ذات معنى ضمن شركة في مرحلة مبكرة وعالية النمو.
  • ملكية كبيرة على اتجاه البحث واتخاذ القرارات التقنية.
  • فرصة لموازنة البحث الأكاديمي مع التأثير الإنتاجي في العالم الحقيقي.
  • الوصول إلى مجموعات بيانات واسعة النطاق متعددة اللغات، وبنية تحتية حديثة للذكاء الاصطناعي، ودورات تجريبية سريعة.
  • بيئة تعاونية تقدر الابتكار والتميز البحثي والتعلم المستمر.
  • فرصة للنشر في مؤتمرات AI و NLP الدولية الرائدة مع المساهمة في مبادرات مفتوحة المصدر ذات تأثير.
عرض النص الأصلي للإعلان
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Researcher - Multilingual Data based in Saudi Arabia.

Join a cutting-edge research environment where multilingual AI innovation meets real-world impact. In this role, you will drive the development of high-quality multilingual datasets and research strategies that power next-generation language models across diverse languages and domains. Working at the intersection of research and engineering, you will transform scientific discoveries into scalable production solutions while contributing to state-of-the-art advancements in natural language processing. This position offers the opportunity to publish influential research, collaborate with highly skilled experts, and shape the future of multilingual AI in a fast-paced, innovation-driven environment. If you are passionate about solving complex language challenges and advancing machine learning research, this is an opportunity to make a meaningful global impact.

Accountabilities

  • Design and conduct research focused on multilingual datasets, including data collection, filtering, deduplication, quality assessment, and optimization.
  • Develop innovative strategies for low-resource and long-tail languages through advanced sampling, data augmentation, and curriculum learning techniques.
  • Research and improve multilingual large language models by enhancing cross-lingual transfer, alignment, robustness, and representation learning.
  • Build, maintain, and refine multilingual evaluation benchmarks to measure model quality and performance across languages.
  • Collaborate closely with machine learning engineers and researchers to influence training pipelines, model architectures, and production deployment strategies.
  • Publish research findings at leading AI and NLP conferences while contributing to open-source initiatives when appropriate.
  • Translate research outcomes into practical improvements that enhance production-ready AI systems.

Requirements

  • Advanced background in Natural Language Processing, Machine Learning, Artificial Intelligence, or a closely related field.
  • Proven research experience in multilingual or cross-lingual language modeling with publications at recognized conferences or journals such as ACL, EMNLP, NeurIPS, ICML, or ICLR.
  • Hands-on experience working with large-scale multilingual text datasets and modern machine learning workflows.
  • Strong understanding of multilingual tokenization, vocabulary design, transfer learning, multilingual representation learning, dataset quality assessment, filtering techniques, and bias mitigation.
  • Proficiency in Python and modern deep learning frameworks such as PyTorch or JAX.
  • Ability to work independently, manage research initiatives, and deliver high-quality results in a fast-moving startup environment.
  • Experience with low-resource languages, non-Latin scripts, multilingual evaluation benchmarks (XTREME, FLORES, TyDi QA), open-source NLP projects, or large language model training is considered a strong advantage.

Benefits

  • Competitive compensation package.
  • Meaningful equity opportunity within an early-stage, high-growth company.
  • Significant ownership over research direction and technical decision-making.
  • Opportunity to balance academic research with real-world production impact.
  • Access to large-scale multilingual datasets, modern AI infrastructure, and rapid experimentation cycles.
  • Collaborative environment that values innovation, research excellence, and continuous learning.
  • Opportunity to publish at leading international AI and NLP conferences while contributing to impactful open-source initiatives.

How Jobgether Works

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

المصدر: LinkedIn - أُضيفت للموقع في 27 يوليو 2026

وظائف أخرى لدى Jobgether