تفاصيل الوظيفة
تعلن شركة Tandem Search عن توفر وظيفة Principal Computer Vision & AI Engineer في الخبر، السعودية. هذا الدور قيادي تقني يتطلب خبرة عملية عميقة في رؤية الكمبيوتر وتحليل الفيديو، مع التركيز على بناء ونشر أنظمة AI في بيئات الإنتاج.
المهام والمسؤوليات
- تصميم وبناء أنظمة رؤية كمبيوتر وتحليل فيديو في بيئة الإنتاج.
- تطوير حلول تشمل اكتشاف الأجسام، التجزئة، التتبع، فهم الأحداث والاستدلال المكاني.
- تصميم ونشر أنظمة VLM/multimodal AI تعمل على فيديو مباشر ومسجل.
- الإشراف على دورة حياة التعلم الآلي: استراتيجية البيانات، التجارب، الضبط الدقيق، التقييم، النشر والمراقبة.
- تحسين النماذج للاستدلال على الحافة (Edge) والسحابة مع موازنة الدقة، زمن الاستجابة، الإنتاجية والتكلفة.
- إنشاء تقييم آلي للنماذج، اختبارات الانحدار ومراقبة الإنتاج.
- تطوير أنظمة AI وكيلية (Agentic) قادرة على التفكير فوق الفيديو والتفاعل مع الأدوات، APIs، الكاميرات وأجهزة الاستشعار.
- وضع المعايير التقنية والهندسية مع البقاء ملتزماً بالبرمجة العملية.
الشروط والمتطلبات
- خبرة عملية عميقة في رؤية الكمبيوتر / فيديو AI.
- خبرة إنتاجية قوية في العمل مع فيديو مباشر أو دفق فيديو.
- خبرة قوية في PyTorch و/أو TensorFlow.
- خبرة في اكتشاف الأجسام، التجزئة والتتبع باستخدام تقنيات مثل YOLO، DETR/RT-DETR، GroundingDINO أو ما يعادلها.
- خبرة إنتاجية مع نماذج VLM/multimodal مثل Qwen-VL، InternVL، LLaVA، Gemini أو ما يعادلها.
- خبرة في نشر وتحسين نماذج AI في الإنتاج.
- معرفة ب Edge AI، ويفضل TensorRT، ONNX Runtime، NVIDIA Jetson أو ما شابه.
- فهم قوي لتقييم النماذج، MLOps، التعلم النشط ومراقبة الإنتاج.
- القدرة على العمل كسلطة تقنية على مستوى Principal/Staff مع البقاء عملياً.
المهارات المطلوبة
- أنظمة Agentic AI / tool-calling مثل LangGraph، LangChain أو ما يعادلها.
- أنظمة كاميرات CCTV، بث فيديو فوري (real-time).
- خبرة في المجال الصناعي، الإنشاءات، الروبوتات، الأنظمة الذاتية أو AI ذات السلامة الحرجة.
- RAG / الاسترجاع متعدد الوسائط والبحث المتجه (vector search).
- خادم VLM/LLM مستضاف ذاتياً مثل vLLM أو NIM.
- متطلب حاسم: يجب أن يمتلك المرشح خبرة عملية واسعة في أنظمة رؤية كمبيوتر تعتمد على الفيديو. الملفات الشخصية التي تركز فقط على GenAI/LLM دون خبرة قوية في فيديو الإنتاج لن تكون مناسبة.
عرض النص الأصلي للإعلان
Principal Computer Vision & AI Engineer
Location: Al Khobar, Saudi Arabia
Role Overview
We are looking for a highly experienced Principal Computer Vision & AI Engineer to lead the technical direction of production AI systems operating over live video streams in real-world environments.
This is a hands-on technical leadership role rather than a people-management position. You will architect, build, optimise and deploy advanced computer vision, video understanding and multimodal/Vision-Language Model (VLM) systems while providing technical direction to senior engineers.
Key Responsibilities
- Architect and build production computer vision and video analytics systems.
- Develop solutions across object detection, segmentation, tracking, event understanding and spatial reasoning.
- Design and deploy VLM/multimodal AI systems operating over live and recorded video.
- Own the full ML lifecycle: data strategy, experimentation, fine-tuning, evaluation, deployment and monitoring.
- Optimise models for edge and cloud inference, balancing accuracy, latency, throughput and cost.
- Establish automated model evaluation, regression testing and production monitoring.
- Develop agentic AI systems capable of reasoning over video and interacting safely with tools, APIs, cameras and sensors.
- Set technical architecture and engineering standards while remaining highly hands-on with code.
What We're Looking For
- Deep hands-on experience in Computer Vision / Video AI.
- Strong production experience working with live or streaming video.
- Strong PyTorch and/or TensorFlow experience.
- Expertise across detection, segmentation and tracking using technologies such as YOLO, DETR/RT-DETR, GroundingDINO or equivalent.
- Production experience with Vision-Language Models / multimodal models such as Qwen-VL, InternVL, LLaVA, Gemini or equivalent.
- Experience deploying and optimising AI models in production.
- Exposure to edge AI, ideally TensorRT, ONNX Runtime, NVIDIA Jetson or similar.
- Strong understanding of model evaluation, MLOps, active learning and production monitoring.
- Ability to operate as a Principal/Staff-level technical authority while remaining hands-on.
Nice to Have
- Agentic AI / tool-calling systems such as LangGraph, LangChain or equivalent.
- Real-time CCTV, camera or video-streaming systems.
- Industrial, construction, robotics, autonomous systems or safety-critical AI experience.
- RAG / multimodal retrieval and vector search.
- Self-hosted VLM/LLM serving such as vLLM or NIM.
Critical requirement: Candidates must have substantial hands-on experience with video-based computer vision systems. Pure GenAI/LLM profiles without strong production video/CV experience will not be suitable.
رقم الإعلان لدى المصدر: 4456916167