مهندس برمجيات أول، تقييم الذكاء الاصطناعي / وكلاء البرمجة

Turing
هندسة البرمجيات
عن بُعد
تعاقد

نُشرت في 9 سبتمبر 2026 · التقديم حتى 18 أكتوبر 2026

قدّم على Turing

أحوّلك إلى صفحة Turing الرسمية. التقديم مجاني وبالإنجليزية.

أحصل على عمولة من المنصة عند توظيف مرشح وجّهته. لا يغيّر ذلك شيئاً بالنسبة لك.

يقوم مهندس البرمجيات الأول بتقييم وتحسين نماذج برمجة الذكاء الاصطناعي من خلال مراجعة الكود المُنتَج، وتحديد أنماط الفشل، وإنشاء مؤشرات تقييم. يحتاج المرشحون إلى خمس سنوات على الأقل من الخبرة العملية في هندسة البرمجيات، وكفاءة عالية في لغة إنتاجية رئيسية، ومهارات قوية في مراجعة الكود.

الوصف بالإنجليزية كما نشرته Turing.

Freelance · Remote · North America, LATAM, or India

About Turing

Turing is one of the world’s leading AGI infrastructure companies, working with frontier AI labs to accelerate model development through high-quality training data, evaluations, and engineering talent.

About the Role

We’re looking for experienced, hands-on software engineers to help evaluate and improve AI coding models.

Rather than primarily building production applications, you’ll work with coding agents across real-world repositories and assess the quality of their work. You’ll review generated code and agent behavior, determine whether solutions are technically correct, identify failure modes, and create the evaluation signals and feedback used to improve model performance.

Think of the coding agent as another engineer whose work you’re reviewing: Can it understand the task? Did it choose the right approach? Is the resulting code correct, robust, and maintainable? Can you explain precisely where it succeeded or failed?

What You’ll Do

  • Evaluate AI-generated code and solutions across real-world software repositories

  • Review agent behavior, tool usage, and code changes for correctness and quality

  • Identify technical errors, weak approaches, and recurring model failure modes

  • Compare model outputs and explain why one solution is better than another

  • Create and refine rubrics and evaluation criteria for coding tasks

  • Produce high-quality evaluation and preference data used to improve coding models

  • Build and maintain pipelines and infrastructure supporting data generation, collection, and evaluation workflows

  • Synthesize findings from data work into clear write-ups, updates, and recommendations for the team

  • Collaborate closely with researchers and engineers to translate qualitative judgment into scalable processes

  • Share clear, actionable findings with AI researchers and engineers

What We’re Looking For

  • 5+ years of hands-on software engineering experience

  • Strong proficiency in Python, TypeScript/JavaScript, Go, or another major production language

  • Experience working in substantial real-world codebases

  • Strong code-review skills and technical judgment

  • Ability to clearly explain why an implementation is correct, incorrect, or could be improved

  • Strong written communication

  • Experience using modern LLMs or AI coding tools

Experience with LLM evaluation, coding agents, RLHF, preference data, rubric design, or post-training is a plus, but not required.

Engagement Details

  • Compensation: Market rate; please provide a specific hourly rate expectation

  • Availability: 40 hours/week preferred, with at least 6 hours of Pacific Time overlap

  • Type: Independent contractor

  • Duration: Approximately 3 months

  • Start: As soon as possible

  • Location: North America, LATAM, or India

Evaluation Process

  • AI interview (~25 minutes)

  • Practical code/AI evaluation exercise (~30 minutes)

  • Hiring manager interview (~20 minutes)

The practical exercise focuses on your ability to review and evaluate AI-generated code, not competitive programming or algorithm puzzles.

المهارات المطلوبة

  • Python
  • Typescript
  • JavaScript
  • Go
  • Java
  • GitHub
  • CI/CD
  • Open Source
  • LLM
  • Code Reviews
  • Code Analysis

عن Turing

توظف Turing خبراء عن بُعد لمشاريع الذكاء الاصطناعي والهندسة، من البرمجة إلى الطب. الإعلان الأصلي متاح على موقعهم.

عرض الوظيفة على Turing

وظائف أخرى في هندسة البرمجيات

مدير هندسة بايثون، تدريب وتقييم نماذج اللغة الكبيرة

Turing
هندسة البرمجيات
عن بُعد
دوام كامل

تتضمن هذه الوظيفة المستقلة عن بعد قيادة فرق كبيرة من مهندسي بايثون وعلماء البيانات لدعم تدريب نماذج اللغة الكبيرة الأساسية وسير عمل التقييم. يجب أن تمتلك خبرة خمس سنوات أو أكثر في هندسة البرمجيات مع كفاءة قوية في بايثون ومهارات إدارية مثبتة.

  • Python
  • Software Development

نُشرت أمس

خبير المجال، الهندسة

Turing
هندسة البرمجيات
عن بُعد
دوام كامل

تتضمن هذه الوظيفة عن بعد بناء ميزات الواجهة الأمامية القابلة للتطوير باستخدام React وJavaScript. يحتاج المرشحون إلى ثلاث سنوات على الأقل من الخبرة كمهندس واجهة أمامية وشهادة في علوم الكمبيوتر أو ما يعادلها.

  • Data Engineering

نُشرت منذ 2 يوماً

مساهم في المصادر المفتوحة (جيت هاب)

micro1
هندسة البرمجيات
خبير
عن بُعد
تعاقد

يقوم المساهم في المصادر المفتوحة باستكشاف وإصلاح وتعديل مكونات الواجهة الأمامية والخلفية لمشروع تدريب الذكاء الاصطناعي. يحتاج المتقدمون إلى ملف تعريف قابل للتحقق على جيت هاب يتضمن مساهمات ذات مغزى في المصادر المفتوحة وخبرة قوية في هندسة البرمجيات عبر لغات برمجة متعددة.

  • Python3
  • JAVA
  • Rust
  • +3

نُشرت منذ 2 يوماً‏100 US$ إلى ‏150 US$/ساعة

مهندس برمجيات أول للمنظومة المتكاملة

micro1
هندسة البرمجيات
خبير
عن بُعد
تعاقد

يقوم مهندس برمجيات المنظومة المتكاملة الأول ببناء وإصلاح وتقییم مكونات التطبيقات الأمامية والخلفية لمشروع تدريب ذكاء اصطناعي. يتطلب الدور خبرة مهنية قوية في تطوير تطبيقات المنظومة المتكاملة الإنتاجية وخبرة عملية عبر لغات برمجة متعددة.

  • Python
  • React
  • JavaScript
  • +5

نُشرت منذ 2 يوماً‏50 US$ إلى ‏100 US$/ساعة