أخصائي العلوم والتكنولوجيا والهندسة والرياضيات والمعايير التقنية، تصميم المعايير

SME Careers
العلوم والبحث
عن بُعد
الولايات المتحدة فقط
تعاقد

‏100 US$/ساعةنطاق تقريبي تقدمه SME Careers

نُشرت في 29 سبتمبر 2026 · التقديم حتى 6 نوفمبر 2026

قدّم على SME Careers

أحوّلك إلى صفحة SME Careers الرسمية. التقديم مجاني وبالإنجليزية.

أحصل على عمولة من المنصة عند توظيف مرشح وجّهته. لا يغيّر ذلك شيئاً بالنسبة لك.

كيّف سيرتي الذاتية مع هذا العرض

يصمم أخصائي العلوم والتكنولوجيا والهندسة والرياضيات والمعايير التقنية معايير تقييم متقدمة، ويخلق أسئلة تقنية صعبة، ويقيم المحتوى المُولَّد بالذكاء الاصطناعي لتحسين النماذج الرائدة. يشترط حصول المرشحين على درجة البكالوريوس في تخصص تقني وامتلاك ثلاث سنوات على الأقل من الخبرة في المجال.

الوصف بالإنجليزية كما نشرته SME Careers.

In this hourly, remote contractor role, you will work as a STEM & Technical Subject Matter Expert (SME) to create challenging expert-level questions, evaluate AI-generated technical content, and design detailed rubrics that define what a correct, rigorous, and expert-quality answer should contain.

You will develop problems within your area of expertise that test advanced reasoning rather than simple factual recall. You will identify weaknesses in frontier AI models, including incorrect assumptions, incomplete reasoning, mathematical or scientific errors, missed edge cases, and answers that appear plausible but fail under expert scrutiny.

A major focus of this role is rubric design and writing. You will translate complex technical judgment into clear, specific, and measurable evaluation criteria, including required concepts, reasoning steps, acceptable alternative approaches, critical errors, and partial-credit considerations. You may also critique and improve rubrics developed by other experts.

This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world’s largest AI companies and foundation-model labs. Your technical expertise directly helps improve the world’s premier AI models by making their reasoning more accurate, rigorous, and reliable.

Responsibilities

  • Create expert-level technical questions: Develop difficult questions that test genuine technical reasoning and expose weaknesses in advanced AI systems.
  • Design evaluation rubrics: Write structured criteria defining what a correct, rigorous, and complete response must contain.
  • Define partial-credit standards: Identify essential reasoning steps, acceptable alternatives, minor errors, and critical failures.
  • Evaluate AI-generated solutions: Assess outputs for correctness, reasoning quality, completeness, methodology, and technical precision.
  • Identify hidden reasoning errors: Detect cases where an answer reaches the correct conclusion using invalid assumptions or flawed logic.
  • Critique peer rubrics: Review other experts’ criteria for technical accuracy, clarity, completeness, and scoring consistency.
  • Develop reference content: Produce expert answers, explanations, critiques, and other gold-standard content.
  • Support AI training: Create questions, rubrics, evaluations, preference judgments, and reference-answer pairs suitable for RL and SFT workflows.
  • Maintain technical rigor: Ensure content reflects appropriate professional or academic standards within the relevant discipline.

Requirements

  • Bachelor’s degree or higher in Mathematics, Physics, Chemistry, Engineering, Computer Science, Statistics, Applied Science, or another technical discipline.
  • Strong professional proficiency in English, minimum C1, with the ability to write precise technical explanations and evaluation criteria.
  • 3+ years of professional, academic, research, or industry experience in your stated technical domain.
  • Demonstrated advanced expertise in a clearly defined technical field or specialization.
  • Ability to create challenging technical questions that require multi-step reasoning, domain expertise, or professional judgment.
  • Strong ability to design detailed evaluation rubrics defining required reasoning, correct methodology, acceptable alternatives, partial-credit criteria, and critical errors.
  • Ability to distinguish between a correct final answer and an answer supported by valid versus flawed reasoning.
  • Comfortable reviewing and critiquing other experts’ rubrics for ambiguity, missing criteria, redundancy, technical inaccuracies, or poor scoring design.
  • High attention to detail when evaluating formulas, assumptions, units, methodology, edge cases, logical consistency, and technical terminology.
  • Experience with exam writing, academic grading, peer review, research review, technical QA, standards development, or assessment design is strongly preferred.
  • Prior experience with AI evaluation, RLHF, SFT, benchmarking, data annotation, prompt design, or LLM evaluation is preferred.
  • Reliable, self-directed, and able to deliver consistent quality in an hourly, remote contractor workflow.

المهارات المطلوبة

  • STEM
  • Mathematics
  • Physics
  • Chemistry
  • Engineering
  • Computer Science
  • Technical Reasoning
  • AI Evaluation
  • Rubric Design
  • rubric writing
  • Benchmarking
  • Model Evaluation
  • Question Writing
  • Expert Review
  • English
  • Assessment Design
  • Technical Writing
  • Academic Grading
  • Exam Item Writing
  • Peer Review
  • Research Review
  • Quality Assurance (QA)
  • Standards Development
  • Scoring Methodology
  • Partial Credit Scoring
  • Error Analysis
  • Critical Thinking
  • Multi-step Reasoning
  • Statistical Analysis
  • Applied Mathematics

فقط: الولايات المتحدة

عن SME Careers

SME Careers هي منصة الخبراء التابعة لـ SuperAnnotate، توظف مختصين عن بُعد لتدريب نماذج الذكاء الاصطناعي وتقييمها، من اللغات إلى القانون. الإعلان الأصلي متاح على موقعهم.

عرض الوظيفة على SME Careers

وظائف أخرى في العلوم والبحث

أخصائي نفسي ثنائي اللغة بالكانتونية ودكتوراه

micro1
العلوم والبحث
خبير
عن بُعد
تعاقد

تتضمن هذه الوظيفة عن بُعد تحليل دراسات الحالة النفسية، وتطوير سيناريوهات ثنائية اللغة بالكانتونية والإنجليزية، وتقييم المحتوى المُولَّد بواسطة الذكاء الاصطناعي للتأكد من دقته ومعاييره الأخلاقية. يجب أن يكون المرشح حاصلاً على درجة الدكتوراه في علم النفس وأن يمتلك طلاقة تامة أو شبه تامة في كل من اللغتين الكانتونية والإنجليزية.

  • bilingual communication
  • ethical decision-making
  • cultural sensitivity
  • +5

نُشرت منذ 5 يوماً‏100 US$ إلى ‏200 US$/ساعة

طبيب نفسي ثنائي اللغة بالكانتونية

micro1
العلوم والبحث
خبير
عن بُعد
تعاقد

تتضمن هذه الوظيفة عن بُعد تطوير وتقييم دراسات الحالة النفسية للمساعدة في تدريب نماذج الذكاء الاصطناعي. يجب أن يحمل المرشح شهادة طبية مع إقامة مكتملة في الطب النفسي، ورخصة طبية سارية المفعول، وطلاقة تامة في اللغة الكانتونية.

  • bilingual communication
  • ethical decision-making
  • cultural sensitivity
  • +5

نُشرت منذ 5 يوماً‏100 US$ إلى ‏200 US$/ساعة

خبير الصحة السلوكية، سلامة الذكاء الاصطناعي وتقييم النماذج

Mercor
العلوم والبحث
عن بُعد
الولايات المتحدة فقط
دوام جزئي

يقوم خبراء الصحة السلوكية بتقييم محادثات نماذج الذكاء الاصطناعي من حيث السلام الحيادي والحياد والحكم السليم في التفاعلات الحساسة للمستخدمين. يجب أن يحمل المرشحون شهادة في مجال ذي صلة وأن يمتلكوا خبرة مهنية لا تقل عن ثلاث سنوات في الصحة العقلية، أو الإرشاد النفسي، أو الخدمات الاجتماعية.

نُشرت منذ 5 يوماً‏45 US$ إلى ‏70 US$/ساعة

أخصائية نفسية ثنائية اللغة بالفيتنامية ودكتوراه

micro1
العلوم والبحث
خبير
عن بُعد
تعاقد

تتضمن هذه الوظيفة عن بُعد تحليل دراسات الحالة النفسية وتطوير مواد تدريبية حساسة ثقافياً لأنظمة الذكاء الاصطناعي. يجب أن تكون المرشحة حاصلة على درجة الدكتوراه في علم النفس وأن تمتلك طلاقة تامة أو شبه تامة في اللغتين الفيتنامية والإنجليزية.

  • bilingual communication
  • ethical decision-making
  • cultural sensitivity
  • +5

نُشرت منذ 6 يوماً‏100 US$ إلى ‏200 US$/ساعة