عضو الطاقم الفني، هندسة الأبحاث

micro1
الذكاء الاصطناعي وتعلم الآلة
خبير
عن بُعد
دوام كامل
Core team

‏400.000 US$ إلى ‏800.000 US$/سنةنطاق تقريبي تقدمه micro1

نُشرت في 14 سبتمبر 2026 · التقديم حتى 13 نوفمبر 2026

قدّم على micro1

أحوّلك إلى صفحة micro1 الرسمية. التقديم مجاني وبالإنجليزية.

أحصل على عمولة من المنصة عند توظيف مرشح وجّهته. لا يغيّر ذلك شيئاً بالنسبة لك.

يقوم مهندس الأبحاث بتطوير بيئات التعلم التعزيزي، وخطوط التدريب، وأنظمة التقييم الآلي للارتقاء بنماذج الذكاء الاصطناعي الحديثة. يجب أن يمتلك المرشحون خبرة عميقة في التعلم التعزيزي وتصميم البيئات، إضافة إلى سجل قوي في توسيع نطاق أطر التجريب.

الوصف بالإنجليزية كما نشرته micro1.

Job Title: Member of Technical Staff, Research Engineering

Job Type: Full-time

Location: Remote

The Role

We are seeking a Research Engineer to operate at the frontier of Reinforcement Learning (RL), developing novel environments, training pipelines, and evaluation systems that advance the capabilities of modern AI models. This role sits at the intersection of research and production, translating experimental ideas into scalable, high-performance systems.

What You’ll Work On

  • Architect self-contained RL environments that capture complex, real-world tasks, including reward functions, verifiers, and evaluation logic.
  • Design and scale episode pipelines and multi-component training processes (MCPs) to support reproducible experimentation.
  • Build automated data generation systems, leveraging synthetic data to accelerate training cycles without compromising quality.
  • Develop and integrate AI-driven evaluation and quality assurance systems for automated grading, validation, and feedback loops.
  • Fine-tune and optimize open-source RL models using internally generated datasets and custom training strategies.
  • Establish benchmarking frameworks to measure model capability, robustness, and data quality across tasks.
  • Contribute to the release and analysis of evaluations on internal and external benchmark platforms (e.g., micro1 benchmarks).

What We're Looking For

  • Deep experience in Reinforcement Learning, including environment design and training dynamics.
  • Strong track record of building and scaling RL systems, pipelines, or experimentation frameworks.
  • Proficient in automation and data generation, including synthetic data pipelines.
  • Familiar with automated evaluation systems, model validation, and quality assurance workflows.
  • Experienced in fine-tuning and evaluating open-source ML models.
  • Clear, concise communicator with strong technical writing skills.
  • Comfortable operating in fast-paced, research-driven, and highly collaborative environments.

Preferred

  • Experience publishing benchmarks, evaluations, or research artifacts.
  • Familiarity with evaluation ecosystems (e.g., micro1 benchmarks or similar frameworks).
  • Background in scalable infrastructure for large-scale RL experimentation.

Compensation & Benefits Notice

The national pay range for this full-time position is base salary of $200,000, $300,000 USD. All employees are eligible for equity compensation, and employees may also receive performance-based bonuses, dependent on role and subject to company policies. micro1 provides a comprehensive benefits package, including up to 100% reimbursement for health-insurance premiums, paid time off, a 401(K) plan with a company match, and additional benefits designed to support a high-performing, remote-first workforce.

micro1 is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, sexual orientation, or gender identity), national origin, age, disability, genetic information, veteran status, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance and/or a reasonable accommodation during the application process, reach out to support@micro1.ai.

Our hiring process utilizes artificial intelligence tools to assist in candidate screening and assessment. Our AI tools are designed to complement, not replace, human decision-making.

Disclaimer

The information contained in this job posting, including but not limited to role responsibilities, qualifications, compensation, and benefits, is provided for informational purposes only and does not constitute a binding offer of employment. micro1 reserves the right to amend, modify, or withdraw any portion of this posting at its sole discretion and without prior notice. All employment decisions are made in accordance with applicable laws and regulations.

المهارات المطلوبة

  • Reinforcement Learning
  • ML-Oriented Data Design
  • RL Environments
  • RL Workflows

عن micro1

micro1 مختبر بيانات ذكاء اصطناعي يوظف خبراء عن بُعد لتدريب النماذج وتقييمها. الإعلان الأصلي متاح على موقعهم.

عرض الوظيفة على micro1

وظائف أخرى في الذكاء الاصطناعي وتعلم الآلة

مهندس مراقبة الجودة، إدارة DataOS

Turing
الذكاء الاصطناعي وتعلم الآلة
عن بُعد
دوام كامل

تتضمن هذه الوظيفة كمقاول عن بعد الإشراف على خطوط أنابيب مراقبة الجودة، وتقييم مخرجات المشاريع التقنية، وضمان معايير عالية لتدفقات بيانات الذكاء الاصطناعي. يجب أن يمتلك المرشحون أساسيات هندسة برمجيات قوية وخبرة سابقة كخبير مراقبة جودة أو قائد مشروع تقني.

  • QA
  • Quality Management

نُشرت منذ 3 يوماً

مدرب LLM لتنفيذ وظائف الوكيل

Turing
الذكاء الاصطناعي وتعلم الآلة
عن بُعد
دوام كامل

يقوم مدرب LLM بتصميم المحادثات متعددة الجولات ومحاكاة استخدام الأدوات لمساعدة الشركات التأسيسية للذكاء الاصطناعي في تدريب نماذجها وتقييمها. تتطلب الوظيفة مهارات استدلال تقني قوية، وخبرة في واجهات برمجة التطبيقات وتنسيقات البيانات، وثلاث سنوات من الخبرة التقنية المهنية.

  • Python
  • Java
  • JavaScript

نُشرت منذ 4 يوماً

عضو الطاقم التقني للذكاء الاصطناعي للمؤسسات

micro1
الذكاء الاصطناعي وتعلم الآلة
خبير
عن بُعد
دوام كامل
Core team

يندمج هذا الدور داخل أنظمة الذكاء الاصطناعي للمؤسسات لتشخيص الإخفاقات، وتصميم مجموعات بيانات التيقييم، وتشغيل التجارب. يحتاج المرشحون إلى درجة الماجستير في علوم الكمبيوتر أو التعلم الآلي وخبرة قوية في تصميم أطر تقييم التعلم الآلي.

  • Research Signal Judgment
  • ML-Oriented Data Design
  • Ops-to-Research Translation
  • +1

نُشرت منذ 4 يوماً‏300.000 US$ إلى ‏700.000 US$/سنة

مهندس منشَر ميدانياً

micro1
الذكاء الاصطناعي وتعلم الآلة
عن بُعد
دوام كامل
Core team

يتعاون المهندس المنشَر ميدانياً مع مختبرات الذكاء الاصطناعي الرائدة والشركات لبناء خطوط أنابيب التعلم الآلي، وأنظمة ذكاء البيانات، وسير العمل الوكيلي. يجب أن يكون المتقدمون مهندسي بايثون أقوياء ذوي خبرة مهنية في بناء أنظمة الإنتاج والعمل مع نماذج اللغة الكبيرة.

  • Python
  • LLM Systems
  • ML Infrastructure
  • +1

نُشرت منذ 4 يوماً‏300.000 US$ إلى ‏650.000 US$/سنة