Membre du personnel technique, Recherche en programmation
8 $US à 9 $US/hFourchette indicative communiquée par micro1
Publié le 14 septembre 2026 · Candidatures jusqu'au 13 novembre 2026
Je vous redirige vers la page officielle de micro1. La candidature est gratuite et se fait en anglais.
Je touche une commission de la plateforme si un candidat que j'ai orienté est recruté. Cela ne change rien pour vous.
Ce rôle à distance implique la conception de cadres d'évaluation et d'initiatives de recherche pour les agents de programmation de pointe. Les candidats doivent posséder de solides antécédents en génie logiciel et au moins trois années d'expérience en recherche en intelligence artificielle, en apprentissage automatique ou en génie logiciel.
Description en anglais, telle que publiée par micro1.
Job Title: Member of Technical Staff, Coding Research
Job Type: Full-time
Location: Remote
The Role
We are seeking a Member of Technical Staff to help advance the evaluation and development of frontier coding agents. Sitting at the intersection of AI research, software engineering, and model evaluation, you will design the benchmarks, methodologies, and data systems that shape how next-generation coding models are measured and improved.
What You'll Do
- Design and own evaluation frameworks for coding agents, including benchmark specifications, scoring methodologies, rubrics, and quality standards.
- Lead end-to-end research initiatives focused on measuring and improving coding model performance across diverse software engineering tasks.
- Develop high-quality datasets, golden examples, and evaluation protocols that enable reliable assessment of frontier coding systems.
- Analyze model behavior and failure modes, identifying systematic weaknesses and translating findings into actionable improvements for training and evaluation.
- Build tooling and infrastructure that support large-scale experimentation, data generation, review workflows, and evaluation pipelines.
- Establish best practices for coding-agent assessment, ensuring methodological rigor, reproducibility, and measurement quality.
- Partner closely with researchers, engineers, and applied AI teams to design experiments and evaluate emerging model capabilities.
- Contribute to technical reports, benchmark studies, and client-facing research initiatives that communicate model performance and insights.
What We're Looking For
- Strong software engineering background with expertise in Python, C++, or comparable programming languages.
- 3+ years of experience in software engineering, machine learning, AI research, evaluation, or related technical disciplines.
- Experience designing, reviewing, or validating technical assessments, benchmarks, coding tasks, or evaluation methodologies.
- Familiarity with large language models, coding agents, reinforcement learning, model evaluation, or related AI systems.
- Proven ability to build tooling, automate workflows, and improve technical processes through systematic experimentation.
- Strong analytical skills with the ability to investigate model behavior and derive insights from complex technical systems.
- Excellent written and verbal communication skills, including the ability to clearly articulate technical findings to diverse audiences.
- Comfortable operating in fast-moving research environments with significant ambiguity and evolving priorities.
Preferred
- Experience working on frontier AI systems, coding agents, or model evaluation research.
- Deep interest in understanding how data, evaluations, and feedback mechanisms influence model capabilities.
- Track record of independently driving ambiguous technical or research projects from conception to execution.
- Experience designing benchmarks or datasets for machine learning systems at scale.
- Familiarity with agentic workflows, tool use, reinforcement learning, or post-training methodologies.
- Publications, open-source contributions, or demonstrated technical leadership in AI, machine learning, or software engineering.
Compensation & Benefits Notice
The national pay range for this full-time position is base salary of $200,000, $260,000 USD. All employees are eligible for equity compensation, and employees may also receive performance-based bonuses, dependent on role and subject to company policies. micro1 provides a comprehensive benefits package, including up to 100% reimbursement for health-insurance premiums, paid time off, a 401(K) plan with a company match, and additional benefits designed to support a high-performing, remote-first workforce.
micro1 is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, sexual orientation, or gender identity), national origin, age, disability, genetic information, veteran status, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance and/or a reasonable accommodation during the application process, reach out to support@micro1.ai.
Our hiring process utilizes artificial intelligence tools to assist in candidate screening and assessment. Our AI tools are designed to complement, not replace, human decision-making.
Compétences recherchées
- LLMs
- Coding Evaluation
- AI Evaluation
- ML Systems
À propos de micro1
micro1 est un labo de données IA qui recrute des experts à distance pour entraîner et évaluer des modèles. L'offre originale est consultable sur leur site.
Voir l'offre sur micro1Autres postes en IA et machine learning
Ingenieur CQ - Gestion DataOS
Ce role de prestataire distant consiste a superviser les pipelines de controle qualite, a evaluer les resultats des projets techniques et a garantir des normes elevees pour les flux de donnees d'IA. Les candidats doivent posseder de solides bases en genie logiciel et une experience prealable en tant qu'expert en controle qualite ou chef de projet technique.
- QA
- Quality Management
Publié il y a 3 j
Entraineur LLM - Appel de fonction d'agent
L'entraineur LLM concoit des conversations a tours multiples et simule l'utilisation d'outils pour aider les entreprises d'IA fondamentale a entrainer et evaluer leurs modeles. Le poste exige de solides competences en raisonnement technique, une experience avec les API et les formats de donnees, ainsi que trois annees d'experience technique professionnelle.
- Python
- Java
- JavaScript
Publié il y a 4 j
Membre du personnel technique, IA d'entreprise
Ce poste s'intègre au sein des systèmes d'IA d'entreprise pour diagnostiquer des défaillances, concevoir des jeux de données d'évaluation et exécuter des expériences. Les candidats doivent détenir un master en informatique ou en apprentissage automatique et justifier d'une solide expérience dans la conception de cadres d'évaluation pour le machine learning.
- Research Signal Judgment
- ML-Oriented Data Design
- Ops-to-Research Translation
- +1
Publié il y a 4 j300 000 $US à 700 000 $US/an
Ingénieur déployé sur le terrain
L'ingénieur déployé sur le terrain collabore avec les principaux laboratoires d'IA et entreprises pour créer des pipelines de ML, des systèmes d'intelligence des données et des flux de travail agentiques. Les candidats doivent être de solides ingénieurs Python possédant une expérience professionnelle dans la création de systèmes de production et l'utilisation de grands modèles de langage.
- Python
- LLM Systems
- ML Infrastructure
- +1
Publié il y a 4 j300 000 $US à 650 000 $US/an