Member of Technical Staff, Coding Research

micro1
AI and machine learning
Expert
Remote
Full time
Core team

$8 to $9/hIndicative range provided by micro1

Posted on September 14, 2026 · Applications until November 13, 2026

Apply on micro1

I am sending you to the official micro1 page. Applying is free and in English.

I earn a commission from the platform when a candidate I referred gets hired. It changes nothing for you.

This remote role involves designing evaluation frameworks and research initiatives for frontier coding agents. Candidates must have a strong software engineering background and at least three years of experience in AI research, machine learning, or software engineering.

Description in English, as published by micro1.

Job Title: Member of Technical Staff, Coding Research

Job Type: Full-time

Location: Remote

The Role

We are seeking a Member of Technical Staff to help advance the evaluation and development of frontier coding agents. Sitting at the intersection of AI research, software engineering, and model evaluation, you will design the benchmarks, methodologies, and data systems that shape how next-generation coding models are measured and improved.

What You'll Do

  • Design and own evaluation frameworks for coding agents, including benchmark specifications, scoring methodologies, rubrics, and quality standards.
  • Lead end-to-end research initiatives focused on measuring and improving coding model performance across diverse software engineering tasks.
  • Develop high-quality datasets, golden examples, and evaluation protocols that enable reliable assessment of frontier coding systems.
  • Analyze model behavior and failure modes, identifying systematic weaknesses and translating findings into actionable improvements for training and evaluation.
  • Build tooling and infrastructure that support large-scale experimentation, data generation, review workflows, and evaluation pipelines.
  • Establish best practices for coding-agent assessment, ensuring methodological rigor, reproducibility, and measurement quality.
  • Partner closely with researchers, engineers, and applied AI teams to design experiments and evaluate emerging model capabilities.
  • Contribute to technical reports, benchmark studies, and client-facing research initiatives that communicate model performance and insights.

What We're Looking For

  • Strong software engineering background with expertise in Python, C++, or comparable programming languages.
  • 3+ years of experience in software engineering, machine learning, AI research, evaluation, or related technical disciplines.
  • Experience designing, reviewing, or validating technical assessments, benchmarks, coding tasks, or evaluation methodologies.
  • Familiarity with large language models, coding agents, reinforcement learning, model evaluation, or related AI systems.
  • Proven ability to build tooling, automate workflows, and improve technical processes through systematic experimentation.
  • Strong analytical skills with the ability to investigate model behavior and derive insights from complex technical systems.
  • Excellent written and verbal communication skills, including the ability to clearly articulate technical findings to diverse audiences.
  • Comfortable operating in fast-moving research environments with significant ambiguity and evolving priorities.

Preferred

  • Experience working on frontier AI systems, coding agents, or model evaluation research.
  • Deep interest in understanding how data, evaluations, and feedback mechanisms influence model capabilities.
  • Track record of independently driving ambiguous technical or research projects from conception to execution.
  • Experience designing benchmarks or datasets for machine learning systems at scale.
  • Familiarity with agentic workflows, tool use, reinforcement learning, or post-training methodologies.
  • Publications, open-source contributions, or demonstrated technical leadership in AI, machine learning, or software engineering.

Compensation & Benefits Notice

The national pay range for this full-time position is base salary of $200,000, $260,000 USD. All employees are eligible for equity compensation, and employees may also receive performance-based bonuses, dependent on role and subject to company policies. micro1 provides a comprehensive benefits package, including up to 100% reimbursement for health-insurance premiums, paid time off, a 401(K) plan with a company match, and additional benefits designed to support a high-performing, remote-first workforce.

micro1 is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, sexual orientation, or gender identity), national origin, age, disability, genetic information, veteran status, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance and/or a reasonable accommodation during the application process, reach out to support@micro1.ai.

Our hiring process utilizes artificial intelligence tools to assist in candidate screening and assessment. Our AI tools are designed to complement, not replace, human decision-making.

Required skills

  • LLMs
  • Coding Evaluation
  • AI Evaluation
  • ML Systems

About micro1

micro1 is an AI data lab that hires remote experts to train and evaluate models. The original posting is on their site.

View this job on micro1

More jobs in AI and machine learning

QC Engineer - DataOS Management

Turing
AI and machine learning
Remote
Full time

This remote contractor role involves overseeing quality control pipelines, evaluating technical project outputs, and ensuring high standards for AI data workflows. Candidates must possess strong software engineering fundamentals and previous experience as a quality control expert or technical project lead.

  • QA
  • Quality Management

Posted 3 days ago

LLM Trainer - Agent Function call

Turing
AI and machine learning
Remote
Full time

The LLM Trainer designs multi-turn conversations and simulates tool use to help foundational AI companies train and evaluate their models. The role requires strong technical reasoning skills, experience with APIs and data formats, and three years of professional technical experience.

  • Python
  • Java
  • JavaScript

Posted 4 days ago

Member of Technical Staff, Enterprise AI

micro1
AI and machine learning
Expert
Remote
Full time
Core team

This role embeds within enterprise AI systems to diagnose failures, design evaluation datasets, and run experiments. Candidates need a Master's degree in Computer Science or Machine Learning and strong experience in designing ML evaluation frameworks.

  • Research Signal Judgment
  • ML-Oriented Data Design
  • Ops-to-Research Translation
  • +1

Posted 4 days ago$300,000 to $700,000/yr

Forward Deployed Engineer

micro1
AI and machine learning
Remote
Full time
Core team

The Forward Deployed Engineer collaborates with leading AI labs and enterprises to build ML pipelines, data intelligence systems, and agentic workflows. Applicants must be strong Python engineers with professional experience building production systems and working with large language models.

  • Python
  • LLM Systems
  • ML Infrastructure
  • +1

Posted 4 days ago$300,000 to $650,000/yr