Engineering
Member of Technical Staff, Coding Research
micro1
Full-Time
Senior
$200k – $260k/yr
Remote
Posted 2w ago
Skills & tools
Python
Job Description
Core team
$600K \- $1\.3M/yr compensation
**Required Skills**
-------------------
LLMs
Coding Evaluation
AI Evaluation
ML Systems
### **About micro1**
micro1 is the leading AI data lab for training frontier models and evaluating AI agents. Experts contribute their diverse subject matter knowledge across domains such as finance, healthcare, STEM engineering, and more. micro1 transforms that real\-world expertise into high\-quality training data, evaluations, and feedback loops that improve how AI systems learn, reason, and perform.
Our platform identifies and vets top talent through an AI recruiter, enabling high\-quality expert contributions at scale. We aim to enable 1 billion people to do meaningful work by applying their expertise to AI. As our global expert network grows, micro1 is building the human intelligence layer for frontier AI.
**Job Title:** Member of Technical Staff, Coding Research
**Job Type:** Full\-time
**Location:** Remote
**The Role**
We are seeking a Member of Technical Staff to help advance the evaluation and development of frontier coding agents. Sitting at the intersection of AI research, software engineering, and model evaluation, you will design the benchmarks, methodologies, and data systems that shape how next\-generation coding models are measured and improved.
**What You'll Do**
* Design and own evaluation frameworks for coding agents, including benchmark specifications, scoring methodologies, rubrics, and quality standards.
* Lead end\-to\-end research initiatives focused on measuring and improving coding model performance across diverse software engineering tasks.
* Develop high\-quality datasets, golden examples, and evaluation protocols that enable reliable assessment of frontier coding systems.
* Analyze model behavior and failure modes, identifying systematic weaknesses and translating findings into actionable improvements for training and evaluation.
* Build tooling and infrastructure that support large\-scale experimentation, data generation, review workflows, and evaluation pipelines.
* Establish best practices for coding\-agent assessment, ensuring methodological rigor, reproducibility, and measurement quality.
* Partner closely with researchers, engineers, and applied AI teams to design experiments and evaluate emerging model capabilities.
* Contribute to technical reports, benchmark studies, and client\-facing research initiatives that communicate model performance and insights.
**What We're Looking For**
* Strong software engineering background with expertise in Python, C\+\+, or comparable programming languages.
* 3\+ years of experience in software engineering, machine learning, AI research, evaluation, or related technical disciplines.
* Experience designing, reviewing, or validating technical assessments, benchmarks, coding tasks, or evaluation methodologies.
* Familiarity with large language models, coding agents, reinforcement learning, model evaluation, or related AI systems.
* Proven ability to build tooling, automate workflows, and improve technical processes through systematic experimentation.
* Strong analytical skills with the ability to investigate model behavior and derive insights from complex technical systems.
* Excellent written and verbal communication skills, including the ability to clearly articulate technical findings to diverse audiences.
* Comfortable operating in fast\-moving research environments with significant ambiguity and evolving priorities.
**Preferred**
* Experience working on frontier AI systems, coding agents, or model evaluation research.
* Deep interest in understanding how data, evaluations, and feedback mechanisms influence model capabilities.
* Track record of independently driving ambiguous technical or research projects from conception to execution.
* Experience designing benchmarks or datasets for machine learning systems at scale.
* Familiarity with agentic workflows, tool use, reinforcement learning, or post\-training methodologies.
* Publications, open\-source contributions, or demonstrated technical leadership in AI, machine learning, or software engineering.
**Compensation \& Benefits Notice**
The national pay range for this full\-time position is base salary of $200,000 –$260,000 USD. All employees are eligible for equity compensation, and employees may also receive performance\-based bonuses, dependent on role and subject to company policies. micro1 provides a comprehensive benefits package, including up to 100% reimbursement for health\-insurance premiums, paid time off, a 401(K) plan with a company match, and additional benefits designed to support a high\-performing, remote\-first workforce.
micro1 is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, sexual orientation, or gender identity), national origin, age, disability, genetic information, veteran status, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance and/or a reasonable accommodation during the application process, reach out to support@micro1\.ai.
Our hiring process utilizes artificial intelligence tools to assist in candidate screening and assessment. Our AI tools are designed to complement, not replace, human decision\-making.
**Apply now**
-------------
Get jobs like this in your inbox
Join thousands of digital nomads getting the best remote jobs delivered weekly. Free, no spam.
Similar Jobs
Senior Platform Engineer, GitLab Orbit
GitLab IncGitLab IncFull-Time · RemoteDirect
VueTypeScriptRust
$139k – $235k2w agoView details$139k – $235k2w ago
Senior Software Engineer I, Full Stack
Wpromote, LLCWpromote, LLCFull-Time · RemoteDirect
ReactDjangoPython
$135k – $155k2w agoView details$135k – $155k2w ago
Test Automation Developer
LynxLynxFull-Time · RemoteDirect
PythonDockerCI/CD
$65k – $70k2w agoView details$65k – $70k2w ago
Engineering Director
College BoardCollege BoardFull-Time · RemoteDirect
ReactNode.jsJavaScript
$140k – $175k2w agoView details$140k – $175k2w ago
Forensic Engineer
YA GroupYA GroupFull-Time · RemoteDirect
$80k – $275k2w ago
Director of Product Management, Agentic Software Delivery
GitLab IncGitLab IncFull-Time · RemoteDirect
$203k – $346k2w ago
Remote Water/Wastewater Engineer
AtwellAtwellFull-Time · RemoteDirect
$100k – $116k2w ago
Senior Forensic Engineer - Mechanical
YA GroupYA GroupFull-Time · RemoteDirect
$140k – $275k2w ago
TL
IT Security Systems Administrator
TopDog LawTopDog LawFull-Time · RemoteDirect
PythonAzure
$115k – $135k2w agoView details$115k – $135k2w ago
E
Forward Deployed Engineer
EVBEVBFull-Time · RemoteDirect
Python
$180k – $250k2w agoView details$180k – $250k2w ago