Farooq77Farooq77
HomeAboutServicesAgencyPortfolioTestimonialsJobsArticlesContact
F77

Farooq77

AI Automation Engineer & Agency Lead

Building intelligent automation systems with Python, AI agents, APIs, and workflow automation that help businesses save time, reduce costs, and scale efficiently.

Company

HomeAboutAgencyContactFAQ

Expertise

ServicesProcessTech StackPortfolio

Resources

JobsArticles

Connect

EmailLinkedInGitHub

Currently Available

Open for Projects & Vendor Partnerships

Start a Project →
Privacy PolicyTerms of ServiceCookie PolicyDisclaimer

© 2026 Farooq77. All rights reserved.

Designed & Developed by Muhammad Farooq

FULL_TIME•Remote•Active

Member of Technical Staff, Coding Research

micro1 is hiring a Member of Technical Staff, Coding Research to advance the evaluation and development of frontier coding agents. This full-time role sits at the intersection of AI research, software engineering, and model evaluation, with responsibility for designing benchmarks, methodologies, datasets, evaluation protocols, and supporting infrastructure used to measure and improve next-generation coding models. The role involves researching model behavior and failure modes, building large-scale experimentation and evaluation tooling, establishing rigorous assessment standards, and collaborating with researchers and engineers on emerging AI capabilities.

Salary

USD200,000 - USD260,000 / year

Department

Coding Research & AI Evaluation

Experience

3+ years of experience in software engineering, machine learning, AI research, model evaluation, or a related technical discipline, with a strong software engineering background and experience designing or validating technical assessments, benchmarks, coding tasks, or evaluation methodologies

Curated by Farooq77 Jobs

Role fit snapshot

Engagement
Full-time remote | micro1 application platform
Experience
Minimum recorded experience: 3+ years of experience in software engineering, machine learning, AI research, model evaluation, or a related technical discipline, with a strong software engineering background and experience designing or validating technical assessments, benchmarks, coding tasks, or evaluation methodologies
Core expertise
Coding Research & AI Evaluation | Large Language Models | Coding Agents | Coding Evaluation
Named tools
Python
Geographic eligibility
Not specified beyond Remote
Listing dates
Posted 2026-09-16 | Valid through 2026-10-16
Compensation
Recorded compensation: $200,000-$260,000/year

Responsibilities

  • Design and own evaluation frameworks for coding agents, including benchmark specifications, scoring methodologies, rubrics, and quality standards.
  • Lead end-to-end research initiatives focused on measuring and improving coding model performance across diverse software engineering tasks.
  • Develop high-quality datasets, golden examples, and evaluation protocols for reliable assessment of frontier coding systems.
  • Analyze coding model behavior and failure modes to identify systematic weaknesses and translate findings into actionable training and evaluation improvements.
  • Build tooling and infrastructure supporting large-scale experimentation, data generation, review workflows, and evaluation pipelines.
  • Establish best practices for coding-agent assessment with an emphasis on methodological rigor, reproducibility, and measurement quality.
  • Partner with researchers, engineers, and applied AI teams to design experiments and evaluate emerging model capabilities.
  • Automate technical workflows and improve evaluation processes through systematic experimentation.
  • Contribute to technical reports, benchmark studies, and client-facing research initiatives communicating model performance and research insights.

Skills

Large Language ModelsCoding AgentsCoding EvaluationAI EvaluationModel EvaluationMachine Learning SystemsSoftware EngineeringPythonC++Benchmark DesignEvaluation FrameworksScoring MethodologiesEvaluation RubricsDataset DevelopmentGolden ExamplesModel Behavior AnalysisFailure Mode AnalysisReinforcement LearningAgentic WorkflowsTool UsePost-TrainingExperimentationEvaluation PipelinesWorkflow AutomationData GenerationResearch MethodologyTechnical ResearchTechnical WritingOpen Source Development

Requirements

  • Strong software engineering background with expertise in Python, C++, or comparable programming languages.
  • 3+ years of experience in software engineering, machine learning, AI research, evaluation, or related technical disciplines.
  • Experience designing, reviewing, or validating technical assessments, benchmarks, coding tasks, or evaluation methodologies.
  • Familiarity with large language models, coding agents, reinforcement learning, model evaluation, or related AI systems.
  • Proven ability to build tooling, automate workflows, and improve technical processes through systematic experimentation.
  • Strong analytical skills with the ability to investigate model behavior and derive insights from complex technical systems.
  • Excellent written and verbal communication skills with the ability to clearly communicate technical findings to diverse audiences.
  • Ability to operate effectively in fast-moving research environments with significant ambiguity and evolving priorities.
  • Experience with frontier AI systems, coding agents, or model evaluation research is preferred.
  • Experience designing benchmarks or datasets for machine learning systems at scale is preferred.
  • Familiarity with agentic workflows, tool use, reinforcement learning, or post-training methodologies is preferred.
  • Publications, open-source contributions, or demonstrated technical leadership in AI, machine learning, or software engineering are preferred.

Benefits

  • Full-time remote position.
  • Base salary of $200,000–$260,000 per year.
  • Eligibility for equity compensation.
  • Potential performance-based bonuses subject to role and company policies.
  • Up to 100% reimbursement for health insurance premiums.
  • Paid time off.
  • 401(k) plan with company match.
  • Additional benefits supporting a high-performing remote-first workforce.
  • Opportunity to work on frontier coding agents and next-generation AI evaluation systems.
  • Opportunity to design benchmarks, datasets, and research methodologies that directly influence coding model development.

Before you apply

  • Confirm that your location is eligible for the role.
  • Confirm the employment or contract type.
  • Verify the current compensation at the official source.
  • Review the required skills and experience.
  • Check that the listing is still open before applying.
  • Never pay a fee to submit a job application.
This listing is curated by Farooq77 and may originate from a third-party employer or platform. The application link may include referral or tracking parameters. The official application source is the final authority for role details, eligibility, compensation, and availability.
Apply Now →

Related jobs

  • Behavioral Analyst

    CONTRACTOR

  • Software Engineer - Open Source Contributions

    CONTRACTOR