AIPayList

Jobs / AfterQuery

Mechanistic Interpretability (llms) Machine Learning Expert

up to $200/hr

source wording: “$150 - $200/hr

Platform
AfterQuery
Category
Coding & software eval
Eligibility
Worldwide
Freshness
posted 4mo ago · seen live 36 min ago

What this role asks for

  • 10–20 hrs/week
  • PhD
  • Master's
  • Machine learning
Read from the posting's own words: show the exact sentences
  • 10–20 hrs/week: “…on a project-by-project basis, with an expected commitment of 10–20 hours per week for the projects you accept. This position offers exceptional pay, exposure to…
  • PhD: “…venue (e.g., NeurIPS, ICML, ICLR, or equivalent) - Master's or PhD in Machine Learning, Artificial Intelligence, Computer Science, or a related…
  • Master's: “…a peer-reviewed venue (e.g., NeurIPS, ICML, ICLR, or equivalent) - Master's or PhD in Machine Learning, Artificial Intelligence, Computer Science, or a related…
  • Machine learning: “…is a remote, project-based role for machine learning researchers with deep expertise in mechanistic interpretability. You will complete tasks…

The role, as AfterQuery describes it

This is a remote, project-based role for machine learning researchers with deep expertise in mechanistic interpretability. You will complete tasks at the frontier of interpretability research — including analyzing internal model representations, reverse-engineering learned circuits, and developing tools and techniques to understand how neural networks compute. Work is over the next 2–3 weeks, asynchronous, and assigned on a project-by-project basis, with an expected commitment of 10–20 hours per week for the projects you accept. This position offers exceptional pay, exposure to cutting-edge AI safety and interpretability research, and a strong addition to your research portfolio.

Read the rest of the description (5 more paragraphs)

Responsibilities - Conduct mechanistic interpretability research on transformer-based and other neural network architectures - Identify, isolate, and analyze computational circuits responsible for specific model behaviors - Apply and extend techniques such as activation patching, probing, sparse autoencoders, and attention analysis - Develop tools and frameworks to automate or scale interpretability workflows across model families - Document methodologies, findings, and technical approaches clearly and reproducibly

Required qualifications - Published researcher with at least one first-author publication in a peer-reviewed venue (e.g., NeurIPS, ICML, ICLR, or equivalent) - Master's or PhD in Machine Learning, Artificial Intelligence, Computer Science, or a related quantitative field - Demonstrated expertise in mechanistic interpretability, model analysis, or AI safety research - Deep familiarity with transformer architectures and modern large language model internals - Strong problem-solving skills and ability to work independently on open-ended research tasks

Preferred qualifications - Hands-on experience with interpretability tools and libraries (e.g., TransformerLens, baukit, or similar) - Familiarity with sparse autoencoders, superposition, and feature geometry research - Background in TA'ing or teaching deep learning, NLP, or AI safety courses

Why apply - Flexible Time Commitment – Work on your schedule while tackling meaningful research challenges - Startup Exposure – Work directly with an early-stage Y Combinator-backed company, gaining hands-on experience that sets you apart - Exceptional Pay – Project-based pay ranges from $150–$200/hour - Portfolio Building – Gain experience on frontier interpretability and AI safety research problems - Professional Growth – Sharpen your skills on varied, challenging model analysis and reverse-engineering tasks

Employment type: Contract Time commitment: 10 hours/week Location: Remote

Posted by AfterQuery, reproduced here so you can judge the role before clicking. Original posting ↗

Source: platform job feed · first seen 22h ago · ID 1776315980481

More like this