Remote CS Research Expert - Benchmark & ML Systems

24-Mag Llc

New York (NY)

Remote

USD 76,000 - 103,000

Part time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Remote work
Flexible schedule
Competitive hourly rate

Job summary

24-MAG LLC is seeking a part-time independent contractor with deep expertise in computer science to author and verify advanced AI benchmarking questions. The role is fully remote and entails developing rigorous problems across CS topics, from algorithms and ML to systems engineering and formal methods.

You will join an AI research initiative focused on high-quality evaluation materials, provide well-supported solutions, and help set evaluation standards.

Qualifications

  • PhD (or current) in CS, EE, CE, or related field.
  • Strong command of graduate CS theory, algorithms, systems design, and ML.
  • Excellent written English and ability to communicate complex concepts clearly.
  • Substantial software engineering or technical industry experience is valued.
  • Background in GPU kernel engineering, architectures, or formal methods is preferred.

Responsibilities

  • Author original multiple-choice questions testing deep technical understanding.
  • Verify questions for correctness, rigor, and solvability.
  • Develop content on GPU kernels, accelerators, and computer architecture.
  • Create rigorous problems in formal verification and automated reasoning.
  • Develop questions on distributed systems, cloud infra, OS, and databases.
  • Ensure solutions include clear reasoning and references.

Skills

PhD in CS
Graduate CS theory
Strong written English
GPU Kernel Engineering
Distributed Systems
Formal Methods
Data Engineering
Machine Learning
Cloud Infrastructure
Web/API Development
Operating Systems

Education

PhD in Computer Science

Job description

24-MAG LLC is seeking a part-time independent contractor with deep expertise in computer science to author and verify advanced AI benchmarking questions. The role is fully remote and entails developing rigorous problems across CS topics, from algorithms and ML to systems engineering and formal methods.

You will join an AI research initiative focused on high-quality evaluation materials, provide well-supported solutions, and help set evaluation standards.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Applied CS Benchmark Specialist - Remote Contract
Applied CS Benchmark Specialist - Remote Contract

Visa Hunt • United States

Remote
USD 91,000 - 116,000
Remote Part-Time Senior Engineering Assessment Consultant
Remote Part-Time Senior Engineering Assessment Consultant

24-Mag Llc • New York (NY)

Remote
USD 69,000 - 96,000
Remote AI Benchmark QA Engineer
Remote AI Benchmark QA Engineer

24-Mag Llc • New York (NY)

Remote
USD 76,000 - 117,000
Remote Applied Mathematician for AI Benchmarking
Remote Applied Mathematician for AI Benchmarking

24-Mag Llc • New York (NY)

Remote
USD 69,000 - 96,000
Remote work
Part-time contract
Flexible scheduling
Remote Medical Benchmarks Architect for AI Assessment
Remote Medical Benchmarks Architect for AI Assessment

24-Mag Llc • New York (NY)

Remote
USD 117,000 - 152,000
Remote Python Engineer for AI Benchmarking
Remote Python Engineer for AI Benchmarking

24-Mag Llc • New York (NY)

Remote
ML Research Engineer: End-to-End AI Evaluation & Experiments
ML Research Engineer: End-to-End AI Evaluation & Experiments

24-Mag Llc • New York (NY)

Remote
USD 100,000 - 155,000
Remote AI Benchmark Engineer & Researcher
Remote AI Benchmark Engineer & Researcher

Pathway • Palo Alto (CA)

Remote
USD 120,000 - 180,000
Remote AI Benchmark Engineer — PhD-Level Question Author
Remote AI Benchmark Engineer — PhD-Level Question Author

Obsidian • Detroit (MI)

On-site
USD 90,000 - 130,000
Remote Subject Matter Expert - AI Benchmarking for Pro Services
Remote Subject Matter Expert - AI Benchmarking for Pro Services

Lilt • United States

Remote