Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Obsidian in New York seeks a senior engineer to advance a frontier engineering-reasoning evaluation project with an AI research lab. You will define simulations, thresholds, and prompts while guiding a model to design artifacts that meet strict specs.
You’ll work in a browser-based studio, collaborate via GitHub, and push the boundaries of model reasoning in engineering domains.
A frontier engineering-reasoning evaluation run in collaboration with a leading AI research lab. The work measures whether state-of-the‑art models can reason from first principles in your engineering domain rather than retrieve facts from training data — and you get visibility into the model's internal reasoning on your own tasks.
Domain experts write simulations and specification sheets; the model attempts to design an artifact — controller gains, circuit parameters, geometry — that satisfies every spec.
You define the simulation, a set of specs with pass/fail thresholds, and the prompt. The model probes your simulation with a limited number of calls, then submits a final design. An agentic grader runs the simulation and scores spec satisfaction.
Onboarding calls run daily, alongside internal tooling built to help you work faster. Prior model-evaluation experience is not required.