Research Scientist, Computer Vision

Attentive.ai

San Jose (CA)

On-site

USD 170,000 - 275,000

Full time

29 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Attentive.ai in the United States is seeking a Research Scientist to advance computer vision, deep learning, and NLP that power automated takeoff and estimation. You will join the early US research team and own research problems end to end, taking models from concept to production.

You will build models for object detection, semantic segmentation, and structured extraction on construction drawings, and design evaluation frameworks while staying current with ML research relevant to field services.

Qualifications

  • 2–8 years of applied ML or AI research with a strong focus on computer vision and deep learning.
  • Depth in at least one of: object detection, image segmentation, image processing, OCR and layout analysis, self-supervised or representation learning, or multimodal / vision-language models
  • Experience training and evaluating neural networks at scale, and at least one model you've taken to production
  • Strong Python and PyTorch, across the modeling, training loop and the data pipeline
  • Practical GPU and distributed training experience
  • Judgment about data: what to label, what to synthesize, what to discard
  • The ability to read current research critically and say what's worth trying

Responsibilities

  • Own research problems end to end: framing, data, experiments, ablations, and the model that ships
  • Advance deep learning models for object detection, semantic segmentation, and structured extraction on vector and raster construction drawings
  • Build multimodal systems combining visual layout, geometry, and text
  • Design evaluation frameworks to continuously improve the model
  • Stay current with computer vision and ML research and evaluate what's relevant to construction-industry problems

Skills

Python
PyTorch
Computer Vision
Deep Learning
NLP

Education

Master's or PhD in CS/ML/EE

Tools

SVG
DXF
DWG

Job description

About Attentive.ai

Attentive.ai builds AI for construction and field services. Our takeoff and estimating platform, Beam AI, is used by 1,200+ contractors across the US and Canada and has completed over 500,000 takeoffs. Attentive.ai is transforming the field services and construction industries with our flagship AI solution - Beam AI. Our platform empowers businesses to double their bidding capacity and accelerate growth through automation and intelligent insights. More than

About Attentive.ai

Attentive.ai builds AI for construction and field services. Our takeoff and estimating platform, Beam AI, is used by 1,200+ contractors across the US and Canada and has completed over 500,000 takeoffs. Attentive.ai is transforming the field services and construction industries with our flagship AI solution - Beam AI. Our platform empowers businesses to double their bidding capacity and accelerate growth through automation and intelligent insights. More than 1k+ businesses across the U.S. and Canada already use our products to boost sales velocity and streamline operations. We have raised $30.5M in Series B funding, accelerating our mission to make advanced AI tools accessible, practical, and impactful in the real world. We are proudly Backed by Insight Partners, Peak XV (Surge), InfoEdge, Tenacity and Vertex Ventures.

About The Role

As a Research Scientist, you'll advance computer vision, deep learning, and NLP that power automated takeoff and estimation. The core problem is unsolved: a plan set is hundreds of pages drafted to dozens of CAD conventions, where the same primitive is a wall, a hatch pattern, or a leader line depending on context a model has to infer.

You're an early member of the US research team, which means real influence over the directions we pursue. You'll own research problems end to end and take models to production.

What You'll Do
  • Own research problems end to end: framing, data, experiments, ablations, and the model that ships
  • Advance deep learning models for object detection, semantic segmentation, and structured extraction on vector and raster construction drawings
  • Build multimodal systems combining visual layout, geometry, and text
  • Design evaluation frameworks to continuously improve the model
  • Stay current with computer vision and ML research and evaluate what's relevant to construction-industry problems
What We're Looking For
  • 2-8 years of applied machine learning or AI research experience with a strong focus on computer vision and deep learning. We'll calibrate level and compensation to what you've built.
  • Depth in at least one of: object detection, image segmentation, image processing, OCR and layout analysis, self-supervised or representation learning, or multimodal / vision-language models
  • Experience training and evaluating neural networks at scale, and at least one model you've taken to production
  • Strong Python and PyTorch, across the modeling, training loop and the data pipeline
  • Practical GPU and distributed training experience
  • Judgment about data: what to label, what to synthesize, what to discard
  • The ability to read current research critically and say what's worth trying
Nice To Have
  • Master's or PhD in Computer Science, Machine Learning, Electrical Engineering, Mathematics, or a related field
  • Familiarity with vector graphics and CAD formats (SVG, DXF, DWG, IFC)
  • LLM and NLP experience like structured extraction, retrieval over long documents
  • Model optimization: quantization, pruning, knowledge distillation
  • Publications or open-source contributions in computer vision or deep learning
  • Experience with cloud environments (GCP, AWS, or Azure)
How We Work
  • Empirical over theoretical - we'd rather run the experiment than argue about it
  • Truth-seeking, including about our own results. A negative result reported early is worth more than a positive one defended late
  • Research is judged by whether it holds up in front of 1,200 contractors, not by whether it's clever
  • Honest feedback, given directly
Compensation And Sponsorship

We are hiring Research Scientists across multiple levels, with opportunities for candidates ranging from approximately 2-8 years of relevant experience; level and compensation will be aligned with experience, research depth, and impact. The base salary range for this role is $170,000 - $275,000, plus equity. The final offer depends on experience and level. We sponsor visas. If you need sponsorship now or in the future, please apply.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist, Computer Vision
Research Scientist, Computer Vision

Beam AI • San Jose (CA)

On-site
USD 170,000 - 275,000
Equity
Research Intern, Computer Vision
Research Intern, Computer Vision

Attentive.ai • San Jose (CA)

On-site
USD 110,000 - 124,000
Research Intern, Computer Vision
Research Intern, Computer Vision

Beam AI • San Jose (CA)

On-site
USD 110,000 - 124,000
Research Scientist, Computer Vision — AI for Construction
Research Scientist, Computer Vision — AI for Construction

Attentive.ai • San Jose (CA)

On-site
USD 170,000 - 275,000
Construction AI: Research Scientist, Computer Vision
Construction AI: Research Scientist, Computer Vision

Beam AI • San Jose (CA)

On-site
USD 170,000 - 275,000
Equity
Computer Vision Research Engineer
Computer Vision Research Engineer

Bobyard, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Research Intern - Computer Vision for Construction AI
Research Intern - Computer Vision for Construction AI

Beam AI • San Jose (CA)

On-site
USD 110,000 - 124,000
Research Engineer FullTime
Research Engineer FullTime

Higharc Inc. • United States

Remote
USD 90,000 - 130,000
Unlimited PTO
Comprehensive medical, dental, and vision coverage
401K plan
Product Manager
Product Manager

Bobyard, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 170,000
Research Engineer - 3D Vision and Generation, Self-Driving
Research Engineer - 3D Vision and Generation, Self-Driving

Applied Intuition Inc. • Sunnyvale (CA)

On-site
USD 126,000 - 423,000