Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Thinking Machines in San Francisco, CA, is hiring an engineer to architect and scale core data infrastructure for distributed training and multimodal data catalogs. You’ll work with researchers to accelerate experiments and build from the ground up using Spark, Kafka, Beam, Ray, and Delta Lake.
Responsibilities include data ingestion, quality checks, deduplication, and search, plus end-to-end lifecycle traceability and robust monitoring to improve platform reliability.
Thinking Machines in San Francisco, CA, is hiring an engineer to architect and scale core data infrastructure for distributed training and multimodal data catalogs. You’ll work with researchers to accelerate experiments and build from the ground up using Spark, Kafka, Beam, Ray, and Delta Lake.
Responsibilities include data ingestion, quality checks, deduplication, and search, plus end-to-end lifecycle traceability and robust monitoring to improve platform reliability.