Sr. Data Engineer – Ontology & Semantic Modeling

SoundThinking

Fremont (CA)

Hybrid

USD 130,000 - 170,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading data and AI firm is seeking a Sr. Data Engineer to design scalable data pipelines and contribute to an ontology-driven semantic layer. The role requires experience with Databricks, SQL, and Postgres, along with proficiency in Python. Candidates should have a minimum of 5 years in data engineering and be comfortable working in AWS environments. This position is office-based in Fremont, CA, with a hybrid work model expected, requiring onsite work three days a week.

Qualifications

  • Experience working in production environments with Databricks and Spark.
  • Advanced SQL and strong data modeling skills required.
  • Experience building reliable ETL/ELT pipelines.

Responsibilities

  • Design and maintain ontologies, taxonomies, and semantic models.
  • Build and operate scalable data pipelines using Databricks.
  • Analyze existing Postgres schemas to identify performance bottlenecks.

Skills

Data Engineering
Databricks
SQL
Postgres
Python
AWS

Education

5+ years in Data Engineering or Data Architecture

Tools

Spark
Graph databases (Neo4j)

Job description

We’re building a modern data + AI platform where ontologies and semantic models create a consistent understanding of entities, relationships, and meaning across systems. We are seeking a Sr. Data Engineer (Ontology & Semantic Modeling) to design scalable data pipelines, contribute to an ontology-driven semantic layer, and help improve database schema and performance across our platform.

Ontology & Semantic Layer
  • Design, evolve, and maintain ontologies, taxonomies, and semantic models
  • Define entities and relationships; create mappings from source systems into semantic structures
  • Establish practical governance practices (versioning, documentation, naming standards)
  • Ensure semantic consistency across pipelines, APIs, and downstream applications
Data Engineering (Databricks)
  • Build and operate scalable pipelines using Databricks
  • Design reliable, well-structured datasets and transformation frameworks
  • Improve data quality through validation, deduplication, and monitoring
  • Optimize Spark and SQL workloads for performance and cost efficiency
Database Architecture & Performance (Postgres)
  • Analyze existing Postgres schemas and query patterns to identify performance bottlenecks
  • Improve table structures, indexing strategies, and data access patterns
  • Review query execution plans and optimize joins and filtering logic
  • Evaluate trade‑offs between normalized, dimensional, and denormalized models
  • Partner with engineering teams to resolve latency and scaling challenges
AI & Agent Enablement
  • Structure datasets and metadata to support LLM and agentic AI workflows
  • Support retrieval use cases (RAG / hybrid search) by preparing clean, linked, high‑signal data
  • Collaborate with AI teams to validate semantic consistency and correctness
Cloud & Collaboration
  • Work within AWS environments (S3, IAM, compute, managed databases)
  • Collaborate across data engineering, AI, product, and platform teams
  • Document architecture decisions, ontology standards, and best practices
Minimum Qualifications
  • 5+ years in Data Engineering or Data Architecture
  • Strong experience with Databricks + Spark in production environments
  • Advanced SQL and strong data modeling skills (conceptual/logical/physical)
  • Strong experience with Postgres, including schema design and performance tuning
  • Experience building reliable ETL/ELT pipelines
  • Proficiency in Python (or Scala) for data engineering workflows
  • Experience working in AWS environments
  • Experience implementing data quality practices (validation, deduplication, monitoring)
Nice to Have
  • Comfortable working with ontology / semantic modeling concepts
  • Knowledge graph concepts; exposure to RDF/OWL/SPARQL
  • Graph databases (Neo4j, etc.)
  • Experience supporting AI/ML or LLM‑based systems (RAG/hybrid retrieval)
  • Experience with orchestration tools (Airflow, Dagster, dbt)
  • Experience with governance/lineage tooling (e.g., Unity Catalog)

Travel: 15%

Location: Fremont, CA, Office

Hybrid Workplace

Employees are expected to work onsite for a minimum of 3 days per week, unless the advertised role has a specific on‑site requirement.

SoundThinking provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability or genetics. In addition to federal law requirements, SoundThinking complies with applicable state and local laws governing nondiscrimination in employment in every location in which the company has facilities. This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation and training.

SoundThinking expressly prohibits any form of workplace harassment based on race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability or veteran status.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

Samsara • United States

On-site
USD 120,000 - 150,000
Senior AI Data Engineer
Senior AI Data Engineer

Dematic • Wauwatosa (WI)

Hybrid
USD 134,250 - 179,000
Career Development
Competitive Compensation and Benefits
Pay Transparency
+1
Senior AI Data Engineer
Senior AI Data Engineer

Dematic • Atlanta (GA)

Hybrid
USD 134,250 - 179,000
Career Development
Competitive Compensation and Benefits
Pay Transparency
+1
Senior Data Modeler
Senior Data Modeler

Aptonet • United States

Remote
MXN 1,776,000 - 2,488,000
Lead Software Engineer
Lead Software Engineer

NICE • Seattle (WA)

On-site
USD 150,000 - 190,000
Data Engineer (in person)
Data Engineer (in person)

SEP • Westfield (IN)

On-site
USD 90,000 - 110,000
Flexible work schedules
Opportunities to learn and develop
Community of friendly peers
+1
Senior Data Engineer
Senior Data Engineer

Further • Cleveland (OH)

On-site
USD 100,000 - 130,000
Net-zero cost medical option
Company contributions to HSA
Fertility support
+2
Data Engineer (Founding Team)
Data Engineer (Founding Team)

Fabrion • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive salary
Early-stage equity
Senior Data Engineer
Senior Data Engineer

iFlow Inc. • Normal (IL)

On-site
USD 140,000 - 180,000
Software Engineer, Data Infrastructure
Software Engineer, Data Infrastructure

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 350,000 - 475,000
Health, dental, vision benefits
Unlimited PTO
Parental leave
+1