Staff Data Engineer

Pearson

Bengaluru

Hybrid

INR 1,800,000 - 4,000,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Pearson in Bengaluru is seeking a Staff Data Engineer to design, build, and operate analytics and automation capabilities with governed enterprise data. You will deliver Power BI reporting, establish Fabric/BigQuery environments, and develop Python-based automations and copilots.

In this hands-on role, you will own solution delivery, platform best practices, and governance, collaborating across teams to drive scalable data architecture and secure operations.

Qualifications

  • 5+ years in analytics, BI, or data engineering.
  • 3+ years hands-on Power BI development.
  • Strong experience with Microsoft Fabric (Lakehouse/Warehouse).
  • Proficient in DAX, SQL, and data modeling.
  • Hands-on Python delivering ML/GenAI solutions in production.
  • Knowledge of RAG concepts and data governance.

Responsibilities

  • Design, develop, and maintain reports and dashboards.
  • Develop semantic models with star schema.
  • Tune DAX for performance and usability.
  • Implement Power BI and Looker deployment pipelines across environments.
  • Define Dev/Test/Prod strategies and governance.
  • Lead collaboration and translate business requirements into scalable solutions.
  • Design agentic AI solutions powering analytics copilots and AI workflows.

Skills

Analytics experience
Power BI development
DAX proficiency
SQL proficiency
Python for data tasks
GenAI in production
Data governance
Environments & deployments

Tools

Microsoft Fabric
Looker
Power Automate
Power Apps
Dataverse
Azure OpenAI
GCP BigQuery
CI/CD for BI

Job description

Job Title: Staff Data Engineer
About the Role

We are seeking an Analytics Engineer to design, build, and operate our analytics and automations as well as build of AI-powered automations and copilots using governed enterprise data. This role is responsible for delivering high-quality Power BI reporting, establishing and maintaining Microsoft Fabric and/or GCP BigQuery, and building business automations and applications using Python, Power Automate and Power Apps.

You will be part of the Cloud & Service Management organization helping to evolve our self-service analytics, scalable data architecture, and automations—while ensuring security, performance, and governance across the platform.

This is a hands-on role with ownership of both solution delivery and platform best practices.

Key Responsibilities
Analytics & Reporting
  • Design, develop, and maintain reports and dashboards
  • Build and optimize semantic models using strong dimensional modeling (star schema)
  • Write and tune DAX measures with a focus on performance and usability
  • Implement Power BI and Looker deployment pipelines and promote content across environments
Microsoft Fabric Platform
  • Establish and maintain Microsoft Fabric architecture, including:
  • Lakehouse and/or Warehouse
  • Dataflows Gen2
  • OneLake data organization
  • Manage Fabric capacities, workspaces, and permissions
  • Monitor performance, cost, and reliability of Fabric workloads
  • Develop and maintain Python-based data transformations and notebooks within Fabric
  • Use Python for data preparation, enrichment, validation, and advanced analytics
  • Define and enforce data modeling and medallion architecture standards
Automation & Applications
  • Build and maintain automation flows for business processes, approvals, and integrations
  • Work with Dataverse, connectors, and security roles
  • Implement error handling, logging, and operational support patterns
Platform Governance & Operations
  • Define Dev/Test/Prod environment strategy for reporting and automation platform
  • Implement Application Life best practices (solutions, pipelines, source control where applicable)
  • Establish governance standards to prevent platform sprawl
  • Partner with security and IT teams on access control and compliance
  • Provide guidance and enablement to analysts and citizen developers
Collaboration & Leadership
  • Translate business requirements into scalable technical solutions
  • Contribute to platform roadmap and continuous improvement efforts
Agentic AI & ML Enablement
  • Design and deliver agentic AI solutions that automate multi-step business workflows (tool use, planning, and human-in-the-loop approvals) using enterprise data and governed actions.
  • Build RAG (retrieval-augmented generation) patterns over Fabric/OneLake (document ingestion, chunking, embeddings, retrieval evaluation) to power analytics copilots and self-service Q&A.
  • Develop and operate ML pipelines (feature engineering, training, evaluation, batch/real-time inference) using Python and approved ML frameworks.
  • Establish LLMOps/ModelOps practices: prompt/version control, offline evaluation, regression testing, monitoring (quality, drift, cost, latency), and safe rollback.
  • Implement AI security and governance : data access controls, prompt/data leakage prevention, PII handling, model risk reviews, and audit logging for agent actions.
  • Partner with stakeholders to identify high-value use cases and deliver measurable outcomes (time saved, defect reduction, SLA improvements).
Required Qualifications
  • 5+ years of experience in analytics, BI, or data engineering roles
  • 3+ years of hands-on Power BI development experience
  • Strong experience with Microsoft Fabric (Lakehouse, Warehouse, Dataflows)
  • Proficient in DAX, SQL, and data modeling
  • Hands-on experience with:
  • Power Automate (cloud flows, approvals, integrations)
  • Power Apps (Canvas apps)
  • Dataverse
  • Hands-on Python experience delivering ML or GenAI solutions in production (notebooks-to-service, APIs, scheduled jobs, or integrated automations).
  • Working knowledge of RAG concepts (embeddings, vector search, retrieval, grounding, evaluation).
  • Experience implementing monitoring and testing for data/ML/GenAI systems (data quality checks, model/prompt evaluation, logging/telemetry).
  • Experience managing environments, security, and deployments
  • Strong understanding of data governance and analytics best practices
Preferred Qualifications
  • Experience designing enterprise-scale analytics platforms
  • Familiarity with Azure services (Azure SQL, Data Factory, Synapse)
  • Familiarity with GCP BigQuery and Looker
  • Experience with CI/CD concepts for Power BI and Looker
  • Power Platform or Microsoft analytics certifications
  • Experience working in a Center of Excellence (CoE) model
  • Experience with Azure OpenAI / Azure AI Foundry (or equivalent) and enterprise deployment patterns.
  • Experience with orchestration frameworks (e.g., Semantic Kernel, LangChain, Autogen) and tool/function calling.
Who we are:

At Pearson, our purpose is simple: to help people realize the life they imagine through learning. We believe that every learning opportunity is a chance for a personal breakthrough. We are the world's lifelong learning company. For us, learning isn't just what we do. It's who we are. To learn more: We are Pearson.

Pearson is an Equal Opportunity Employer and a member of E-Verify. Employment decisions are based on qualifications, merit and business need. Qualified applicants will receive consideration for employment without regard to race, ethnicity, color, religion, sex, sexual orientation, gender identity, gender expression, age, national origin, protected veteran status, disability status or any other group protected by law. We actively seek qualified candidates who are protected veterans and individuals with disabilities as defined under VEVRAA and Section 503 of the Rehabilitation Act.

If you are an individual with a disability and are unable or limited in your ability to use or access our career site as a result of your disability, you may request reasonable accommodations by emailing TalentExperienceGlobalTeam@grp.pearson.com.

Job: Engineering

Job Family: TECHNOLOGY

Organization: OCTO

Schedule: FULL_TIME

Workplace Type: Hybrid

Req ID: 25935

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer - Analytics & Automation
Senior Data Engineer - Analytics & Automation

Pearson India Education Services Pvt Ltd • Bengaluru

On-site
INR 3,000,000 - 6,000,000
Staff Data Engineer
Staff Data Engineer

Jobtailor • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Staff Cloud Engineer
Staff Cloud Engineer

Pearson • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Data Engineer III
Data Engineer III

Pearson • Bengaluru

On-site
INR 1,000,000 - 1,700,000
Software Engineer III
Software Engineer III

Pearson • Chennai District

On-site
INR 1,500,000 - 2,100,000
Microsoft Fabric Data Analyst
Microsoft Fabric Data Analyst

Applicantz • Bengaluru

On-site
INR 2,626,000 - 4,379,000
Software Engineer III - AI & Agentic Automation
Software Engineer III - AI & Agentic Automation

Pearson • Bengaluru

Hybrid
INR 350,000 - 600,000
Microsoft Fabric Data Analyst
Microsoft Fabric Data Analyst

Applicantz • Bengaluru

On-site
INR 900,000 - 1,500,000
Senior MS Fabric & Power BI Engineer - SE
Senior MS Fabric & Power BI Engineer - SE

Endava • Bengaluru

On-site
INR 1,500,000 - 2,200,000
Data Engineer III
Data Engineer III

Pearson • Chennai District

Hybrid
INR 1,800,000 - 3,200,000