Data Engineer - PySpark Developer

Barclays

Bengaluru

On-site

INR 1,500,000 - 2,400,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Barclays in Bengaluru seeks a Data Engineer - PySpark Developer to evolve our digital landscape. You will design and deliver ETL pipelines, data products, and foundational data assets in collaboration with engineers and stakeholders.

The role emphasizes cloud-based ETL, AWS services, Databricks, security, and scalable architectures to meet business needs while adhering to Barclays standards.

Qualifications

  • Experience in ETL / data processing / data warehousing / data engineering / RDBMS / PySpark.
  • Experience in Data Processing and ETL in On-prem and Cloud based systems.
  • AWS cloud development with EC2, S3, Lambda, RDS, DynamoDB; Python.
  • IaC tools like AWS CloudFormation, Terraform.
  • CI/CD tools and pipelines (Jenkins, GitLab CI) for automation and deployment.
  • Cloud security best practices including IAM, VPC, security groups, encryption.
  • Monitoring/logging/troubleshooting with CloudWatch/CloudTrail.
  • Version control with Git, GitHub, Bitbucket.
  • Strong communication and collaboration in a team environment.
  • Translate business needs into technical solutions; align to solution architecture.
  • Scale to the BUK CDAO tech tooling strategy e.g. Databricks.

Responsibilities

  • Build and maintain data pipelines, data warehouses, and data lakes.
  • Design scalable data architectures with appropriate security measures.
  • Develop processing and analysis algorithms for complex data.
  • Collaborate with data scientists to deploy ML models.

Skills

ETL development
PySpark
AWS cloud
IaC tools
CI/CD pipelines
Cloud security
Monitoring & logging
Git version control
Communication

Tools

Terraform
AWS CloudFormation
Jenkins
GitLab CI
GitHub
Databricks

Job description

Join us as a Data Engineer - PySpark Developer at Barclays, where you'll spearhead the evolution of our digital landscape, driving innovation and excellence. This role is responsible for delivering the technology stack, using strong analytical and problem-solving skills to understand the business requirements and deliver quality solutions coming up with the ETL process design / data pipeline of the Foundational / Consolidated and Business data Products for the Tesco Tenancy on EDP. The resource will be working on complex technical problems that will involve detailed analytical skills and analysis. This will be done in conjunction with fellow engineers, business analysts and business stakeholders.

To be a successful Data Engineer - PySpark Developer, you should have experience with:

  • Experience in ETL / data processing / data warehousing / data engineering / RDBMS / PySpark.
  • Experience in Data Processing and ETL in On-prem and Cloud based systems.
  • Preferred AWS cloud development, with a strong understanding of AWS services (e.g., EC2, S3, Lambda, RDS, DynamoDB, etc.). Proficient in at least one Python programming language.
  • Hands-on experience with Infrastructure as Code (IaC) tools like AWS CloudFormation, Terraform etc.
  • Experience with CI/CD tools and pipelines (Jenkins, GitLab CI, etc.) for automation and deployment.
  • Good understanding of cloud security best practices and experience with IAM, VPC, security groups, and encryption mechanisms.
  • Good knowledge of monitoring, logging, and troubleshooting tools (e.g. CloudWatch, CloudTrail).
  • Experience with version control systems (Git, GitHub, Bitbucket).
  • Strong communication skills and ability to work in a collaborative team environment.
  • The role works with business stakeholders to translate needs into technical solutions and align to solution architecture.
  • The role should be able to scale up and align to the BUK CDAO tech tooling strategy e.g. EDP capabilities, Databricks etc.
  • It incorporates security principles into designs, deliverables.
  • Ensures compliance with standards (CSO, CDO, CDA etc.), executes and tests end-to-end Data Pipeline.
  • Meets non-functional requirements like scalability, security, and resilience.
  • Data solutions must follow 'data product principles' and the CDA team's best practices for data modelling, compliance, and security.
Some Other Highly Valued Skills May Include
  • Experience with Cloud based services and integrations, Exposure to Databricks, Agile development practices and Familiarity with databases (SQL/NoSQL), Release management

You may be assessed on key critical skills relevant for success in role, such as risk and controls, change and transformation, business acumen, strategic thinking and digital and technology, as well as job-specific technical skills.

This role is based in Bengaluru.

Purpose of the role

To build and maintain the systems that collect, store, process, and analyse data, such as data pipelines, data warehouses and data lakes to ensure that all data is accurate, accessible, and secure.

Accountabilities
  • Build and maintenance of data architectures pipelines that enable the transfer and processing of durable, complete and consistent data.
  • Design and implementation of data warehoused and data lakes that manage the appropriate data volumes and velocity and adhere to the required security measures.
  • Development of processing and analysis algorithms fit for the intended data complexity and volumes.
  • Collaboration with data scientist to build and deploy machine learning models.
Assistant Vice President Expectations
  • To advise and influence decision making, contribute to policy development and take responsibility for operational effectiveness. Collaborate closely with other functions/ business divisions.
  • Lead a team performing complex tasks, using well developed professional knowledge and skills to deliver on work that impacts the whole business function. Set objectives and coach employees in pursuit of those objectives, appraisal of performance relative to objectives and determination of reward outcomes
  • If the position has leadership responsibilities, People Leaders are expected to demonstrate a clear set of leadership behaviours to create an environment for colleagues to thrive and deliver to a consistently excellent standard. The four LEAD behaviours are: L – Listen and be authentic, E – Energise and inspire, A – Align across the enterprise, D – Develop others.
  • OR for an individual contributor, they will lead collaborative assignments and guide team members through structured assignments, identify the need for the inclusion of other areas of specialisation to complete assignments. They will identify new directions for assignments and/ or projects, identifying a combination of cross functional methodologies or practices to meet required outcomes.
  • Consult on complex issues; providing advice to People Leaders to support the resolution of escalated issues.
  • Identify ways to mitigate risk and developing new policies/procedures in support of the control and governance agenda.
  • Take ownership for managing risk and strengthening controls in relation to the work done.
  • Perform work that is closely related to that of other areas, which requires understanding of how areas coordinate and contribute to the achievement of the objectives of the organisation sub-function.
  • Collaborate with other areas of work, for business aligned support areas to keep up to speed with business activity and the business strategy.
  • Engage in complex analysis of data from multiple sources of information, internal and external sources such as procedures and practises (in other areas, teams, companies, etc).to solve problems creatively and effectively.
  • Communicate complex information. 'Complex' information could include sensitive information or information that is difficult to communicate because of its content or its audience.
  • Influence or convince stakeholder to achieve outcomes.

All colleagues will be expected to demonstrate the Barclays Values of Respect, Integrity, Service, Excellence and Stewardship – our moral compass, helping us do what we believe is right. They will also be expected to demonstrate the Barclays Mindset – to Empower, Challenge and Drive – the operating manual for how we behave.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineer - PySpark/SQL
Engineer - PySpark/SQL

Barclays • Bengaluru

On-site
INR 1,500,000 - 3,500,000
Data Engineer - PySpark
Data Engineer - PySpark

Barclays • Maharashtra

On-site
INR 1,000,000 - 1,500,000
Data Engineer
Data Engineer

Barclays • Pune City

On-site
INR 800,000 - 1,500,000
Data Engineer - PySpark Developer
Data Engineer - PySpark Developer

1203 Barclays Global Serv. Cent • Bengaluru

On-site
INR 900,000 - 1,500,000
Senior Data Engineer
Senior Data Engineer

Barclays • Maharashtra

On-site
INR 3,500,000 - 6,500,000
Data Engineer
Data Engineer

Hackajob • Bengaluru, Pune District

On-site
INR 1,000,000 - 1,500,000
Cloud Data Engineer
Cloud Data Engineer

Barclays • Pune District

On-site
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

Barclays • Maharashtra

On-site
INR 900,000 - 1,900,000
Engineer - PySpark/SQL
Engineer - PySpark/SQL

Barclays • Maharashtra

On-site
INR 1,200,000 - 1,800,000
Cloud Data Engineer
Cloud Data Engineer

Barclays • Maharashtra

On-site
INR 1,200,000 - 1,800,000