Senior Data & Python Software Engineer

United States Digital Space LLC

Berlin

Vor Ort

EUR 70.000 - 100.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

United States Digital Space LLC in Berlin seeks a Data & Python Software Engineer to drive scalable data pipelines and web crawling technologies, collaborating with CTO and Head of Engineering to advance our data engineering efforts for digital security and brand protection.

You will design high-performance scraping systems, implement APIs, ensure data quality and governance, and optimize reliability at scale across ingestion, processing, and storage layers.

Qualifikationen

  • Experience with web scraping
  • Strong SQL skills
  • Strong Python experience
  • Experience with PostgreSQL or similar relational databases
  • Experience designing and building scalable APIs and backend services (e.g. FastAPI, Django, or similar frameworks)
  • Experience deploying and operating systems in the cloud (AWS, GCP, or Azure)
  • Experience with Docker and containerized environments
  • Experience with workflow orchestration tools such as Airflow
  • Experience with browser-based automation tools (Playwright, Selenium, or similar)

Aufgaben

  • Design, build, and maintain high-performance web scraping systems as well backend services and data pipelines supporting web data extraction and brand protection use cases
  • Implement and maintain scraping focused APIs and other data services that power internal products and external integrations
  • Build reliable ingestion, processing, and storage workflows for large-scale web data
  • Handle cleaning of web data and ensure data quality, validation, and governance across ingestion, storage, and serving layers
  • Optimize scraping systems for performance, scalability, reliability, and cost efficiency
  • Monitor, debug, and improve scraping system reliability using observability tools (logging, metrics, tracing)
  • Collaborate closely with product and engineering teams to deliver features from design through full end-to-end production deployment
  • Take independent ownership of systems in production, including maintenance, iteration and performance management

Kenntnisse

Web scraping
Strong SQL
Strong Python

Tools

FastAPI
Django
PostgreSQL
Docker
AWS

Jobbeschreibung

At the company, we lead the way in AI-powered brand protection, copyright law, and digital security,

safeguarding the integrity of content creators, brands, and enterprises worldwide. As we scale

rapidly, we're looking for a Data \& Python Software engineer to drive innovation in our data

pipelines and crawling technologies. In this pivotal role, you'll collaborate with our CTO and

Head of Engineering, steering our Data Engineering Team toward developing groundbreaking

solutions for digital security challenges.

Build Scalable Web Data Extraction Pipelines:

Design and develop web scraping systems that support large-scale web data extraction and brand protection workflows.

Ensure that data moves reliably from collection through processing to storage while maintaining performance, resilience, and operational stability at scale.

Ensure Data Quality and Governance:

Own data validation, consistency, and governance across ingestion, storage, and serving layers.

Establish clear standards for schema design, transformation logic, and monitoring to guarantee trustworthy, production-grade datasets that can be reliably consumed across the organization.

Optimize Performance and Reliability:

Continuously improve scraping system efficiency through performance tuning, cost optimization, and architectural enhancements. Implement logging, metrics, and tracing to monitor production systems, diagnose issues quickly, and maintain high reliability under growing workloads.

Responsibilities:
  • Design, build, and maintain high-performance web scraping systems as well backend services and data pipelines supporting web data extraction and brand protection use cases
  • Implement and maintain scraping focused APIs and other data services that power internal products and external integrations
  • Build reliable ingestion, processing, and storage workflows for large-scale web data
  • Handle cleaning of web data and ensure data quality, validation, and governance across ingestion, storage, and serving layers
  • Optimize scraping systems for performance, scalability, reliability, and cost efficiency
  • Monitor, debug, and improve scraping system reliability using observability tools (logging, metrics, tracing)
  • Collaborate closely with product and engineering teams to deliver features from design through full end-to-end production deployment
  • Take independent ownership of systems in production, including maintenance,
  • iteration and performance management
Core Technical Requirements:
  • Experience with web scraping
  • Strong SQL skills
  • Strong Python experience
  • Experience with PostgreSQL or similar relational databases
  • Experience designing and building scalable APIs and backend services (e.g. FastAPI, Django, or similar frameworks)
  • Experience designing efficient, scalable data models and database schemas
  • Hands-on experience deploying and operating systems in the cloud (AWS, GCP, or Azure)
  • Experience working with Docker and containerized environments
Preferred Technical Requirements:
  • Experience with workflow orchestration tools such as Airflow
  • Experience with browser-based automation tools (Playwright, Selenium, or similar)
  • Experience with DBT or analytics-focused data transformation workflows
  • Experience building or operating high-concurrency systems and task queues
  • Experience designing and deploying cloud-native workflows on AWS
  • Familiarity with CI/CD pipelines and production deployment practices
  • Experience working in a high-growth, early-stage startup environment
  • Experience - University education in a technical field such as Computer Science, Engineering or similar. Masters level preferred. 4+ years ( or 2 year+ in a early stage startup)
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Data Software Engineer
Senior Data Software Engineer

TrioTech Recruitment • Berlin

Vor Ort
EUR 75.000 - 95.000
Free Breakfast & Lunch
Data Engineer / Data Scientist
Data Engineer / Data Scientist

Jobtailor • Berlin

Vor Ort
EUR 70.000 - 110.000
Senior Data Engineer (ADB + Python)
Senior Data Engineer (ADB + Python)

Bonapolia • Deutschland

Remote
EUR 90.000 - 150.000
Big Data Engineer
Big Data Engineer

Embedded Shishya • Deutschland

Hybrid
EUR 70.000 - 110.000
Data & Analytics Engineer
Data & Analytics Engineer

Jobtailor • Hamburg

Vor Ort
EUR 65.000 - 95.000
Staff Data Engineer - Sponsored Content & Products (all genders)
Staff Data Engineer - Sponsored Content & Products (all genders)

ABOUT YOU SE & Co. KG • Hamburg

Vor Ort
EUR 85.000 - 100.000
Senior Data Engineer
Senior Data Engineer

Onapsis • Heidelberg

Vor Ort
EUR 70.000 - 90.000
Competitive compensation
Supportive and humble colleagues
Career growth opportunities
Senior Software Engineer - Data Platform (d/f/m, Berlin)
Senior Software Engineer - Data Platform (d/f/m, Berlin)

Monda • Berlin

Hybrid
Data Engineer
Data Engineer

OpenSC • Deutschland

Vor Ort
EUR 60.000 - 100.000
Senior Data Acquisition Engineer
Senior Data Acquisition Engineer

Bluefish • Köln

Vor Ort
EUR 70.000 - 90.000