Senior Full Stack Data Platform Engineer

Millennium

Bengaluru

Presencial

INR 1 800 000 - 2 800 000

Tempo integral

14 dias+
Gerador de candidaturas

Uma candidatura feita para esta oferta — um currículo e uma carta de apresentação personalizados que vão ao encontro do anúncio.

Ultrapassa os filtros ATS

Resumo da oferta

Millennium in Bengaluru is seeking a Senior Full Stack Software Engineer to architect high-throughput data platforms using Python, Java, and C++. You will build APIs, process unstructured text/audio/video, and enable DAG-based workflows, bridging ML research with real-time trading decisions.

The role requires strong data engineering experience, expertise across SQL/NoSQL, and cloud-native pipelines on AWS/GCP.

Qualificações

  • 5+ years of software engineering experience, preferably building data platforms.
  • Strong proficiency in Python and Java/C++ and ability to switch between them.
  • Experience building data pipelines, ETL processes, or with big data frameworks.
  • Experience with unstructured data types (Text, Audio, Documents) and OCR/transcription.
  • Proficiency in SQL and NoSQL engines (Elasticsearch, Redis, Solr, MongoDB or equivalent).
  • Cloud-native experience in AWS/GCP with serverless data lakes or pipelines.

Responsabilidades

  • High-Performance Data Pipelines: Architect low-latency, high-throughput platform that ingests and normalizes unstructured data.
  • AI & ML Integration: Build infrastructure that serves NLP/ML models in real-time within data streams.
  • Backend Microservices: Develop services for metadata management, search, and retrieval.
  • System Optimization: Tune databases, serialization, and network calls for speed in financial contexts.
  • Data Strategy: Implement storage strategies using vector databases and distributed file systems.

Conhecimentos

Data Platform Experience
Python/Java/C++ proficiency
Data Engineering
Unstructured Data Processing
SQL/NoSQL
Cloud Native

Ferramentas

Kafka
Airflow
Apache Parquet
Arrow
Iceberg
KDB

Descrição da oferta de emprego

Founded in 1989, Millennium is a global alternative investment management firm. Millennium seeks to pursue a diverse array of investment strategies across industry sectors, asset classes and geographies. The firm’s primary investment areas are Fundamental Equity, Equity Arbitrage, Fixed Income, Commodities and Quantitative Strategies. We solve hard and interesting problems at the intersection of computer science, finance, and mathematics. We are focused on innovating and rapidly applying innovations to real world scenarios. This enables engineers to work on interesting problems, learn quickly and have deep impact to the firm and the business.

At Millennium, we are redefining how investment decisions are made. We don't just look at balance sheets; we harness the chaos of the real world. By analyzing vast amounts of unstructured data—from news briefings and earnings call audio to regulatory documents,we provide our Portfolio Managers (PMs) with the "informational edge" (Alpha) they need to outperform the market.

The Role

We are seeking a Senior Full Stack Software Engineer with deep expertise in building high-throughput data platforms. In this role, you will architect scalable data platforms using Python, Java, C++, build robust APIs, and enable processing of data using genAI techniques. You will build and optimize a config-driven, plugin-enabled data platform that will allow the construction of DAGs for data processing. You will then apply the platform to build reusable components and pipelines that will ingest gigabytes of unstructured text, audio, and video. You will enable a variety of rich data consumption use-cases by building the right abstractions and APIs for data consumers. You will be the bridge between complex ML research and real-time trading decisions, working in a poly-language environment (Python, Java, C++) where performance is paramount.

Key Responsibilities
  • High-Performance Data Pipelines: Architect low-latency, high-throughput platform that enables rapid development of pipelines to ingest and normalize unstructured data (PDFs, news feeds, audio streams).
  • AI & ML Integration: Build the infrastructure that wraps and serves NLP and ML models. You will ensure that model inference happens in real-time within the data stream.
  • Backend Microservices: Develop robust backend services to handle metadata management, search, and retrieval of processed alternative data.
  • System Optimization: Tune the platform for speed. In financial markets, milliseconds matter; you will optimize database queries, serialization, and network calls to ensure data reaches the PMs instantly.
  • Data Strategy: Implement storage strategies for unstructured data, utilizing Vector Databases for semantic search and Distributed File Systems for raw storage.
Required Qualifications
  • Data Platform Experience: Minimum 5+ years of software engineering experience, preferably building data platforms.
  • Core Languages: Strong proficiency in both Python and Java/C++ is required. You should be comfortable switching between these languages for different use cases (e.g., Python for data processing, Java or C++ for high-concurrency, scalable services).
  • Data Engineering: Proven experience building data pipelines, ETL processes, or working with big data frameworks (e.g., Kafka, Airflow, Apache Parquet, Arrow, Iceberg , KDB etc).
  • Unstructured Data Expertise: Proven experience working with unstructured data types (Text, Audio, Documents). Familiarity with techniques such as OCR, transcription normalization, text extraction.
  • Database Knowledge: Proficiency in SQL and significant experience with search/NoSQL engines (Elasticsearch, Redis, Solr, MongoDB or equivalent).
  • Cloud Native: Experience building serverless data lakes or processing pipelines on AWS/GCP, etc
Preferred Qualifications
  • AI/NLP Exposure: Experience working with Large Language Models (LLMs), Vector Databases (Pinecone, Milvus, Weaviate), or NLP libraries (Hugging Face, spaCy) or similar
  • Frontend Competence: Solid experience with modern frontend frameworks (React, Vue, or Angular) and data visualization libraries (e.g., D3.js, Highcharts, or AG Grid).
  • Financial Knowledge: Understanding of financial instruments (Equities, Fixed Income) or the investment lifecycle.
  • Document Processing: Familiarity with parsing complex document structures (Earnings calls transcripts, 10-K/10-Q filings, Broker Research, Sector and Industry Reports, Central Bank documents, news wires, social media, etc).
Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior Full Stack Data Platform Engineer
Senior Full Stack Data Platform Engineer

Millennium Management LLC • Bengaluru

Presencial
INR 3 500 000 - 7 000 000
Software Engineer, Post-Trade Platform (Java/Angular)
Software Engineer, Post-Trade Platform (Java/Angular)

Millennium Management LLC • Bengaluru

Presencial
INR 1 200 000 - 2 400 000
Platform / Site Reliability Engineer
Platform / Site Reliability Engineer

Millennium Consulting • Bengaluru

Presencial
INR 1 800 000 - 2 400 000
Tech Lead, Post-Trade Platform (Java/Angular)
Tech Lead, Post-Trade Platform (Java/Angular)

Millennium Management LLC • Bengaluru

Presencial
INR 3 000 000 - 6 000 000
Infra Developer
Infra Developer

Millennium Management LLC • Bengaluru

Presencial
INR 1 200 000 - 1 800 000
Infra Developer
Infra Developer

Millennium • Bengaluru

Presencial
INR 2 800 000 - 4 500 000
Data Scientist
Data Scientist

Millennium Management LLC • Bengaluru

Presencial
INR 1 200 000 - 1 900 000
Full Stack Developer - (C# / .NET)
Full Stack Developer - (C# / .NET)

Millennium • Bengaluru

Presencial
INR 2 400 000 - 4 200 000
Senior Data Engineer
Senior Data Engineer

Vertage Global • Chennai District

Presencial
INR 2 400 000 - 4 200 000
Senior Data Engineer - Vice President - Python Development
Senior Data Engineer - Vice President - Python Development

Citibank (Switzerland) AG • Pune District

Híbrido
INR 3 500 000 - 5 200 000