Senior Full Stack Data Platform Engineer

Millennium

Bengaluru

On-site

INR 1,800,000 - 2,800,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Millennium in Bengaluru is seeking a Senior Full Stack Software Engineer to architect high-throughput data platforms using Python, Java, and C++. You will build APIs, process unstructured text/audio/video, and enable DAG-based workflows, bridging ML research with real-time trading decisions.

The role requires strong data engineering experience, expertise across SQL/NoSQL, and cloud-native pipelines on AWS/GCP.

Qualifications

  • 5+ years of software engineering experience, preferably building data platforms.
  • Strong proficiency in Python and Java/C++ and ability to switch between them.
  • Experience building data pipelines, ETL processes, or with big data frameworks.
  • Experience with unstructured data types (Text, Audio, Documents) and OCR/transcription.
  • Proficiency in SQL and NoSQL engines (Elasticsearch, Redis, Solr, MongoDB or equivalent).
  • Cloud-native experience in AWS/GCP with serverless data lakes or pipelines.

Responsibilities

  • High-Performance Data Pipelines: Architect low-latency, high-throughput platform that ingests and normalizes unstructured data.
  • AI & ML Integration: Build infrastructure that serves NLP/ML models in real-time within data streams.
  • Backend Microservices: Develop services for metadata management, search, and retrieval.
  • System Optimization: Tune databases, serialization, and network calls for speed in financial contexts.
  • Data Strategy: Implement storage strategies using vector databases and distributed file systems.

Skills

Data Platform Experience
Python/Java/C++ proficiency
Data Engineering
Unstructured Data Processing
SQL/NoSQL
Cloud Native

Tools

Kafka
Airflow
Apache Parquet
Arrow
Iceberg
KDB

Job description

Founded in 1989, Millennium is a global alternative investment management firm. Millennium seeks to pursue a diverse array of investment strategies across industry sectors, asset classes and geographies. The firm’s primary investment areas are Fundamental Equity, Equity Arbitrage, Fixed Income, Commodities and Quantitative Strategies. We solve hard and interesting problems at the intersection of computer science, finance, and mathematics. We are focused on innovating and rapidly applying innovations to real world scenarios. This enables engineers to work on interesting problems, learn quickly and have deep impact to the firm and the business.

At Millennium, we are redefining how investment decisions are made. We don't just look at balance sheets; we harness the chaos of the real world. By analyzing vast amounts of unstructured data—from news briefings and earnings call audio to regulatory documents,we provide our Portfolio Managers (PMs) with the "informational edge" (Alpha) they need to outperform the market.

The Role

We are seeking a Senior Full Stack Software Engineer with deep expertise in building high-throughput data platforms. In this role, you will architect scalable data platforms using Python, Java, C++, build robust APIs, and enable processing of data using genAI techniques. You will build and optimize a config-driven, plugin-enabled data platform that will allow the construction of DAGs for data processing. You will then apply the platform to build reusable components and pipelines that will ingest gigabytes of unstructured text, audio, and video. You will enable a variety of rich data consumption use-cases by building the right abstractions and APIs for data consumers. You will be the bridge between complex ML research and real-time trading decisions, working in a poly-language environment (Python, Java, C++) where performance is paramount.

Key Responsibilities
  • High-Performance Data Pipelines: Architect low-latency, high-throughput platform that enables rapid development of pipelines to ingest and normalize unstructured data (PDFs, news feeds, audio streams).
  • AI & ML Integration: Build the infrastructure that wraps and serves NLP and ML models. You will ensure that model inference happens in real-time within the data stream.
  • Backend Microservices: Develop robust backend services to handle metadata management, search, and retrieval of processed alternative data.
  • System Optimization: Tune the platform for speed. In financial markets, milliseconds matter; you will optimize database queries, serialization, and network calls to ensure data reaches the PMs instantly.
  • Data Strategy: Implement storage strategies for unstructured data, utilizing Vector Databases for semantic search and Distributed File Systems for raw storage.
Required Qualifications
  • Data Platform Experience: Minimum 5+ years of software engineering experience, preferably building data platforms.
  • Core Languages: Strong proficiency in both Python and Java/C++ is required. You should be comfortable switching between these languages for different use cases (e.g., Python for data processing, Java or C++ for high-concurrency, scalable services).
  • Data Engineering: Proven experience building data pipelines, ETL processes, or working with big data frameworks (e.g., Kafka, Airflow, Apache Parquet, Arrow, Iceberg , KDB etc).
  • Unstructured Data Expertise: Proven experience working with unstructured data types (Text, Audio, Documents). Familiarity with techniques such as OCR, transcription normalization, text extraction.
  • Database Knowledge: Proficiency in SQL and significant experience with search/NoSQL engines (Elasticsearch, Redis, Solr, MongoDB or equivalent).
  • Cloud Native: Experience building serverless data lakes or processing pipelines on AWS/GCP, etc
Preferred Qualifications
  • AI/NLP Exposure: Experience working with Large Language Models (LLMs), Vector Databases (Pinecone, Milvus, Weaviate), or NLP libraries (Hugging Face, spaCy) or similar
  • Frontend Competence: Solid experience with modern frontend frameworks (React, Vue, or Angular) and data visualization libraries (e.g., D3.js, Highcharts, or AG Grid).
  • Financial Knowledge: Understanding of financial instruments (Equities, Fixed Income) or the investment lifecycle.
  • Document Processing: Familiarity with parsing complex document structures (Earnings calls transcripts, 10-K/10-Q filings, Broker Research, Sector and Industry Reports, Central Bank documents, news wires, social media, etc).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform / Site Reliability Engineer
Platform / Site Reliability Engineer

Millennium Consulting • Bengaluru

On-site
INR 1,800,000 - 2,400,000
AI Engineer
AI Engineer

Millennium • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Full Stack Software Engineer - Data Entitlements Platform
Full Stack Software Engineer - Data Entitlements Platform

Millennium • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Senior Associate - Data Engineer
Senior Associate - Data Engineer

Acuity Analytics • Gurugram District

On-site
INR 1,800,000 - 3,000,000
Full Stack Developer - (C# / .NET)
Full Stack Developer - (C# / .NET)

Millennium • Bengaluru

On-site
INR 2,400,000 - 4,200,000
Software Engineer - Applied AI
Software Engineer - Applied AI

Hedgineer • Bengaluru

On-site
INR 1,800,000 - 4,000,000
Delivery Lead-Data Engineer
Delivery Lead-Data Engineer

Acuity Analytics • Gurugram District

On-site
INR 1,500,000 - 2,500,000
Full Stack Software Engineer
Full Stack Software Engineer

Millennium • Bengaluru

On-site
INR 2,000,000 - 4,000,000
Lead Software Engineer -Python or Java, big data
Lead Software Engineer -Python or Java, big data

United States Digital Space LLC • Karnataka

On-site
INR 1,500,000 - 2,500,000
Infrastructure Engineer
Infrastructure Engineer

Millennium • Bengaluru

On-site
INR 600,000 - 900,000