As a member of the Data team, you will be working with cutting‑edge data engineering and distributed computing problems, improving throughput, reducing latency, maintaining uptime of data pipelines and web services, and writing test‑driven code for processing terabytes of data in multi‑region distributed systems.
Here are some of the challenging projects we are working on as part of the Data team.
- Scaling the current data pipeline to handle 5x of the present scale within the next one year.
- Moving from batch‑oriented processing to a near‑real‑time processing engine.
- Building performance monitoring systems for databases, web services and processing engines.
Roles and Responsibilities
- Thinking big and executing with great focus with a milestone‑based approach rather than a big bang.
- Designing and coding with scale, high availability, and cost efficiency in mind.
- Mentoring and reviewing the code of fellow colleagues.
- Leading a micro team within the team and adopting good technology processes and tools.
- Owning problem statements and solutions built to solve them.
- Open to working on a polyglot tech stack.
Requirements
- 5–7 years of proven experience in developing scalable REST/gRPC services or streaming pipelines and data‑intensive applications.
- Expertise in Java programming language and frameworks.
- Hands‑on experience with data modelling, database design and performance.
- Hands‑on experience with web frameworks such as Vert.X, SpringBoot, Quarkus (a plus).
- Hands‑on experience with data processing technologies such as Kafka, Flink, Spark (a plus).
- Hands‑on experience with containerization, Docker, Kubernetes (a plus).