AI & Data Engineering Intern

Growth For Impact

Bengaluru

Ibrido

INR 1.800.000 - 3.000.000

Tempo pieno

14 giorni+
Generatore di candidature

Non inviare un curriculum generico — genera un curriculum e una lettera di presentazione personalizzati per questo specifico impiego.

Supera i filtri ATS

Vantaggi offerti da questo lavoro

Health insurance coverage
Unlimited leaves & flexible working
Role-based remote work and work-from-
Relocation assistance
Professional Mental Wellness services
Employee Stock Options for all hires

Descrizione del lavoro

CleanMax in Bengaluru, India, seeks a data/ML engineer to drive satellite change detection projects. You will design datasets, build scalable pipelines, and develop evaluation dashboards to showcase results for competition judges.

You will optimize inference, develop captioning for changes, and build interfaces to interpret outputs. This role requires independent work, strong Python skills, and a passion for fast-paced ML innovation.

Competenze

  • Strong programming in Python with PyTorch/TensorFlow, NumPy, Pandas.
  • Experience with data pipelines, data processing, or data engineering workflows.
  • Understanding of end-to-end ML lifecycle from training to deployment.
  • Ability to work independently on well-defined workstreams.
  • Attention to detail for high-quality data.
  • Eagerness to learn quickly in a fast-paced, competition-driven environment.
  • Bachelor’s degree in CS/ML/DS/Engineering/Physics or related field.

Mansioni

  • Design and build training and benchmarking datasets across multiple sensors and tasks.
  • Develop scalable data pipelines for ingestion, preprocessing, and dataset export.
  • Create interactive dashboards to visualize change detection results for judges.
  • Optimize model inference latency, throughput, and memory usage.
  • Develop a change captioning system using time-step imagery and masks.
  • Build interfaces to query and reason about changes via conversational or programmatic interfaces.
  • Benchmark models against ground truth and evaluation metrics.
  • Iterate on model architectures and training strategies based on feedback.
  • Write clean, efficient Python code and contribute to shared ML utilities.
  • Collaborate with cross-functional teams and document processes.

Conoscenze

Python
PyTorch
TensorFlow
NumPy
Pandas
Data pipelines
End-to-end ML lifecycle
Independent work
Attention to detail
Fast learner

Formazione

Bachelor’s degree in Computer Science | ML | DS | Engineering | Physics

Strumenti

Descrizione del lavoro

Overview

Role description and responsibilities focusing on data, modeling, and collaboration for satellite change detection competitions.

Responsibilities
  • Designed and built training and benchmarking datasets across two satellite sensors and eight change detection tasks, establishing consistent annotation standards and quality control processes.
  • Built and optimized scalable data pipelines for scene ingestion, preprocessing, chip generation, and dataset export, enabling efficient handling of multi-sensor satellite data.
  • Developed an interactive dashboard to evaluate and demonstrate change detection results for competition judges, providing clear visualizations of detected changes, confidence scores, and model performance metrics.
  • Optimized model inference performance by improving latency, throughput, and memory utilization to meet competition runtime and resource constraints.
  • Develop a change captioning system that generates natural language descriptions of changes by leveraging two time-step satellite images and a change detection mask.
  • Build the agentic interface layer for interacting with change detection outputs, enabling users to query, interpret, and reason about detected changes through conversational or programmatic interfaces.
  • Evaluate trade-offs between model complexity, accuracy, and inference performance.
  • Benchmark and validate models against ground truth datasets and competition evaluation metrics.
  • Continuously iterate on model architectures and training strategies based on performance feedback.
  • Write clean, efficient, and maintainable Python code for data processing and model development.
  • Contribute to shared utilities and machine learning libraries used across the competition team.
  • Collaborate closely with cross-functional teams by: Providing feedback on dataset quality and annotation issues, Partnering with the ML team on model evaluation and performance improvements, Coordinating with data engineers to optimize data pipelines.
  • Document datasets, models, pipelines, and technical processes comprehensively to ensure reproducibility and provide clear technical insights for team members and competition judges.
Qualifications
  • Strong programming skills in Python with hands-on experience using libraries and frameworks such as PyTorch, TensorFlow, NumPy, Pandas, or equivalent.
  • Practical experience or strong academic coursework in data pipelines, data processing, or data engineering workflows.
  • Basic understanding of the end-to-end machine learning lifecycle, including model training, evaluation, and deployment at scale.
  • Ability to work independently on well-defined workstreams, take ownership of deliverables, and proactively seek guidance from senior scientists when needed.
  • Strong attention to detail, with an understanding that high-quality data is critical to model performance.
  • Eagerness to learn quickly, adapt to new challenges, and thrive in a fast-paced, competition-driven environment.
  • Currently pursuing or recently completed a Bachelor’s degree in Computer Science, Machine Learning, Data Science, Engineering, Physics, or a related field.
Benefits
  • Health insurance coverage
  • Unlimited leaves & flexible working hours
  • Role-based remote work and work-from-home benefit
  • Relocation assistance
  • Professional Mental Wellness services
  • Employee Stock Options for all hires
Ottieni la revisione del curriculum gratis e riservata.

o trascina qui il file.

Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Geospatial Data Scientist
Geospatial Data Scientist

Growth For Impact • Bengaluru

Ibrido
INR 600.000 - 900.000
Health insurance
Unlimited leaves
Remote work option
+3
Data Scientist - 3 / Team Lead
Data Scientist - 3 / Team Lead

SatSure Analytics India • Bengaluru

In loco
INR 3.500.000 - 6.800.000
Medical health cover
Mental health support
Learning allowance
+1
Data Scientist II
Data Scientist II

Eagleview • Bengaluru

In loco
INR 1.800.000 - 2.400.000
Expert Team Lead, Engineering
Expert Team Lead, Engineering

Jobgether SRL • India

In loco
INR 1.800.000 - 3.400.000
Competitive compensation
Equity opportunities
Remote work options
Data Scientist - 2
Data Scientist - 2

SatSure Analytics India • Bengaluru

In loco
INR 1.800.000 - 2.600.000
Medical health cover
Mental health support
Learning allowances
+1
Applied AI Researcher
Applied AI Researcher

GalaxEye • Bengaluru

In loco
INR 800.000 - 1.200.000
Expert Team Lead, SWE
Expert Team Lead, SWE

Jobgether SRL • India

Remoto
INR 1.800.000 - 3.000.000
Competitive compensation
Equity participation
Medical, dental, and vision coverage
+2
Data Scientist
Data Scientist

TP • Bengaluru Urban

In loco
INR 1.200.000 - 1.800.000
AI Engineer/Lead AI Engineer
AI Engineer/Lead AI Engineer

Salesforce • Bengaluru, Hyderabad

In loco
INR 4.000.000 - 8.000.000
AI Engineer
AI Engineer

LMD Consulting • Jaipur

In loco
INR 600.000 - 1.400.000
Competitive salary
Five-day work week
Cutting-edge AI projects
+3