A leading tech firm is seeking an AI/ML Data Scientist for a remote position. The successful candidate will partner with business teams to implement AI/ML solutions and have a strong background in Python, SQL, and cloud technologies. This role requires a Master's degree plus experience in data science. Responsibilities include training models for forecasting, building data pipelines, and creating dashboards to deliver insights. The position offers a full-time contract and is ideal for those with a passion for AI/ML applications.
Qualifications
Masters + 1+ years or PhD in a quantitative field is required.
Strong Python and SQL skills are mandatory.
Experience with CI/CD for ML/data projects is essential.
Responsibilities
Partner with teams to create AI/ML solutions and success metrics.
Train and deploy models for various applications including forecasting.
Build reliable data pipelines across GCP and AWS platforms.
Skills
Python
SQL
scikit-learn
XGBoost
LightGBM
PyTorch
TensorFlow
GenAI/LLM applications
Cloud experience
Education
Masters in Data Science, Computer Science, Applied Math/Statistics, Econometrics
PhD in relevant field
Tools
GCP
AWS
Streamlit
Plotly Dash
Job description
AI/ML Data Scientist (Remote)
We are seeking an AI/ML Data Scientist for our client for a remote position. US Citizens or Green Card Holders ONLY!!No C2CNo Third Party Agencies.
What you'll do
Partner with business and engineering teams to translate problems into measurable AI/ML solutions and success metrics.
Design, train, validate, and deploy models for classification, regression, recommendation, and time‑series forecasting; pick algorithms, features, and evaluation strategies that match business goals.
Develop and evaluate GenAI agent applications using frameworks like Langchain and Google ADK, leveraging techniques such as RAG, prompt engineering, and vector DB integration.
Build and operate reliable data pipelines and model inference endpoints across GCP and AWS (BigQuery, Vertex AI, Cloud Run, S3, Lambda, SageMaker, etc.).
Implement CI/CD, automated testing, and monitoring for ML/data projects (GitHub Actions, Cloud Build, CodeBuild/CodePipeline).
Create lightweight dashboards and internal apps (Streamlit, Plotly Dash) to deliver models and insights to stakeholders.
Write clear model documentation: problem formulation, modeling approach, validation, data needs, and deployment steps.
Advocate for coding best practices, reproducibility, and shared documentation across a global data science organization.
Required qualifications
Masters + 1+ years, or a PhD in Data Science, Computer Science, Applied Math/Statistics, Econometrics, or related quantitative field.
Strong Python and SQL skills; experience with scikit-learn, XGBoost/LightGBM, and a deep learning framework (PyTorch or TensorFlow).
Hands‑on experience building GenAI/LLM applications and agentic workflows.
Practical experience building reliable data pipelines and ensuring data quality/lineage on GCP (BigQuery).
Practical experience building internal dashboards/apps with Streamlit and/or Plotly Dash.
Cloud experience with GCP (BigQuery, Vertex AI, Cloud Run, Gemini Enterprise) and AWS (S3, Lambda, ECS, SageMaker).
Solid comprehension of the mechanisms behind widely used AI/ML algorithms - including their intuition, assumptions, statistical theory, computational complexity, strengths/weaknesses, and when to use each.