Get more replies from employers
Send a job-specific resume in minutes.
Object Tech, Inc. seeks a Machine Learning Engineer Intern to join a team building an AI agent for scientific data. You will explore real‑world tables, infer relationships, and help integrate ML techniques with LLMs to deliver practical AI systems.
This fully remote internship offers hands‑on experience, collaboration with domain experts, and opportunities for future growth, including recommendations and referrals. CPT/OPT support is provided for students.
“Why be a star when you can make a constellation?” - Mariam Kaba
Object Tech, Inc. is a leading technology company providing full-stack AI solutions to reform tech enterprises’ technology R&D and production. With the fusion of experienced scientists and engineers in materials science, nanotechnology, semiconductor manufacturing, machine learning, computer vision and software engineering, we are thriving on cutting-edge end-to-end AI solutions to help accelerate deep tech R&D and promote high-yield production. We are committed to fostering a dynamic and inclusive work environment where creativity and innovation thrive.
We’re building an AI agent that helps scientists make sense of large collections of historical experimental data. The first component ingests many heterogeneous data tables and automatically works out how they relate to one another. The role sits at the intersection of machine learning, large language models, and data engineering, in close collaboration with academic domain experts.
This internship provides an opportunity to gain hands‑on experience applying AI techniques to real‑world scientific data challenges, while learning how machine learning, LLMs, and data engineering can work together to build practical AI systems.
Exploring and profiling messy, real‑world scientific tables (CSV/Excel), extracting structure and metadata.
Contributing to a relationship‑inference pipeline that combines classical data‑integration signals (schema and value matching, key detection) with LLM‑based semantic reasoning over column meanings and provenance.
Supporting the development of evaluation methods that measure system performance against expert‑provided ground truth.
Load, clean, and normalize heterogeneous tabular datasets and build reusable data‑profiling tooling.
Prototype and iterate on an LLM/agent pipeline that classifies how pairs of tables relate (parallel / shared / hierarchical).
Design and maintain benchmarks and metrics to evaluate system accuracy, and run experiments to improve performance.
Work with database schemas, including joins, keys, and data modeling, to represent and query discovered relationships.
Collaborate with mentors and domain scientists to understand requirements and turn feedback into concrete improvements.
Document findings, experiments, and technical approaches throughout the project.
BS/MS (in progress or completed) in Data Science, Computer Science, or a closely related field.
Solid grounding in machine learning fundamentals.
Hands‑on database experience, including SQL, schema design, and joins.
Experience with LLMs / AI agents (prompting, RAG, tools like LangChain).
Prior work with scientific or experimental datasets.
Familiarity with data integration, schema matching, or entity resolution.
Good software habits such as version control, testing, and clear documentation.
Fully remote with flexible schedule
Collaborate with team members from leading tech firms (including MAMAA)
Work on high‑impact, real‑world AI marketing projects
Great for students (supports CPT/OPT)
Opportunity for recommendation letters, referrals, and future growth
Direct mentorship from professionals in product, marketing, and AI
Object Tech, Inc. is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.