Stand out for this role — generate a tailored resume and cover letter in about a minute.
Altamira Technologies is seeking a Data Scientist to design, implement, and maintain data pipelines and ML projects on cloud-native platforms. You will interpret complex data, build predictive models, and communicate insights to stakeholders in a defense/security environment.
The ideal candidate has strong programming, statistics, and domain knowledge, with experience in Python/R, SQL, and visualization tools. A SECRET clearance starting status is preferred or the ability to obtain TS/SCI.
Altamira Technologies has a long and successful history providing innovative solutions throughout the U.S. National Security community. Headquartered in McLean, Virginia, Altamira serves the defense, intelligence and homeland security communities worldwide by focusing on creating innovative solutions leveraging common standards in architecture, data and security. Altamira believes that our people and the culture of our company differentiate us from other companies.
A data scientist will have skills sets of both data analysts and data engineers. Data scientists are responsible for designing, implementing, and maintaining a data pipeline. In addition, data scientists shall interpret and analyze complex sets of data, as well as plan, execute, and manage ML projects with cloud-native platforms and advanced ML solutions. They understand some of the most challenging processes, technologies, and can leverage a vast array of methodologies in the field, such as data mining, natural language programming, and machine learning. Data scientists must have a combination of skills that include programming, mathematical modeling, statistics, and domain knowledge. They must combine an advanced math and statistics background with programming, domain knowledge, and communication skills to analyze data, create applied mathematical models, and present results in a form useful to the organization. They must also be able to understand and manipulate structured and unstructured large data sets, which requires proficiency in distributed SQL programming, relational and non-relational data queries, general programming languages (such as Python and R,) and machine learning techniques.