Data Scientist
Indexed description
What will you do?
The Data Scientist will work with a team of DevOps engineers, software developers, data engineers, and system operators to identify data needs and prototype a range of novel solutions. This data scientist would be involved at all levels of the data life cycle from onboard management of data to its use in application development and back to application integration and gathering test data.
- Leverage third-party tools to architect and prototype a modern data management and application development pipeline in a local and/or a cloud environment
- Perform data analytics of simulated and real-world data
- Integrate structured and unstructured data from disparate data sources
- Develop applications and models supporting various users
- Provide technical input to program managers and government representatives
- Bachelor's Degree, majoring in majoring in Computer Science, Data Science, Information Systems, or a related field
- 4+ years of experience as a Data Scientist including experience in statistical modeling and machine learning based on the analysis of large sets of data
- Experience with data storage and management tools (S3, SQL, MongoDB, Hbase, Apache Atlas, Kafka, etc.)
- Programming experience in Python, R, or similar data manipulation languages and associated libraries (e.g. pandas, numpy, polars, dask)
- Experience with data science and analytics toolsets (e.g. JupyterHub / Jupyter Notebooks, Apache Spark, MATLAB)
- Knowledge of data modeling principles
- Experience in knowledge extraction and insights from data in various forms, both structured and unstructured
- Cloud development experience, preferably in AWS
- Excellent verbal and written communication skills
- with current TOP SECRET/SCI Eligible Clearance or ability to obtain a TOP SECRET/SCI clearance
- Successful completion of background check
- Master's Degree or higher in Computer Science, Data Science, or Information Systems
- Experience establishing data pipelines in cloud platforms, such as AWS, Azure, or Google Cloud
- Data visualization experience and associated tools/libraries (e.g. pyplot, seaborn)
- Experience using Git for version control and issue tracking
- Experience with artifact repositories (e.g. Artifactory)
- Experience with CI/CD pipelines (e.g. Jenkins, Gitlab pipelines)
- Experience with AI/ML development tools and libraries (e.g. Sagemaker, ML Studio, Tensorflow, Keras, scikit-learn)
- Experience leading teams and projects
- Programming experience in C++ and Java
- Experience with Linux systems
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search