Data Scientist

Beavercreek, OH, US • Posted 1 day ago • Updated 2 hours ago
Full Time
On-site
Fitment

Dice Job Match Score™

🧠 Analyzing your skills...

Job Details

Skills

  • Data Architecture
  • Prototyping
  • Organizational Skills
  • Analytical Skill
  • DevOps
  • System Integration
  • Data Management
  • Application Development
  • Data Analysis
  • Unstructured Data
  • Statistical Models
  • Data Storage
  • Management
  • Amazon S3
  • SQL
  • MongoDB
  • Apache HBase
  • Apache Atlas
  • Apache Kafka
  • Python
  • R
  • Data Manipulation
  • Pandas
  • NumPy
  • Analytics
  • Jupyter
  • Apache Spark
  • MATLAB
  • Data Modeling
  • Extraction
  • Forms
  • Communication
  • Security Clearance
  • Computer Science
  • Data Science
  • Information Systems
  • Cloud Computing
  • Amazon Web Services
  • Microsoft Azure
  • Google Cloud Platform
  • Google Cloud
  • Data Visualization
  • Git
  • Version Control
  • Issue Tracking
  • Continuous Integration
  • Continuous Delivery
  • Jenkins
  • GitLab
  • Artificial Intelligence
  • Amazon SageMaker
  • Machine Learning (ML)
  • TensorFlow
  • Keras
  • scikit-learn
  • Team Leadership
  • C++
  • Java
  • Linux

Summary

Overview

VTG is seeking a Data Scientist to support our Team in Beavercreek, OH. The Data Scientist will design, prototype, and implement a data management and application development pipeline in support of national defense data science and data architecture prototyping tasks. This role will also include gathering and organizing data, conducting data analytics, and developing data analytic and AI/ML based applications. This is an onsite role due to its classification level.

What will you do?

The Data Scientist will work with a team of DevOps engineers, software developers, data engineers, and system operators to identify data needs and prototype a range of novel solutions. This data scientist would be involved at all levels of the data life cycle from onboard management of data to its use in application development and back to application integration and gathering test data.
  • Leverage third-party tools to architect and prototype a modern data management and application development pipeline in a local and/or a cloud environment
  • Perform data analytics of simulated and real-world data
  • Integrate structured and unstructured data from disparate data sources
  • Develop applications and models supporting various users
  • Provide technical input to program managers and government representatives

Do you have what it takes?

Required qualifications:
  • Bachelor's Degree, majoring in majoring in Computer Science, Data Science, Information Systems, or a related field
  • 4+ years of experience as a Data Scientist including experience in statistical modeling and machine learning based on the analysis of large sets of data
  • Experience with data storage and management tools (S3, SQL, MongoDB, Hbase, Apache Atlas, Kafka, etc.)
  • Programming experience in Python, R, or similar data manipulation languages and associated libraries (e.g. pandas, numpy, polars, dask)
  • Experience with data science and analytics toolsets (e.g. JupyterHub / Jupyter Notebooks, Apache Spark, MATLAB)
  • Knowledge of data modeling principles
  • Experience in knowledge extraction and insights from data in various forms, both structured and unstructured
  • Cloud development experience, preferably in AWS
  • Excellent verbal and written communication skills
  • with current TOP SECRET/SCI Eligible Clearance or ability to obtain a TOP SECRET/SCI clearance
  • Successful completion of background check

Desired qualifications:
  • Master's Degree or higher in Computer Science, Data Science, or Information Systems
  • Experience establishing data pipelines in cloud platforms, such as AWS, Azure, or Google Cloud
  • Data visualization experience and associated tools/libraries (e.g. pyplot, seaborn)
  • Experience using Git for version control and issue tracking
  • Experience with artifact repositories (e.g. Artifactory)
  • Experience with CI/CD pipelines (e.g. Jenkins, Gitlab pipelines)
  • Experience with AI/ML development tools and libraries (e.g. Sagemaker, ML Studio, Tensorflow, Keras, scikit-learn)
  • Experience leading teams and projects
  • Programming experience in C++ and Java
  • Experience with Linux systems
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: RTL806649
  • Position Id: 2a01c0860163baaf47e84d8c890d3abe
  • Posted 1 day ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Dublin, Ohio

Today

Full-time

Virginia

Today

Easy Apply

Full-time

Compensation information provided in the description

North Carolina

Today

Full-time

USD 73,000.00 - 180,000.00 per year

No location provided

Today

Full-time

Search all similar jobs