Ability to hold a position of public trust with the US government
Bachelor's degree and 5 years total experience
2-4 years industry experience coding commercial software and a passion for solving complex problems
2-4 years direct experience in Data Engineering with experience in tools such as:
Big data tools: Hadoop, Spark, Kafka, etc
Relational SQL and NoSQL databases, including Postgres and Cassandra
Data pipeline and workflow management tools: Azkaban, Luigi, Airflow, etc
AWS cloud services: EC2, EMR, RDS, Redshift
Data streaming systems: Storm, Spark-Streaming, etc
Search tools: Solr, Lucene, Elasticsearch
Object-oriented/object function scripting languages: Python, Java, C , Scala, etc
Advanced working SQL knowledge and experience working with relational databases, query authoring and optimization (SQL) as well as working familiarity with a variety of databases
Experience with message queuing, stream processing, and highly scalable ‘big data’ data stores
Experience manipulating, processing, and extracting value from large, disconnected datasets
Experience manipulating structured and unstructured data for analysis
Experience constructing complex queries to analyze results using databases or in a data processing development environment
Experience with data modeling tools and process
Experience architecting data systems (transactional and warehouses)
Experience aggregating results and/or compiling information for reporting from multiple datasets
Experience working in an Agile environment
Experience supporting project teams of developers and data scientists who build web-based interfaces, dashboards, reports, and analytics/machine learning models