Overview
Peraton is a next-generation national security company that drives missions of consequence spanning the globe and extending to the farthest reaches of the galaxy. As the world’s leading mission capability integrator and transformative enterprise IT provider, we deliver trusted, highly differentiated solutions and technologies to protect our nation and allies. Peraton operates at the critical nexus between traditional and nontraditional threats across all domains: land, sea, space, air, and cyberspace. The company serves as a valued partner to essential government agencies and supports every branch of the U.S. armed forces. Each day, our employees do the can’t be done by solving the most daunting challenges facing our customers. Visitperaton.comto learn how we’re keeping people around the world safe and secure.
Responsibilities
TheDatabricks Data Engineerwill be responsible for the hands-on build-out of a Government-owned Databricks workspace and the data ingestion/integration work needed to consolidate agency business system data. This includes designing and implementing data pipelines from core enterprise systems, implementing Unity Catalog for data governance, configuring MLflow for machine learning workflows, and developing comprehensive documentation to support sustainment beyond the pilot.This role is 100% Remote.Other Responsibilities Include:Build out and configure a dedicated Government workspace within the existing Databricks environmentDesign and implement data ingestion pipelines from core agency business systems including financial, HR, CRM, and ITSM systemsLeverage native/built-in connectors where source systems support them; design custom integration approaches for legacy systemsNormalize and prepare ingested data within Databricks for consumption by downstream visualization/reporting toolsImplement Databricks Unity Catalog for centralized data governance, metadata management, active auditing, and end-to-end lineage trackingReview current configuration, assess security controls for CUI/PII/PHI/financial data, and implement improvementsDevelop comprehensive "as-built" documentation including physical/logical architecture diagrams, automated data dictionaries, and SOPsDocument data sources, integration methods, and data lake architecture decisions to support sustainment
Qualifications
Required Qualifications:Min 8 years with BS/BA, Min 6 years with MS/MA, Min 3 years with PhDUS CitizenshipActive DoD Secret clearance2 years working on the Databricks platformStrong proficiency in PySpark, Spark SQL, and Python for large-scale data processing and pipeline developmentHands-on experience with Unity Catalog administration, including metastore management, access policies, and data lineageExperience with pipeline orchestration tools (Airflow, Databricks Workflows)Experience with Delta Lake, schema evolution, time travel, and optimization techniquesExperience integrating heterogeneous enterprise systems, including legacy/custom integrationsFamiliarity with cloud platforms (Azure/AWS) and infrastructure-as-code practicesUnderstanding of data governance principles, compliance frameworks, and CUI/PII handling practicesFamiliarity with financial, HR, CRM, or ITSM data structuresGit/CI-CD pipeline experienceDoD 8570 certificationPreferred Qualifications:Databricks Certified Data Engineer Associate or Professional certification