
Data Engineer
hackajob · United States
- Hybrid
- Full-time
- $120,000 / year
- United States
Tailored resume — keyword-matched to this role.
Hiring manager — we find who's hiring.
Intro email — drafted to reach them directly.
Job highlights
- Design and implement data processing and streaming solutions.
- Develop AI-ready data pipelines for multi-source fusion.
- Manage data lineage, metadata, and governance processes.
- Build data services with Spark, Kafka, Trino, Iceberg.
- Requires Bachelor's degree and 4+ years experience.
About the role
Data Engineer
MANTECH is seeking a motivated, career, and customer-oriented Data Engineer for a new initiative within the National Capital Region. This effort supports the rapid design, deployment, operation, and sustainment of enterprise-scale AI, data, and mission platform capabilities across a cloud, edge, and classified operational environment. This role supports enterprise data ingestion, transformation, governance, and analytics pipeline development. You will ensure that data is AI-ready and optimized for operational deployment across the platform.Responsibilities
- Design and implement solutions for distributed data processing and streaming architectures.
- Develop and maintain AI-ready data pipelines and capabilities for multi-source data fusion.
- Establish and manage processes for data lineage, metadata management, and data governance.
- Contribute to the implementation and optimization of the Enterprise Data Layer.
- Build data services using technologies like Spark, Kafka, Trino, and Iceberg for large-scale transformation.
Minimum Qualifications
- Bachelor’s degree in Data Science, Engineering, Computer Science, or a related technical field.
- 4 or more years of experience in data engineering, ETL/ELT pipeline development, or streaming architectures.
- Expertise in distributed processing frameworks (e.g., Spark) and data warehousing concepts.
- Demonstrated proficiency in SQL and Python for advanced data manipulation and scripting.
- Experience with data quality assurance, governance, and preparing data for AI/ML models.
Preferred Qualifications
- Hands-on experience with emerging data technologies like ClickHouse, Iceberg, and/or OpenSearch.
- Familiarity with streaming platforms such as Kafka or Trino.
- Prior experience supporting Intelligence Community data management initiatives.
Clearance Requirements
For onsite work, a TS/SCI clearance with polygraph will be required.Physical Requirements
- The person in this position must be able to remain in a stationary position 50% of the time.
- Frequently communicates with co-workers, management, and customers, which may involve delivering presentations.
- Constantly operates a computer and other office productivity machinery.
Key skills/competency
Data Engineer, AI, Data Pipelines, Spark, Kafka, SQL, Python, Data Governance, ETL, Data WarehousingSkills & topics
- Data Engineer
- Data Engineering
- ETL
- ELT
- Spark
- Kafka
- SQL
- Python
- Data Pipelines
- Data Governance
- Cloud
- Edge Computing
- AI
- Machine Learning
- Trino
- Iceberg
- ClickHouse
- OpenSearch
- TS/SCI
- Polygraph
How to get hired
- Tailor your resume: Highlight your experience with Spark, Kafka, SQL, Python, and AI-ready data pipelines.
- Showcase your qualifications: Emphasize your Bachelor's degree and 4+ years of relevant data engineering experience.
- Address clearance requirements: If applicable, mention your TS/SCI clearance with polygraph for onsite roles.
- Prepare for technical interviews: Be ready to discuss distributed processing, data warehousing, and data governance concepts.
- Demonstrate project impact: Use examples to illustrate your success in designing and implementing data solutions.
Technical preparation
Master Spark for distributed data processing.,Practice complex SQL queries and Python scripting.,Build sample ETL/ELT pipelines.,Understand data warehousing and governance concepts.
Behavioral questions
Describe a challenging data pipeline you built.,How do you ensure data quality and governance?,Explain your experience with AI-ready data.,How do you collaborate on platform capabilities?
Frequently asked questions
- What are the main responsibilities of a Data Engineer at MANTECH?
- As a Data Engineer at MANTECH, you'll be responsible for designing and implementing distributed data processing and streaming architectures, developing AI-ready data pipelines, managing data lineage and governance, and building data services using technologies like Spark, Kafka, Trino, and Iceberg.
- What educational background is required for the Data Engineer role?
- A Bachelor's degree in Data Science, Engineering, Computer Science, or a related technical field is the minimum educational requirement for this Data Engineer position.
- What kind of experience is needed for this Data Engineer job?
- You'll need 4 or more years of experience in data engineering, ETL/ELT pipeline development, or streaming architectures. Expertise in distributed processing frameworks like Spark and proficiency in SQL and Python are also essential.
- Are there specific technologies I should be familiar with for this Data Engineer role?
- Yes, expertise in distributed processing frameworks like Spark is crucial. Familiarity with data warehousing concepts, SQL, Python, and preferred technologies such as ClickHouse, Iceberg, Kafka, Trino, or OpenSearch is highly beneficial.
- What is the work arrangement for this Data Engineer position?
- The job description mentions requirements for onsite work, including a TS/SCI clearance with polygraph. This suggests a preference or requirement for an onsite presence, though the exact arrangement may vary. For onsite work, a TS/SCI clearance with polygraph is required.
- What security clearance is needed for the Data Engineer role?
- For onsite work, a TS/SCI clearance with a polygraph will be required for this Data Engineer position.
- How does MANTECH use data engineering skills for AI and mission platforms?
- MANTECH leverages data engineering expertise to support the rapid design, deployment, and sustainment of enterprise-scale AI, data, and mission platform capabilities. Your work will ensure data is AI-ready and optimized for operational deployment.
- What are the preferred qualifications for this Data Engineer position?
- Preferred qualifications include hands-on experience with emerging data technologies like ClickHouse, Iceberg, and OpenSearch, familiarity with streaming platforms such as Kafka or Trino, and prior experience supporting Intelligence Community data management initiatives.