Job Requirements
Washington Dc Brm, DC
Top Secret Polygraph not specified
Mid Level Career (5+ yrs experience)
Salary not specified
Join Premium to unlock estimated salaries
Job Description
Key Technical Responsibilities
Design, develop, and maintain data pipelines and ETL/ELT workflows supporting machine learning, NLP, retrieval, and advanced analytics.
Prepare, transform, curate, and validate structured and unstructured grants and financial data for AI/ML applications.
Perform feature engineering, document processing, chunking, embeddings generation, and dataset preparation for AI-assisted analytics and retrieval workflows.
Develop reproducible data transformations and workflows using Python, SQL, and applicable data engineering technologies.
Implement and maintain data quality, validation, provenance, lineage, versioning, metadata, and source traceability throughout the data lifecycle.
Prepare and integrate data for search, retrieval-augmented generation (RAG), visualization, analytics, and AI-assisted decision support.
Troubleshoot data pipeline, transformation, integration, and data-quality issues across development and testing environments.
Collaborate with multidisciplinary data engineering, data science, software engineering, and AI/ML teams during iterative development and testing.
Develop technical documentation covering data architecture, pipelines, transformations, schemas, features, embeddings, dependencies, and operational procedures.
Support deployment and operation of data engineering capabilities within secure Federal, on-premises, cloud, or hybrid environments.
Ensure data engineering solutions comply with Federal security, privacy, accessibility, records-management, data-ownership, and governance requirements.
Required Technical Qualifications
Demonstrated experience building, integrating, and operating data pipelines for machine learning, NLP, information retrieval, or advanced analytics.
Hands-on experience with feature engineering, document processing, embeddings, curated datasets, data transformation, and reproducible data workflows.
Strong proficiency in Python and SQL for data engineering, data transformation, automation, and model-support workflows.
Experience working with structured and unstructured data and preparing data for AI/ML or analytics applications.
Experience implementing data quality controls, data validation, provenance, metadata, versioning, lineage, and source traceability for AI/ML or data-intensive systems.
Experience troubleshooting and optimizing data pipelines, integrations, transformations, and data-processing workflows.
Ability to develop clear technical documentation covering data pipelines, schemas, transformations, dependencies, and operational processes.
Strong technical communication, problem-solving, and cross-functional collaboration skills.
Design, develop, and maintain data pipelines and ETL/ELT workflows supporting machine learning, NLP, retrieval, and advanced analytics.
Prepare, transform, curate, and validate structured and unstructured grants and financial data for AI/ML applications.
Perform feature engineering, document processing, chunking, embeddings generation, and dataset preparation for AI-assisted analytics and retrieval workflows.
Develop reproducible data transformations and workflows using Python, SQL, and applicable data engineering technologies.
Implement and maintain data quality, validation, provenance, lineage, versioning, metadata, and source traceability throughout the data lifecycle.
Prepare and integrate data for search, retrieval-augmented generation (RAG), visualization, analytics, and AI-assisted decision support.
Troubleshoot data pipeline, transformation, integration, and data-quality issues across development and testing environments.
Collaborate with multidisciplinary data engineering, data science, software engineering, and AI/ML teams during iterative development and testing.
Develop technical documentation covering data architecture, pipelines, transformations, schemas, features, embeddings, dependencies, and operational procedures.
Support deployment and operation of data engineering capabilities within secure Federal, on-premises, cloud, or hybrid environments.
Ensure data engineering solutions comply with Federal security, privacy, accessibility, records-management, data-ownership, and governance requirements.
Required Technical Qualifications
Demonstrated experience building, integrating, and operating data pipelines for machine learning, NLP, information retrieval, or advanced analytics.
Hands-on experience with feature engineering, document processing, embeddings, curated datasets, data transformation, and reproducible data workflows.
Strong proficiency in Python and SQL for data engineering, data transformation, automation, and model-support workflows.
Experience working with structured and unstructured data and preparing data for AI/ML or analytics applications.
Experience implementing data quality controls, data validation, provenance, metadata, versioning, lineage, and source traceability for AI/ML or data-intensive systems.
Experience troubleshooting and optimizing data pipelines, integrations, transformations, and data-processing workflows.
Ability to develop clear technical documentation covering data pipelines, schemas, transformations, dependencies, and operational processes.
Strong technical communication, problem-solving, and cross-functional collaboration skills.
group id: 10105424