Senior Data Engineer - Remote (USA)

Reston, VA (Remote)

$90K/yr – $152K/yrSenior Level7+ years expFull time

Posted 1 day ago

Top 1% of applicants do these two things.

The other 99% submit and hope. You're about to do something different.

Job Summary

Own and enhance end-to-end data pipelines and platforms using Spark, Hive, Airflow, and Databricks, delivering data access APIs, visualizations, and robust data quality, with DevOps integration and security practices.

Job Description

Responsibilities

  • Design, develop, and maintain scalable data pipelines using Spark, Hive, and Airflow
  • Develop and deploy data processing workflows on the Databricks platform
  • Develop API services to facilitate data access and integration
  • Create interactive data visualizations and reports using AWS QuickSight
  • Builds required infrastructure for optimal extraction, transformation and loading of data from various data sources using AWS and SQL technologies
  • Monitor and optimize the performance of data infrastructure and processes
  • Develop data quality and validation jobs
  • Assembles large, complex sets of data that meet non-functional and functional business requirements
  • Write unit and integration tests for all data processing code
  • Work with DevOps engineers on CI, CD, and IaC

Requirements

  • Bachelor's degree
  • 7+ years of hands-on software development experience
  • 4+ years of experience building data pipelines using Python, Java, and cloud technologies, with hands-on experience leveraging Spark and Hive for large-scale data processing.
  • Candidate must be able to obtain and maintain a Public Trust clearance
  • Candidate must reside in the US, be authorized to work in the US, and work must be performed in the US
  • Must have lived in the US 3 full years out of the last 5 years

Preferred

  • Experience building job workflows with the Databricks platform
  • Strong understanding of AWS products including S3, Redshift, RDS, EMR, AWS Glue, AWS Glue DataBrew, Jupyter Notebooks, Athena, QuickSight, EMR, and Amazon SNS
  • Familiar with work to build processes that support data transformation, workload management, data structures, dependency and metadata
  • Experienced in data governance process to ingest (batch, stream), curate, and share data with upstream and downstream data users.
  • Experienced in data pipeline builder and data wrangler who enjoys optimizing data systems and building them from the ground up.
  • Demonstrated understanding using software and tools including relational NoSQL and SQL databases including Cassandra and Postgres; workflow management and pipeline tools such as Airflow, Luigi and Azkaban; stream-processing systems like Spark-Streaming and Storm; and object function/object-oriented scripting languages including Scala, C++, Java and Python.
  • Familiar with DevOps methodologies, including CI/CD pipelines (Github Actions) and IaC (Terraform)
  • Ability to obtain and maintain a Public Trust; residing in the United States
  • Experience with Agile methodology, using test-driven development.
ICF logo

ICF

3.6 Glassdoor

ICF is a leading global solutions and technology provider. We combine unmatched expertise with cutting-edge technology to help clients solve their most complex challenges, navigate change, and shape the future.

Founded in 1969
Reston, Virginia, USA
11K employees (647 in marketing)
71% recommend to a friend
80% CEO approval

Traffic Signals

Monthly Visitors
172.8K
Monthly Google Ads Budget
$29K
Traffic Source Mix
Search
30.6%
Direct
47.7%
Referral
12.3%
Social
7.6%
Paid
1.6%

Headcount Trend

Current headcount: ~10.7K

Marketing Team

Marketing Team Size
510 (5% of company)
Median Tenure
4.5 years
Marketing Roles Posted (Last 30 Days)
11

Funding

Total Funding
$59M
Last Raise
$29M
Grant, 3 years ago

Funding History

  • Feb 2023
    Grant • $29M
  • Mar 2021
    Grant • $30M

Recent News