{"id":1650926,"url":"https://alion.io/job/steampunk-data-engineer-databricks-2","title":"Data Engineer - Databricks","company":{"id":3854754,"name":"Steampunk","domain":"steampunk.com","url":"https://alion.io/company/steampunk-5","size_band":"1001-5000","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"iCIMS","truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":"junior","employment_type":null,"work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["McLean, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":125000,"max":160000,"currency":"USD","period":"year","gross":null,"usd_annual":160000},"salary_estimate":null,"experience_years_min":2,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Agile","optional":false},{"name":"Airflow","optional":false},{"name":"Alation","optional":false},{"name":"Amazon EC2","optional":false},{"name":"Amazon S3","optional":false},{"name":"Apache Iceberg","optional":false},{"name":"AWS","optional":false},{"name":"AWS Step Functions","optional":false},{"name":"Azure","optional":false},{"name":"C++","optional":false},{"name":"CI/CD","optional":false},{"name":"Collibra","optional":false},{"name":"Databricks","optional":false},{"name":"Delta Lake","optional":false},{"name":"ETL/ELT","optional":false},{"name":"GCP","optional":false},{"name":"Informatica","optional":false},{"name":"Java","optional":false},{"name":"MS SQL","optional":false},{"name":"MySQL","optional":false},{"name":"pySpark","optional":false},{"name":"Python","optional":false},{"name":"Scala","optional":false},{"name":"Spark","optional":false},{"name":"SQL","optional":false},{"name":"Machine Learning","optional":true}],"status":"live","first_seen_at":"2026-10-02T00:14:38Z","employer_posted_date":"2026-10-02","last_verified_at":"2026-10-04T23:21:51Z","board_verified":true,"closed_at":null,"days_open":3,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":3},"description":"Overview\nWe are looking for seasoned Data Engineer to work with our team and our clients to develop enterprise grade data platforms, services, and pipelines in Databricks. We are looking for more than just a \"Data Engineer\", but a technologist with excellent communication and customer service skills and a passion for data and problem solving.\nContributions\nLead and architect migrations of data using Databricks with focus on performance, reliability, and scalability.\nAssess and understand ETL jobs, workflows, data marts, BI tools, and reports \nAddress technical inquiries concerning customization, integration, enterprise architecture and general feature/functionality of data products \nExperience working with database/data warehouse/data mart solutions in cloud (Preferably AWS. Alternatively Azure, GCP). \nKey must have skill sets - Databricks, SQL, PySpark/Python, AWS \nSupport an Agile software development lifecycle \nYou will contribute to the growth of our AI & Data Exploitation Practice! \nQualifications\nAbility to hold a position of public trust with the US government. \n2-4 years industry experience coding commercial software and a passion for solving complex problems. \n2-4 years direct experience in Data Engineering with experience in tools such as: \nBig data tools: Databricks, Apache Spark, Delta Lake, etc. \nRelational SQL (Preferably T-SQL. Alternatively pgSQL, MySQL). \nData pipeline and workflow management tools: Databricks Workflows, Airflow, Step Functions, etc. \nAWS cloud services: Databricks on AWS, S3, EC2, RDS (or Azure equivalents). \nObject-oriented/object function scripting languages: PySpark/Python, Java, C++, Scala, etc. \nExperience working with Data Lakehouse architecture and Delta Lake/Apache Iceberg \nAdvanced working SQL knowledge and experience working with relational databases, query authoring and optimization (SQL) as well as working familiarity with a variety of databases. \nExperience manipulating, processing, and extracting value from large, disconnected datasets. \nAbility to inspect existing data pipelines, discern their purpose and functionality, and re-implement them efficiently in Databricks. \nExperience manipulating structured and unstructured data. \nExperience architecting data systems (transactional and warehouses). \nExperience the SDLC, CI/CD, and operating in dev/test/prod environments. \nExperience with data cataloging tools such as Informatica EDC, Unity Catalog, Collibra, Alation, Purview, or DataZone is a plus. \nCommitment to data governance. \nExperience working in an Agile environment. \nExperience supporting project teams of developers and data scientists who build web-based interfaces, dashboards, reports, and analytics/machine learning models \nAbout steampunk\nSteampunk relies on several factors to determine salary, including but not limited to geographic location, contractual requirements, education, knowledge, skills, competencies, and experience. The projected compensation range for this position is $125,000 to $160,000. The estimate displayed represents a typical annual salary range for this position. Annual salary is just one aspect of Steampunk’s total compensation package for employees. Learn more about additional Steampunk benefits here.\nIdentity Statement\nAs part of the application process, you are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud.\nSteampunk is a Change Agent in the Federal contracting industry, bringing new thinking to clients in the Homeland, Federal Civilian, Health and DoD sectors. Through our Human-Centered delivery methodology, we are fundamentally changing the expectations our Federal clients have for true shared accountability in solving their toughest mission challenges. If you want to learn more about our story, visit http://www.steampunk.com.\nWe are an equal opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, disability status, protected veteran status, or any other characteristic protected by law. Steampunk participates in the E-Verify program.","description_format":"text","description_chars":4186,"description_truncated":false,"requirements":{"experience_years_min":2,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-10-02T00:14:38Z"}],"visa":[],"liveness":{"score":90,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.903,"p_room":1,"age_days":2,"expected_fill_days":21,"reasons":["conf:0","velocity","win:early","comp:junior,brand"],"computed_at":"2026-10-04T05:45:00Z"},"pay":{"stated_usd_annual":160000,"is_top_pay":true},"html_url":"https://alion.io/job/steampunk-data-engineer-databricks-2","json_url":"https://alion.io/job/steampunk-data-engineer-databricks-2.json","meta":{"generated_at":"2026-10-05T00:28:42Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":533,"day_limit":5000,"remaining_today":4467,"minute_limit":60,"resets_at":"2026-10-06T00:00:00Z"}}}