{"id":2064856,"url":"https://alion.io/job/cyberproof-data-engineer-3","title":"Data Engineer","company":{"id":1884346,"name":"CyberProof","domain":"cyberproof.com","url":"https://alion.io/company/cyberproof","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":null,"employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Bentonville, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":92000,"max":138000,"currency":"USD","period":"year","gross":null,"usd_annual":138000},"salary_estimate":null,"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Airflow","optional":false},{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"Azure Data Factory","optional":false},{"name":"BigQuery","optional":false},{"name":"Databricks","optional":false},{"name":"ETL/ELT","optional":false},{"name":"GCP","optional":false},{"name":"Google BigQuery","optional":false},{"name":"Java","optional":false},{"name":"Python","optional":false},{"name":"Spark","optional":false},{"name":"SQL","optional":false},{"name":"Apache Kafka","optional":true}],"status":"closed","first_seen_at":"2026-10-08T03:11:22Z","employer_posted_date":null,"last_verified_at":"2026-10-08T09:21:30Z","board_verified":false,"closed_at":"2026-10-08T09:21:30Z","days_open":0,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":0},"description":"UST is a mission-driven technology company that creates transformative experiences and human-centered solutions for clients and partners. The Lead II - Data Engineering role focuses on designing, developing, and supporting scalable data pipelines and enterprise data platforms using GCP, Airflow, Java/Spark, and SQL. The position also involves optimizing data processing, ensuring data quality and governance, and collaborating with analysts, scientists, and application teams.\n\nResponsibilities\nDevelop reusable data processing frameworks and components ensuring data quality, consistency, and reliability across data platforms. o high-volume data integration and migration initiatives\nOptimize Spark jobs for performance, scalability, and cost efficiency\nCode reviews, testing, deployment, and production support activities of monitoring & troubleshooting airflow workflow and production pipelines\nShould have experience in query performance through indexing, portioning and data model optimization\nDesign, develop, and maintain scalable ETL/ELT pipelines to ingest, transform, and load data from multiple sources into data warehouses and data lakes\nBuild and optimize data solutions using SQL, Python, Spark, and cloud technologies such as Azure, AWS, or GCP\nDevelop and manage data integration workflows using tools like Azure Data Factory, Databricks, Airflow, or Informatica\nEnsure data quality, integrity, security, and governance through validation, monitoring, and compliance best practices\nCollaborate with business stakeholders, data analysts, and data scientists to deliver reliable, analytics-ready datasets\nMonitor, troubleshoot, and optimize data pipelines and database performance to support high-volume data processing and reporting requirements\nThe role is expected to Work closely with business analysts, data scientists, and application teams across the globe with a mix of associates and suppliers\n\nSkills\nUST is looking for a skilled Data Engineer with 6+ years of experience in designing, developing, and supporting scalable data pipelines and enterprise data platforms\nThe ideal candidate will have strong expertise in Google Cloud Platform (GCP), Apache Airflow DAGs, Java/Spark development, and SQL-based data engineering to support large-scale analytics and data transformation initiatives\nExperience range: 7-10 Years\nMandatory: BigQuery, Apache Airflow, Spark/Java/ PySpark\nShould have 6+ years of experience in Data Engineering with demonstrated hands on experience of 4+ years of experience in Design, develop, and maintain scalable data pipelines on Google Cloud Platform (GCP 3+ years of recent experience in Develop, schedule, and manage workflows using Apache Airflow DAGs\n3+ years of experience in Build and optimize batch and streaming data processing solutions using Java and Apache Spark\n4+ years of experience with writing complex and optimized SQL queries for data extraction, transformation and reporting\nShould have Implement ETL/ELT processes for ingesting, transforming, and loading large-scale datasets\nNice to have: GCP, GCS, Relational Databases Data Modeling, Kafka , Terraform\n\nBenefits\nFull-time, regular employees accrue a minimum of 10 days of paid vacation per year.\nFull-time, regular employees receive 6 days of paid sick leave each year (pro-rated for new hires throughout the year).\nFull-time, regular employees receive 10 paid holidays.\nFull-time, regular employees are eligible for paid bereavement leave and jury duty.\nFull-time, regular employees are eligible to participate in the Company’s 401(k) Retirement Plan with employer matching.\nFull-time, regular employees and their dependents residing in the US are eligible for medical, dental, and vision insurance.\nFull-time, regular employees receive Company-paid basic life insurance, accidental death and disability insurance, and short- and long-term disability benefits for employees only.\nRegular employees may purchase additional voluntary short-term disability benefits.\nRegular employees may participate in a Health Savings Account (HSA) and a Flexible Spending Account (FSA) for healthcare, dependent child care, and/or commuting expenses as allowable under IRS guidelines.\nPart-time employees receive 6 days of paid sick leave each year (pro-rated for new hires throughout the year).\nPart-time employees are eligible to participate in the Company’s 401(k) Retirement Plan with employer matching.\nFull-time temporary employees receive 6 days of paid sick leave each year (pro-rated for new hires throughout the year).\nFull-time temporary employees are eligible to participate in the Company’s 401(k) program with employer matching.\nFull-time temporary employees and their dependents residing in the US are eligible for medical, dental, and vision insurance.\nPart-time temporary employees receive 6 days of paid sick leave each year (pro-rated for new hires throughout the year).\nAll US employees who work in a state or locality with more generous paid sick leave benefits than specified here will receive the benefit of those sick leave laws.\n\nCompany Overview\nCyberProof is a security services company that intelligently manages your incident detection and response requirements. It is a sub-organization of UST. It was founded in 2017, and is headquartered in Aliso Viejo, California, USA, with a workforce of 501-1000 employees. Its website is https://www.cyberproof.com.","description_format":"text","description_chars":5396,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Life insurance","Retirement plans","Vision insurance"],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-10-08T03:20:00Z"},{"event":"close","at":"2026-10-08T09:21:30Z"}],"visa":[],"liveness":null,"pay":{"stated_usd_annual":138000,"is_top_pay":false},"html_url":"https://alion.io/job/cyberproof-data-engineer-3","json_url":"https://alion.io/job/cyberproof-data-engineer-3.json","meta":{"generated_at":"2026-10-08T18:05:47Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":222,"day_limit":5000,"remaining_today":4778,"minute_limit":60,"resets_at":"2026-10-09T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":1884346},"rest":"https://alion.io/mcp/rest/get_company?id=1884346"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fcyberproof-data-engineer-3"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fcyberproof-data-engineer-3"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fcyberproof-data-engineer-3"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/cyberproof-data-engineer-3\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fcyberproof-data-engineer-3"}]}