{"id":1054223,"url":"https://alion.io/job/dentsu-data-engineer-2","title":"Data Engineer","company":{"id":888,"name":"dentsu","domain":"dentsu.com","url":"https://alion.io/company/dentsu-polska","size_band":"5000+","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":{"grade":"B","score":80,"open_postings":116,"ghost_share":0,"stale_share":0.776,"repost_share":0.034,"time_to_fill_p50_days":24,"computed_at":"2026-10-04T05:45:00Z"}},"role":"Data Science","role_family":"Data Science","seniority":"middle","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Bengaluru, India","Pune, India","India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":12500,"max_usd":30000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":13},"experience_years_min":3,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Azure","optional":false},{"name":"Databricks","optional":false},{"name":"Power BI","optional":false},{"name":"pySpark","optional":false},{"name":"Spark","optional":false},{"name":"SQL","optional":false},{"name":"Tableau","optional":false},{"name":"Airflow","optional":true},{"name":"Azure AKS","optional":true},{"name":"CI/CD","optional":true},{"name":"Istio","optional":true},{"name":"Kubernetes","optional":true},{"name":"Microsoft Entra ID","optional":true},{"name":"Okta","optional":true},{"name":"Python","optional":true},{"name":"Service Mesh","optional":true}],"status":"live","first_seen_at":"2026-09-15T00:00:00Z","employer_posted_date":"2026-09-15","last_verified_at":"2026-10-05T01:43:53Z","board_verified":true,"closed_at":null,"days_open":20,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":20},"description":"Job Description:\nAbout the Role\nWe are looking for a Data Engineer to join our data engineering team, building and maintaining the platforms that ingest, transform, and serve data for our media and sustainability analytics products. You will work on self-serve data platforms used by internal teams and clients to connect data sources, apply business logic and taxonomies, and deliver clean, trusted data into reporting and analytics tools. This is a hands-on engineering role where you'll take ownership of well-scoped components while working closely with senior engineers on broader architectural decisions.\nWhat You'll Do\nBuild and maintain data pipelines that ingest data from third-party APIs and internal sources into cloud data lake and lakehouse environments\nDevelop data transformation logic (Spark/PySpark, SQL) to standardize, model, and enrich raw data into analytics-ready datasets\nBuild and maintain orchestration workflows to schedule, monitor, and troubleshoot data pipeline execution\nSupport data governance and access control models (e.g. Unity Catalog, ABAC-based policies) to help ensure data is secure and appropriately scoped by tenant, client, or market\nWork with product managers and senior engineers to implement platform features such as connector frameworks, taxonomy/rules engines, and data export capabilities\nSupport integration with visualization and reporting tools (e.g. Power BI, Tableau) and help ensure downstream data consumers have reliable, well-documented access\nContribute to architecture documentation (e.g. C4 model diagrams) and participate in design reviews\nTroubleshoot data quality, pipeline failures, and performance issues, tracing errors from source to destination\nWork with DevOps/security teams on service account management, credential handling, and infrastructure migrations (e.g. containerization)\nParticipate in on-call/support rotations as needed for production data pipelines\n What You'll Bring\n3+ years of experience as a Data Engineer building production-grade data pipelines\nSolid hands-on experience with Apache Spark (PySpark) and SQL for data transformation at scale\nExperience with cloud platforms (Azure preferred) and cloud-native data storage (e.g. Data Lake / Blob Storage)\nExperience with Databricks, including familiarity with Unity Catalog or similar data governance/catalog tools\nFamiliarity with data governance and access control models (RBAC/ABAC), and working with sensitive, multi-tenant data\nExperience integrating data pipelines with BI/visualization tools (Power BI, Tableau, or similar)\nComfortable working with API-based data ingestion tools/connectors (e.g. Adverity or similar ingestion platforms) is a plus\nSolid understanding of software engineering practices: version control, CI/CD, testing, code review\nGood communication skills and ability to work cross-functionally with product, engineering, and client-facing stakeholders\nNice to Have\nExperience with workflow orchestration tools such as Apache Airflow\nExperience with identity/access management integrations (Okta, Entra ID)\nExperience with service mesh technologies (Istio) and containerized deployments (AKS/Kubernetes)\nExposure to sustainability, ESG, or carbon accounting data models\nExperience with C4 model architecture documentation (PlantUML or similar\nLocation:\nDGS India - Bengaluru - Manyata N1 BlockBrand:\nMerkleTime Type:\nFull timeContract Type:\nPermanent","description_format":"text","description_chars":3414,"description_truncated":false,"requirements":{"experience_years_min":3,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"India","iso":"IN","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-09-19T01:32:48Z"}],"visa":[],"liveness":{"score":45,"band":"ok","label":"Likely open","p_open":1,"p_active":0.594,"p_room":0.75,"age_days":19,"expected_fill_days":24,"reasons":["conf:0","stale_co","wave","velocity","win:late","comp:brand"],"computed_at":"2026-10-04T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/dentsu-data-engineer-2","json_url":"https://alion.io/job/dentsu-data-engineer-2.json","meta":{"generated_at":"2026-10-05T02:58:50Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":4077,"day_limit":5000,"remaining_today":923,"minute_limit":60,"resets_at":"2026-10-06T00:00:00Z"}}}