{"id":1428239,"url":"https://alion.io/job/mcaffeine-data-engineer","title":"Data Engineer","company":{"id":2261879,"name":"MCaffeine","domain":"mcaffeine.com","url":"https://alion.io/company/mcaffeine","size_band":"501-1000","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":"middle","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Mumbai, India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":13000,"max_usd":32000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":13},"experience_years_min":4,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Amazon CloudWatch","optional":false},{"name":"Amazon EC2","optional":false},{"name":"Amazon Kinesis","optional":false},{"name":"Amazon Redshift","optional":false},{"name":"Amazon S3","optional":false},{"name":"AWS","optional":false},{"name":"AWS Glue","optional":false},{"name":"AWS Lambda","optional":false},{"name":"Confluence","optional":false},{"name":"ETL/ELT","optional":false},{"name":"Git","optional":false},{"name":"PostgreSQL","optional":false},{"name":"pySpark","optional":false},{"name":"Python","optional":false},{"name":"Scala","optional":false},{"name":"SQL","optional":false},{"name":"Spark","optional":true}],"status":"live","first_seen_at":"2026-09-28T21:53:36Z","employer_posted_date":null,"last_verified_at":"2026-09-28T21:53:36Z","board_verified":false,"closed_at":null,"days_open":11,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":11},"description":"Company overview:\nPEP is a dynamic personal care company that proudly houses two innovative brands – mCaffeine & Hyphen. With a passion for creating high-performance, conscious, and consumer-loved products, we are redefining the way personal care is experienced. While mCaffeine is India's first caffeinated personal care brand, loved for its energizing and playful approach, Hyphen is built on the philosophy of simplifying skincare through science-backed formulations. Together, our brands reflect PEP's mission to deliver quality, creativity, and care to millions of consumers. We believe in Confidence over all skin & body biases.\n\nCome, join the pack!\n\nAbout the role:\nWe are seeking a highly skilled Data Engineer with hands-on experience in AWS cloud services to join our team. The ideal candidate will be responsible for building and maintaining scalable data pipelines, integrating data systems, and ensuring data integrity and availability. You will work closely with data analysts and stakeholders to support the company's data-driven decision-making processes. Strong documentation skills are also essential to ensure that processes, pipelines, and systems are clearly defined and easily understood by both technical and non-technical stakeholders.\n\nKey Responsibilities:\nData Pipeline Development: Design, build, and maintain efficient and scalable data pipelines to process structured and unstructured data.\nAWS Cloud Services: Utilize AWS services such as S3, Lambda, CloudWatch, Kinesis, EMR, RDS and EC2 to store, process, and analyze large datasets.\nData Integration: Collaborate with internal teams to integrate and consolidate data from various sources into a centralized data warehouse or data lake. This includes integrating third-party data sources via API, webhooks and automating the process of pushing data into your systems\nETL Processes: Develop ETL (Extract, Transform, Load) processes for data integration and transformation, leveraging AWS tools like AWS Glue or custom Python/Scala scripts.\nData Warehousing: Implement and manage data warehousing solutions, with experience in AWS Redshift or other cloud-based databases.\nAutomation & Optimization: Automate data processing tasks and optimize the performance of data workflows.\nCollaboration: Work closely with data analysts and stakeholders to ensure that data solutions meet business requirements.\nMonitoring & Maintenance: Monitor data pipelines and ensure data availability, quality, and performance. Troubleshoot issues as needed.\nDocumentation: Document data engineering processes, systems, pipelines, and workflows. Ensure all code, processes, and procedures are clearly explained and accessible for future reference or team members.\nBest Practices: Maintain high standards for code quality and documentation to ensure ease of collaboration, support, and long-term maintainability of data systems.\n\nQualifications and Skills:\nEducation: Bachelor's degree in Computer Science, Engineering, Mathematics, or a related field (or equivalent work experience).\nAWS Certifications (e.g. AWS Certified Data Engineer)\nExperience: 4+ years of experience in a data engineering role with a focus on AWS cloud services, infrastructure management, and data warehousing solutions\nExperience in integrating third-party systems and services via webhooks for data ingestion.\nExperience in developing ETL processes and working with large-scale datasets.\nStrong experience with SQL and database technologies, particularly in relational databases.\n2+ years of experience in handling Postgres Database.\nExperience with Python, PySpark, Scala, or other programming languages for data engineering tasks.\nProven ability to write clear, comprehensive documentation for complex data engineering systems and workflows.\nVersion Control: Knowledge of Git or other version control systems\nProblem-Solving: Strong analytical and problem-solving skills.\nDocumentation Tools: Familiarity with documentation tools like Confluence, Markdown, or Jupyter Notebooks for writing clear, organized technical documentation.\nExcellent communication skills to collaborate with cross-functional teams.","description_format":"text","description_chars":4143,"description_truncated":false,"requirements":{"experience_years_min":4,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Consumer Goods","Personal Hygiene & Toiletries","Personal Care"],"lifecycle":[{"event":"open","at":"2026-09-29T00:08:44Z"}],"visa":[],"liveness":{"score":81,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.815,"p_room":1,"age_days":10,"expected_fill_days":30,"reasons":["seen:10","win:early"],"computed_at":"2026-10-09T06:01:00Z"},"pay":null,"html_url":"https://alion.io/job/mcaffeine-data-engineer","json_url":"https://alion.io/job/mcaffeine-data-engineer.json","meta":{"generated_at":"2026-10-10T01:07:24Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1874,"day_limit":5000,"remaining_today":3126,"minute_limit":60,"resets_at":"2026-10-11T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":2261879},"rest":"https://alion.io/mcp/rest/get_company?id=2261879"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fmcaffeine-data-engineer"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fmcaffeine-data-engineer"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fmcaffeine-data-engineer"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/mcaffeine-data-engineer\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fmcaffeine-data-engineer"}]}