{"id":1226744,"url":"https://alion.io/job/myworkdaysite-lead-software-engineer","title":"Lead Software Engineer","company":{"id":2691756,"name":"Myworkdaysite","domain":"myworkdaysite.com","url":"https://alion.io/company/myworkdaysite","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"Backend","role_family":"Backend","seniority":"lead","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Bengaluru, India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":35000,"max_usd":75000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":8},"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AIOps","optional":false},{"name":"Anomaly Detection","optional":false},{"name":"Ansible","optional":false},{"name":"CI/CD","optional":false},{"name":"Commander.js","optional":false},{"name":"Grafana","optional":false},{"name":"Incident Management","optional":false},{"name":"Kubernetes","optional":false},{"name":"Linux","optional":false},{"name":"OpenShift","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Prometheus","optional":false},{"name":"Python","optional":false},{"name":"Self-Healing","optional":false},{"name":"SLI/SLO/SLA","optional":false},{"name":"Splunk","optional":false},{"name":"Terraform","optional":false},{"name":"JavaScript","optional":true},{"name":"Node JS","optional":true}],"status":"live","first_seen_at":"2026-09-25T02:36:38Z","employer_posted_date":null,"last_verified_at":"2026-09-25T02:36:38Z","board_verified":false,"closed_at":null,"days_open":2,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":2},"description":"About this role:\nWells Fargo is seeking a Lead Software Engineer.\n\nIn this role, you will:\nLead complex technology initiatives including those that are companywide with broad impact\nAct as a key participant in developing standards and companywide best practices for engineering complex and large scale technology solutions for technology engineering disciplines\nDesign, code, test, debug, and document for projects and programs\nReview and analyze complex, large-scale technology solutions for tactical and strategic business objectives, enterprise technological environment, and technical challenges that require in-depth evaluation of multiple factors, including intangibles or unprecedented technical factors\nMake decisions in developing standard and companywide best practices for engineering and technology solutions requiring understanding of industry best practices and new technologies, influencing and leading technology team to meet deliverables and drive new initiatives\nCollaborate and consult with key technical experts, senior technology team, and external industry groups to resolve complex technical issues and achieve goals\nLead projects, teams, or serve as a peer mentor\n\nRequired Qualifications:\n5+ years of Software Engineering experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education\n\nDesired Qualification:\nExperience in Software Engineering, SRE, DevOps, or Platform Engineering.\nStrong proficiency in Python for automation and tooling.\nHandson experience with Grafana, Prometheus, and Splunk in production environments.\nSolid understanding of SLIs, SLOs, dashboards, alerting, and observability best practices.\nExperience applying AI/ML concepts to monitoring, alerting, or operational analytics.\nStrong knowledge of Linux, networking, and distributed systems.\nExperience with Cloud platforms and Kubernetes/OpenShift.\nProven experience leading incidents, RCAs, and reliability initiatives\nExperience building custom Prometheus exporters or advanced Grafana dashboards.\nStrong Splunk expertise (search, dashboards, alerts, log pipelines).\nExperience operationalizing ML models for observability (AIOps).\nFamiliarity with CI/CD, Terraform, Ansible, and enterprise automation platforms.\nExperience supporting largescale, regulated, or globally distributed systems.\nImproved reliability and performance against defined SLOs.\nReduced alert noise and faster detection and recovery of incidents.\nIncreased automation and selfhealing adoption using Python and observability signals.\nStrong observability maturity across platforms and applications.\nImproved MTTD and MTTR through effective use of Grafana, Prometheus, and Splunk.\n\nJob Expectation:\n\nReliability & Availability Engineering\nOwn and improve availability, performance, scalability, and resilience of production systems.\nDefine, monitor, and manage SLIs/SLOs and error budgets to guide reliability investments.\nLead capacity planning, performance testing, failover readiness, and disasterrecovery design.\n\nObservability & Monitoring (Grafana / Prometheus / Splunk)\nDesign and operate a comprehensive observability stack using:Prometheus for metrics collection and alerting\nGrafana for dashboards, visualization, and SLO tracking\nSplunk for log aggregation, troubleshooting, and incident forensics\n\nBuild and maintain golden dashboards and actionable alerts aligned to business impact.\nReduce alert fatigue through signalbased monitoring and correlation of metrics, logs, and traces.\nPartner with application teams to define instrumentation standards for metrics and logging.\nUse observability data to improve MTTD, MTTR, and reliability outcomes.\n\nAutomation & Python Engineering\nDevelop Pythonbased automation for monitoring, alert remediation, deployments, scaling, and recovery.\nBuild selfhealing workflows integrated with Prometheus alerts and Splunk signals.\nCreate reusable automation frameworks and internal SRE tooling.\nEmbed automation into CI/CD pipelines to improve deployment safety and reliability.\n\nAI/MLDriven Reliability (AIOps)\nApply AI/ML techniques to observability and operations use cases, including:Anomaly detection on Prometheus metrics\nLog pattern analysis and correlation in Splunk\nPredictive capacity and trend forecasting\nNoise reduction and intelligent alerting\n\nPartner with data and platform teams to operationalize ML models in production.\nEvaluate and integrate AIOps capabilities into the observability ecosystem.\n\nIncident Management & RCA\nServe as incident commander and senior escalation point for P1/P2 incidents.\nLead blameless postincident reviews (PIRs) backed by Grafana metrics and Splunk evidence.\nDrive corrective and preventive actions to completion.\n\nPlatform & Application Partnership\nCollaborate with platform, application, cloud, and SRE teams to embed reliability and observability by design.\nInfluence architectural decisions to ensure systems are observable, scalable, and operable.\nProvide SRE guidance during major releases, migrations, and modernization initiatives.\n\nSecurity, Risk & Compliance\nEnsure observability and automation comply with enterprise security and audit requirements.\nSupport resilience validation, failover drills, and business continuity testing.\n\nTechnical Leadership\nMentor and guide SRE and software engineers.\nDefine standards for observability, automation, reliability, and incident response.\nAct as the technical authority for complex production and platform issues.\n\nPosting End Date:\n\n30 Sep 2026\n\nWe Value Equal Opportunity\n\nWells Fargo is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other legally protected characteristic.\n\nEmployees support our focus on building strong customer relationships balanced with a strong risk mitigating and compliance-driven culture which firmly establishes those disciplines as critical to the success of our customers and company. They are accountable for execution of all applicable risk programs (Credit, Market, Financial Crimes, Operational, Regulatory Compliance), which includes effectively following and adhering to applicable Wells Fargo policies and procedures, appropriately fulfilling risk and compliance obligations, timely and effective escalation and remediation of issues, and making sound risk decisions. There is emphasis on proactive monitoring, governance, risk identification and escalation, as well as making sound risk decisions commensurate with the business unit's risk appetite and all risk and compliance program requirements.\n\nCandidates applying to job openings posted in Canada: Applications for employment are encouraged from all qualified candidates, including women, persons with disabilities, aboriginal peoples and visible minorities. Accommodation for applicants with disabilities is available upon request in connection with the recruitment process.\n\nApplicants with Disabilities\n\nTo request a medical accommodation during the application or interview process, visit.\n\nDrug and Alcohol Policy\n\nWells Fargo maintains a drug free workplace. Please see our to learn more.\n\nWells Fargo Recruitment and Hiring Requirements:\n\na. Third-Party recordings are prohibited unless authorized by Wells Fargo.\nb. Wells Fargo requires you to directly represent your own experiences during the recruiting and hiring process.","description_format":"text","description_chars":7530,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-09-25T13:06:44Z"}],"liveness":{"score":90,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.903,"p_room":1,"age_days":2,"expected_fill_days":24,"reasons":["seen:2","velocity","win:early"],"computed_at":"2026-09-27T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/myworkdaysite-lead-software-engineer","json_url":"https://alion.io/job/myworkdaysite-lead-software-engineer.json","meta":{"generated_at":"2026-09-28T01:48:39Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":972,"day_limit":5000,"remaining_today":4028,"minute_limit":60,"resets_at":"2026-09-29T00:00:00Z"}}}