{"id":1767320,"url":"https://alion.io/job/ntt-site-reliability-engineer","title":"Site Reliability Engineer","company":{"id":121,"name":"NTT","domain":"global.ntt","url":"https://alion.io/company/ntt","size_band":"5000+","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":null,"employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["United States"],"countries":["US"],"hiring_countries":["US"],"hiring_countries_total":1,"salary":{"min":80000,"max":114000,"currency":"USD","period":"year","gross":null,"usd_annual":114000},"salary_estimate":null,"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Agile","optional":false},{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"Bash","optional":false},{"name":"CI/CD","optional":false},{"name":"CircleCI","optional":false},{"name":"CloudFormation","optional":false},{"name":"Docker","optional":false},{"name":"GCP","optional":false},{"name":"GitLab CI","optional":false},{"name":"Grafana","optional":false},{"name":"Incident Management","optional":false},{"name":"Java","optional":false},{"name":"Jenkins","optional":false},{"name":"Kubernetes","optional":false},{"name":"Linux","optional":false},{"name":"New Relic","optional":false},{"name":"PowerShell","optional":false},{"name":"Prometheus","optional":false},{"name":"Python","optional":false},{"name":"Ruby","optional":false},{"name":"Self-Healing","optional":false},{"name":"Terraform","optional":false},{"name":"Unix","optional":false}],"status":"live","first_seen_at":"2026-10-01T00:00:00Z","employer_posted_date":"2026-10-01","last_verified_at":"2026-10-11T16:04:13Z","board_verified":true,"closed_at":null,"days_open":10,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":10},"description":"Make an impact with NTT DATA\nJoin a company that is pushing the boundaries of what is possible. We are renowned for our technical excellence and leading innovations, and for making a difference to our clients and society. Our workplace embraces diversity and inclusion - it’s a place where you can grow, belong and thrive.\nYour day at NTT DATA\nThe Site Reliability Engineer (SRE) is a seasoned subject matter expert, responsible for ensuring the reliability, availability, and performance of company systems and infrastructure.\nThis Site Reliability Engineer (SRE) works closely with development teams, operations teams, and other stakeholders to enhance system resiliency, automate processes, and improve overall system reliability.\nKey responsibilities:\nMonitors system health, performance metrics, and alerts to identify and respond to incidents promptly and diagnoses issues, troubleshoots problems, and restores services in a timely manner.\nImplements incident response processes to minimize downtime and improve system availability.\nDesigns, develops, and maintains automation tools, scripts, and processes to streamline system management tasks, deployments, and configuration changes.\nImplements infrastructure-as-code principles to ensure consistency and repeatability.\nOptimizes system resources, configurations, and processes to enhance performance, scalability, and efficiency.\nUses monitoring tools and performance testing to identify bottlenecks and implement optimizations.\nCollaborates with teams to forecast system resource needs, plans for capacity growth, and ensures adequate scalability.\nLeads incident response efforts, coordinates with cross-functional teams, and drives the resolution of system issues.\nPerforms thorough post-incident analysis to identify root causes and implements preventive measures to minimize future incidents.\nIdentifies opportunities for automation and drives the implementation of self-healing, monitoring, and deployment of automation tools and frameworks.\nContinuously improves operational efficiency, system reliability, and availability through process enhancements and automation.\nEnsures consistency across environments, tracks changes, and enforces configuration standards.\nWorks closely with development teams, operations teams, and other stakeholders to ensure effective collaboration, knowledge sharing, and alignment on reliability goals.\nImplements security best practices, works with security teams to assess and address vulnerabilities, and ensures compliance with security standards and regulations.\nPerforms any other related task as required.\nTo thrive in this role, you need to have:\nSeasoned technical expertise in Linux/Unix systems, networking, and system administration.\nSeasoned proficiency in scripting or programming languages, such as Python, Go, Java, or Ruby.\nSeasoned knowledge of cloud platforms (such as AWS, Azure, or Google Cloud) and associated services.\nSeasoned proven expertise in performance monitoring, optimization, and troubleshooting using tools such as Prometheus, Grafana, or New Relic.\nSeasoned expertise in incident management, root cause analysis, and post-incident reviews\nExcellent problem-solving and analytical skills, with a keen attention to detail.\nExcellent communication, collaboration, and leadership skills.\nSeasoned ability to optimize system performance, scalability, and reliability. experience with performance monitoring and tuning tools (for example, Prometheus, Grafana, or New Relic) to identify bottlenecks, analyze performance data, and implement optimization strategies.\nSeasoned understanding of security principles, best practices, and compliance requirements. experience in designing and implementing security controls, performing security assessments, and ensuring compliance with industry standards.\nAcademic qualifications and certifications:\nBachelor's degree or equivalent in Computer Science, Information Technology, or a related field.\nRelevant certifications, such as AWS Certified DevOps Engineer - Professional, Google Cloud Professional DevOps Engineer, or Certified Kubernetes Administrator (CKA) preferred.\nRequired experience:\nSeasoned hands-on experience in a Site Reliability Engineering role or related roles, including experience in designing and maintaining highly available and scalable systems.\nSeasoned hands-on experience with Linux/Unix systems, networking, and system administration is crucial. In-depth knowledge of cloud platforms (such as AWS, Azure, or Google Cloud) and associated services is essential.\nSeasoned proficiency in multiple programming languages like Python, Java, Go, or Ruby is important for developing and maintaining automation tools, frameworks, and complex system integrations. Expertise in scripting languages like Bash or PowerShell is beneficial.\nSeasoned understanding of complex infrastructure architectures, including scalable and fault-tolerant designs. experience with infrastructure-as-code tools (such as Terraform or CloudFormation) and containerization technologies (such as Docker or Kubernetes) is essential.\nSeasoned experience in designing and implementing robust automation frameworks, CI/CD pipelines, and deployment strategies. Proficiency in tools like Jenkins, GitLab CI/CD, or CircleCI to build, test, and deploy applications with a focus on reliability and scalability.\nSeasoned experience in incident management, troubleshooting complex system issues, and conducting post-incident analysis. Advanced ability to lead incident response efforts, drive root cause analysis, and implement preventive measures.\nSeasoned understanding of DevOps principles, Agile methodologies, and a strong commitment to continuous improvement and learning. experience in promoting a DevOps culture and driving the adoption of best practices\nNTT DATA provides a reasonable range of compensation for U.S.-based positions. The starting pay range for this remote role is $80,000.00 - 114,000.00 - 148,000.00 USD Annual. This range reflects the minimum and maximum target compensation for the position across all US locations. Actual compensation will depend on a number of factors, including the candidate’s actual work location, relevant experience, technical skills, and other qualifications.\nWorkplace type:\nRemote WorkingAbout NTT DATA\nNTT DATA is a $30+ billion business and technology services leader, serving 75% of the Fortune\nGlobal 100. We are committed to accelerating client success and positively impacting society through\nresponsible innovation. We are one of the world’s leading AI and digital infrastructure providers, with\nunmatched capabilities in enterprise-scale AI, cloud, security, connectivity, data centers and\napplication services. Our consulting and industry solutions help organizations and society move\nconfidently and sustainably into the digital future. As a Global Top Employer, we have experts in more\nthan 70 countries. We also offer clients access to a robust ecosystem of innovation centers as well as\nestablished and start-up partners. NTT DATA is part of NTT Group, which invests over $3 billion each\nyear in R&D.\nEqual Opportunity Employer\nNTT DATA is proud to be an Equal Opportunity Employer with a global culture that embraces diversity. We are committed to providing an environment free of unfair discrimination and harassment. We do not discriminate based on age, race, colour, gender, sexual orientation, religion, nationality, disability, pregnancy, marital status, veteran status, or any other protected category. Join our growing global team and accelerate your career with us. Apply today.\nThird parties fraudulently posing as NTT DATA recruiters\nNTT DATA recruiters will never ask job seekers or candidates for payment or banking information during the recruitment process, for any reason. Please remain vigilant of third parties who may attempt to impersonate NTT DATA recruiters whether in writing or by phone in order to deceptively obtain personal data or money from you. All email communications from an NTT DATA recruiter will come from an @nttdata.com email address. If you suspect any fraudulent activity, please contact us.","description_format":"text","description_chars":8148,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"bachelor","optional":false},"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Artificial Intelligence","Information Technology","Internet Services","Systems Integrators"],"lifecycle":[{"event":"open","at":"2026-10-03T14:17:58Z"}],"visa":[],"liveness":{"score":68,"band":"ok","label":"Likely open","p_open":1,"p_active":0.752,"p_room":0.9,"age_days":9,"expected_fill_days":14,"reasons":["conf:1","velocity","win:mid","comp:brand"],"computed_at":"2026-10-10T05:45:15Z"},"pay":{"stated_usd_annual":114000,"is_top_pay":false},"html_url":"https://alion.io/job/ntt-site-reliability-engineer","json_url":"https://alion.io/job/ntt-site-reliability-engineer.json","meta":{"generated_at":"2026-10-11T19:12:19Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler_verified","counted_by":"address","units_charged":1,"used_today":5819,"day_limit":null,"remaining_today":null,"minute_limit":300,"resets_at":"2026-10-12T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":121},"rest":"https://alion.io/mcp/rest/get_company?id=121"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fntt-site-reliability-engineer"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fntt-site-reliability-engineer"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fntt-site-reliability-engineer"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/ntt-site-reliability-engineer\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fntt-site-reliability-engineer"}]}