Infrastructure Solution Engineer- Sr Consultant I
ALLSTATE CORP · Remote
📍 USA - TX (Remote)via workday
Apply on company site ↗
CareerRiver pulls this listing straight from the employer's hiring system — no recruiter middleman, no reposts. Applying takes you directly to ALLSTATE CORP.
At Allstate, great things happen when our people work together to protect families and their belongings from life’s uncertainties. And for more than 90 years, our innovative drive has kept us a step ahead of our customers’ evolving needs. From advocating for seat belts, air bags and graduated driving laws, to being an industry leader in pricing sophistication, telematics, and, more recently, device and identity protection.
Job Description
The Infrastructure Solutions Engineer- Sr Consultant I for NOC Operations is an experienced technical professional in our Standard Incident service team, leading the monitoring, troubleshooting, and resolution of infrastructure incidents across our enterprise technology ecosystem. As a strong contributor to our “Zero Wait” customer obsession initiative, this role drives rapid and effective response to system alerts, ensuring the reliability and performance of Allstate’s critical infrastructure. With demonstrated proficiency in UNIX, Storage, and Backup environments, this position also demonstrates broad versatility across Windows, Nutanix, Azure, and AWS platforms.
Working within our product-centric operating model, the Infrastructure Solutions Engineer- Sr Consultant I applies deep technical knowledge and operational expertise to collaborate with Digital Product Teams (DPTs) and other Service Teams to optimize service delivery, eliminate friction points, and design automation solutions that enhance system reliability and customer experience.
Key Responsibilities
Service Excellence Lead by example in proactive monitoring and rapid response to alerts across multiple technology stacks, consistently achieving strong Mean Time to Acknowledge (MTTA) and Mean Time to Resolve (MTTR) metrics
Champion “Zero Wait” principles in incident response, taking immediate and decisive action on Emergency Command Center (ECC) calls without waiting to be prompted
Critically evaluate and improve Standard Operating Procedures (SOPs) while identifying strategic opportunities to enhance processes and automate complex tasks
Perform shift turnover meetings to ensure seamless 24/7 operational coverage and comprehensive knowledge transfer
Actively contribute to and help prioritize the Service Improvement Backlog (SIB) with strategic ideas that substantially enhance service delivery and eliminate customer friction points
Technical Operations Provide Level 2 support for enterprise infrastructure systems including Storage SAN, NAS and OBS (Brocade, Hitachi, Pure, Scality, Cisco MDS ), Linux (RedHat), with minimal escalation required for common and moderately complex issues
Execute incident remediation strategies for both documented and novel problems, applying advanced technical judgment in complex or ambiguous scenarios
Perform post-incident Retrospective reviews and problem management activities, driving root cause analysis and implementing preventive measures against recurring issues
Partner closely with engineering teams to implement and validate system changes, providing strong operational perspective and ensuring readiness for production
Apply intermediate level techniques with monitoring tools including Netcool, Tivoli, Prism Element / Central, Datadog, and Azure Data Explorer (ADX) to proactively identify, diagnose, and prevent system issues
Continuous Improvement Analyze complex incident patterns and trends to architect automation solutions that reduce manual intervention and improve service outcomes
Develop comprehensive knowledge base strategies and ensure knowledge artifacts meet high quality standards for accuracy and usability across the NOC team
Present at service reviews and demo sessions with technology partners, showcasing NOC operational improvements and contributions
Drive the team’s KPI improvements through technical expertise, process optimization, and a continuous improvement mindset
Participate and contribute to the NOC automation pipeline by identifying, designing, and advocating for automation solutions that improve service delivery
Qualifications 3+ years of experience in technical operations, IT support, or system administration
Intermediate working knowledge of enterprise systems Storage SAN, NAS and OBS (Brocade, Hitachi, Pure, Scality, Cisco MDS), (UNIX/Linux (RedHat), with demonstrated ability to troubleshoot complex issues
Strong understanding of incident management processes, incident / event management frameworks, and service delivery optimization
Strong troubleshooting and analytical skills with demonstrated ability to resolve complex and novel technical challenges
Strong communication abilities, particularly during high-pressure situations and when explaining technical concepts to non-technical stakeholders
Preferred Proficient use and knowledge of automation concepts and tools (GitHub, Ansible, Jenkins, CI/CD pipelines)
Proficient scripting skills (Bash, Python, PowerShell) with the ability to develop and contribute to automation solutions
Strong skills with monitoring tools (Datadog, Azure Data Explorer (ADX)) including custom dashboard usage and alert tuning
Familiarity with AI-assisted tools for incident triage, root cause analysis, or automated remediation workflows
Familiarity with large language model (LLM) tools (e.g., GitHub Copilot, Microsoft Copilot, or similar) to accelerate documentation, scripting, and knowledge base development
Working knowledge of DevOps practices and principles with experience implementing infrastructure as code
Leadership experience in a 24/7 operational environment
Growth Path This position offers clear pathways to advance into senior technical or leadership roles, including:
Senior Consultant II roles with deeper technical specialization and formal mentoring responsibilities
Service Consultant roles focusing on complex technical solutions and automation
Technical leadership in specialized infrastructure domains
Service management roles with
More Remote jobs
Remote jobs · Browse all locations