Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Wells Fargo

Principal Network Engineer – Reliability

Wells Fargo

Principal Network Engineer driving reliability modernization and improvement for enterprise scale services. Collaborating across networks, cloud, and applications for business-critical operations.

Posted 8/1/2026full-timeIrving • Arizona, North Carolina, Texas • 🇺🇸 United StatesLeadWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates extensive experience in network engineering and reliability engineering, focusing on critical enterprise network services. Proficient in incident response, service health measurement, and implementing continuous improvement practices to enhance operational stability.

Highest-signal resume keywords
Network EngineeringReliability EngineeringIncident Response LeadershipSLOs/SLIs ImplementationAutomation Scripting

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
RoutingSwitchingTCP/IPBGPOSPFHigh-Availability DesignLoad BalancingNetwork SecurityCapacity PlanningPerformance Analysis
Soft Skills
Clear Communication Under PressureCross-Functional InfluenceMentoring Engineers
Tools & Technologies
PythonPowerShellBashAnsibleTerraform
Industry Keywords
Operational Readiness ReviewsService Health ReportingRoot Cause AnalysisPost-Incident ReviewsChange Management

Tech Stack

Tools & technologies
AnsibleCloudDNSFirewallsPythonSwitchingTCP/IPTerraform

About the role

Key responsibilities & impact
  • Help shape the reliability strategy for critical network services at enterprise scale
  • Influence how network reliability is measured, engineered, automated, and continuously improved
  • Drive modernization, reduce operational risk, and improve the stability of services that support critical business operations
  • Provide senior technical leadership, hands-on engineering guidance, and cross-functional influence
  • Define and drive reliability targets, SLOs/SLIs, risk measures, service health indicators, and improvement plans for critical network services
  • Lead major incident response and service restoration, and drive root cause analysis, corrective action planning, and prevention of repeat incidents
  • Review, challenge, and guide network designs to improve resiliency, scalability, security, capacity, failure isolation, and operational simplicity
  • Expand telemetry, alerting, service health reporting, configuration management, automated validation, and self-service capabilities to reduce manual toil
  • Strengthen standards, runbooks, documentation, change quality, failure readiness, compliance-aligned practices, and production readiness reviews
  • Partner with network, cloud, security, infrastructure, application, architecture, and vendor teams while mentoring engineers and communicating complex technical issues clearly to senior stakeholders

Requirements

What you’ll need
  • 7+ years of experience in network engineering, reliability engineering, or network operations supporting business-critical enterprise network services
  • 5+ years of expert knowledge of routing, switching, TCP/IP, BGP, OSPF, high-availability design, and related network services including load balancing, DNS, NTP, firewalls, and network security
  • Proven ability to lead complex incident response, guide troubleshooting, restore service, and communicate clearly under pressure
  • Practical experience applying SRE or reliability engineering practices, including SLOs/SLIs, problem management, service health measurement, operational health metrics, and continuous improvement
  • Experience improving operational stability and reducing repeat incidents through root cause analysis, post-incident reviews, corrective action planning, systemic remediation, and measurable recurrence prevention
  • Experience owning corrective actions from post-incident review through implementation, validation, and closure
  • Experience improving change success rates through risk assessment, peer review, validation testing, rollback planning, and post-change verification
  • Experience conducting operational readiness reviews, production readiness assessments, failure-mode reviews, or service acceptance reviews
  • Experience managing vendor escalations, product defects, support cases, and platform lifecycle risks that impact service stability
  • Experience with capacity planning, performance analysis, traffic engineering, and resiliency planning for large-scale enterprise networks
  • Hands-on automation or scripting experience with tools such as Python, PowerShell, Bash, Ansible, Terraform, or equivalent technologies
  • Experience improving monitoring, telemetry, alerting, dashboards, metrics, service health reporting, or performance visibility for critical infrastructure services

Benefits

Comp & perks
  • Hybrid work model with three days per week in the office
  • Provides senior technical escalation for critical incidents and high-risk changes
  • Support critical incident response as needed; any recurring on-call rotation will be clearly defined before offer acceptance
  • 5% or less travel expectations