Senior Network Site Reliability Engineer
פורסם לפני 20 ימים · 28 מועמדים
התפקיד במילים פשוטות
התפקיד כולל ניהול והובלה של תקריות רשת קריטיות, פתרון תקלות מורכבות בסביבות מרובות יצרנים ושיפור הניטור והאוטומציה של הרשת. העבודה מתבצעת במסגרת משמרות של 24/7 ומערבת תמיכה בפריסת אשכולות בינה מלאכותית ורשתות אלחוטיות.
- Minimum 10+ years of hands-on experience in network engineering and operations
- Deep expertise in routing, switching, firewalling, and wireless technologies across multiple vendors
- Strong troubleshooting skills with solid understanding of overlay and underlay networks
- Proficiency in Linux / Unix environments
- Hands-on experience with automation and monitoring platforms
- Strong ownership mindset and problem-solving skills
- Ability to remain calm and effective during high-severity incidents
- Excellent communication and collaboration skills
- Passion for automation, reliability engineering, and continuous improvement
חולץ מתיאור המשרה · מתעדכן אוטומטית
למי זה מתאים
תפקיד זה מתאים למהנדסי רשתות מנוסים מאוד עם לפחות 10 שנות ניסיון מעשי, השולטים בטכנולוגיות ניתוב, אבטחה ואוטומציה. הוא פחות יתאים למי שמחפש משרה ללא תורנויות או עבודה במשמרות.
תיאור המשרה המלא
המשרה המקורית · נשמר לעיוןAbout the Company
If you are interested, please share your updated resume via WhatsApp at +91 9514432939 or email it to ajai_t@hcltech.com.
About the Role
We are seeking a highly experienced Senior Network Site Reliability Engineer (SRE) to join our global network operations team. This role is critical to ensuring the reliability, scalability, and performance of enterprise and cloud-scale network infrastructure. The ideal candidate will lead incident response, troubleshoot complex network issues, drive automation initiatives, and provide strong technical leadership to deliver world-class network services.
Responsibilities
• Network Operations & Incident Management
• Lead and own critical network incidents, managing major outages and ensuring rapid service restoration.
• Provide expert guidance during high-pressure production incidents.
• Participate in 24/7 shift-based operations to ensure continuous availability of critical network services.
•
• Advanced Troubleshooting
• Diagnose and resolve complex issues across routing, switching, firewalling, and wireless domains.
• Perform deep troubleshooting involving overlay and underlay network architectures.
•
• Technical Leadership
• Set technical direction and best practices for network reliability and operations.
• Mentor junior engineers and foster a culture of operational excellence and continuous improvement.
•
• Multi-Vendor Network Engineering
• Operate across complex, multi-vendor environments including:
• Arista, Cisco, Cumulus
• Spectrum Ethernet, InfiniBand
• Palo Alto, Check Point
• Mist, Aruba
• A10, NetScaler, F5
•
• Security & Network Segmentation
• Support network segmentation, policy enforcement, and secure network designs.
• Implement and support VPN solutions such as GlobalProtect and AnyConnect.
•
• Automation & Observability
• Enhance network monitoring and automation using tools such as:
• Grafana, BigPanda, ServiceNow
• ITMP, Syslog, Splunk
• Salt, Ansible, Prometheus
•
• Drive automation to improve reliability, scalability, and operational efficiency.
•
• Innovation & Strategic Initiatives
• Collaborate on wireless network design initiatives.
• Support AI cluster deployments and next-generation network architectures.
•
Qualifications
• Minimum 10+ years of hands-on experience in network engineering and operations.
• Deep expertise in routing, switching, firewalling, and wireless technologies across multiple vendors.
• Strong troubleshooting skills with solid understanding of overlay and underlay networks.
• Proficiency in Linux / Unix environments.
• Hands-on experience with automation and monitoring platforms.
• Proven ability to work independently, set technical direction, and mentor team members.
Required Skills
• Experience with InfiniBand networks and AI cluster deployments.
• Familiarity with network asset management tools such as Nautobot.
• Wireless design and implementation experience with Cisco, Mist, and Aruba solutions.
Preferred Skills
• Strong ownership mindset and problem-solving skills.
• Ability to remain calm and effective during high-severity incidents.
• Excellent communication and collaboration skills.
• Passion for automation, reliability engineering, and continuous improvement.
שאלות על המשרה
- המשרה לא ציינה שכר. אנחנו מציגים שכר רק כשהמעסיק מפרסם אותו.
- Minimum 10+ years of hands-on experience in network engineering and operations, Deep expertise in routing, switching, firewalling, and wireless technologies across multiple vendors, Strong troubleshooting skills with solid understanding of overlay and underlay networks, Proficiency in Linux / Unix environments, Hands-on experience with automation and monitoring platforms