☁️cloudjobboard
← All roles
Vultr logo

Vultr

Senior Network Engineer

📍 ChennaiOn-siteFull-timePosted today

Apply now →

Who We Are

Vultr is on a mission to make high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators around the world. With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal, and Cloud Storage solutions. In December 2024 Vultr announced an equity financing at a $3.5 billion valuation. Founded by David Aninowsky and self-funded for over a decade, Vultr has grown to become the world’s largest privately-held cloud infrastructure company.

Vultr Cares

  • Medical Insurance stipend paid annually
  • 9 Company-Paid Holidays
  • Generous Leave Policy + 1 month paid sabbatical every 5 years + Anniversary Bonus each year
  • Professional Development Reimbursement
  • Internet reimbursement
  • Fitness membership reimbursement
  • Company paid Wellable subscription

Join Vultr

Vultr is expanding its India presence and is building its first Global Integrated Operations Command Center (GIOC) in Chennai — the 24x7 nerve center for monitoring, triage, and incident resolution across Vultr’s global network operations.

We are seeking Senior Network Engineers to own escalated and complex incidents across Vultr’s network estate — routing, switching, DNS, CDN, load balancing, and connectivity. This is a deep-dive engineering role for experienced network operators who lead troubleshooting, drive root-cause analysis, author runbooks, and mentor Network Engineers — and who want to grow toward network architecture, SRE, or platform engineering. You must be comfortable with a rotational shift and on-call model — including nights, weekends, and holidays — to sustain follow-the-sun coverage.

Role Overview

The Senior Network Engineer owns incidents escalated from Network Engineers and leads diagnosis on the most complex network issues. You perform deep troubleshooting, drive incidents to resolution within defined SLAs, determine root cause, and deliver permanent fixes that prevent recurrence. The role ensures deep technical resolution and network reliability that protects customer experience and network availability.

You act as a senior escalation point and network subject-matter expert, leading major-incident bridges, partnering with Engineering SMEs and providers on architecture-level fixes, and raising the team’s capability through runbooks and mentoring. Success is measured by MTTR, root-cause closure, recurrence reduction, and runbook coverage. The role runs on a rotational shift and on-call schedule to sustain 24x7coverage.

Key Responsibilities

Escalated Incident Ownership

  • Own incidents escalated from Network Engineers across routing, switching, DNS, CDN, and connectivity, and drive them to resolution
  • Lead complex troubleshooting and deep-dive diagnosis using packet captures, flow data, routing tables, logs, and change history
  • Meet resolution SLA targets and provide regular, accurate updates throughout the incident lifecycle

Root-Cause Analysis & Permanent Fixes

  • Determine technical root cause and distinguish symptoms from underlying causes
  • Design and implement permanent fixes, configuration changes, and routing/peering adjustments within scope
  • Partner with L 3 / Engineering SMEs, carriers, and transit/peering providers on architecture-level remediation
  • Validate fixes through controlled change and confirm the issue does not recur before closure

Major Incident Response & Technical Leadership

  • Act as technical lead on major-incident (Sev-0 / Sev-1) bridges, coordinating diagnosis across towers and providers
  • Make and document containment, rerouting, and recovery decisions under pressure
  • Serve as senior escalation point and network subject-matter expert for the GIOC

Reliability, Automation & Tooling

  • Identify toil and recurring issues, and build automation to reduce manual effort and MTTR
  • Improve observability — network monitoring, alerting thresholds, flow telemetry, and dashboards — to catch issues earlier
  • Contribute to capacity, performance, and resilience improvements across the network estate
  • Drive permanent reduction of repeat incidents through engineering and process change

Knowledge, Runbooks & Mentoring

  • Author and maintain runbooks, SOPs, and knowledge-base articles that raise L1 first-time resolution
  • Mentor and coach L1 engineers; review triage quality and provide feedback
  • Lead knowledge transfer and shadowing during onboarding and go-live

Post-Incident Review & Continuous Improvement

  • Lead post-incident reviews (PIR) and root-cause write-ups with clear corrective actions
  • Track corrective and preventive actions to closure and measure recurrence
  • Surface systemic risks and drive trend-based reliability improvements

Qualifications & Experience

  • 5-8 years of experience in network engineering, network operations, or infrastructure operations, including Senior Network Engineers escalation and root-cause ownership
  • Deep hands-on networking expertise — routing and switching, BGP/OSPF, DNS, firewalls, load balancing, and CDN
  • Strong network troubleshooting using packet capture, flow analysis, and routing/peering diagnostics
  • Proven experience leading complex incident diagnosis, root-cause analysis, and permanent remediation
  • Network automation/scripting skills (Python, Ansible, or equivalent) and familiarity with monitoring/observability and ITSM tooling
  • Strong written and verbal communication for technical leadership, documentation, and mentoring; willingness and ability to work a rotational 24x7 shift and on-call model, including nights, weekends, and holidays
  • Proficient in English verbal and written communication

Preferred Qualifications

  • ITIL V4 certification; experience with SRE or DevOps practices
  • Networking certification (CCNP / JNCIP or equivalent; CCIE a plus)
  • Experience with network automation (Ansible, Python, NETCONF/YANG) and CI/CD for network changes
  • Exposure to large-scale cloud networking, CDN, and GPU / high-density infrastructure
  • Experience with (physical) troubleshooting, including optical/transport and/or telecom troubleshooting
  • Prior experience as a senior escalation resource in a 24x7 NOC or command-center environment

Inclusion & Privacy

We are an equal opportunity employer and are committed to creating an inclusive environment for all employees. We welcome applications from individuals of all backgrounds and experiences, and we prohibit discrimination based on race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected status under applicable laws. Vultr will consider qualified applicants with arrest or conviction records in accordance with applicable laws and will not conduct a background check until after an offer of employment has been extended and accepted.

We also take your privacy seriously. We handle personal information responsibly and follow applicable laws, including U.S. privacy rules and India’s Digital Personal Data Protection Act, 2023. Your data is used only for legitimate business purposes and is protected with proper security measures.

Where allowed by law, applicants may request details about the data we collect, access or delete their information, withdraw consent for its use, and opt out of nonessential communications. For more details, please see our Privacy Policy.

Applying takes you straight to Vultr's own posting — no middlemen, no reposts.

Apply at Vultr

Similar roles