Senior Software Engineer - Tooling & Automation
- Remote
- All software engineering jobs
- United States
- FullTime
- Engineering
About the role
Who We Are
Vultr is on a mission to make high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators around the world. With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal, and Cloud Storage solutions. In December 2024 Vultr announced an equity financing at a $3.5 billion valuation. Founded by David Aninowsky and self-funded for over a decade, Vultr has grown to become the world’s largest privately-held cloud infrastructure company.
Vultr Cares
100% company-paid insurance premiums for employee medical, dental and vision plans.
401(k) plan that matches 100% up to 4%, with immediate vesting
Professional Development Reimbursement of $2,500 each year
11 Holidays + Paid Time Off Accrual + Rollover Plan
Commitment matters to Vultr! Increased PTO at 3 year and 10 year anniversary + 1 month paid sabbatical every 5 years + Anniversary Bonus each year
$500 stipend for remote office setup in first year + $400 each following year
Internet reimbursement up to $75 per month
Gym membership reimbursement up to $50 per month
Company paid Wellable subscription
Join Vultr
Vultr is seeking a highly skilled and experienced Senior Software Engineer to serve as the technical owner connecting our engineering alerting to Vultr's new Global Integrated Operations Command Center (GIOC) in Chennai, India, and to build the consolidated single-pane dashboard GIOC will operate from. The ideal candidate looks at five disconnected monitoring consoles and a dozen siloed alert feeds, winces, and then methodically builds the one dashboard that replaces all of them. This is a highly visible, focused role requiring deep familiarity with Vultr's current architecture, monitoring stack, and operational workflows. This is your opportunity to build the operational backbone GIOC runs on from day one.
Key Responsibilities
Map every alert source owned by Systems, Network, and Security engineering (Icinga, Grafana, SIEM/Splunk/QRadar/Sentinel, BMC/IPMI, provider-native consoles) and identify the cleanest integration point for each — webhook, syslog forward, API poll, or native connector
Build and maintain the routing/forwarding layer that pushes agreed top-priority alerts from each pillar into GIOC's ticketing and paging tools, with priority/severity preserved and enough context for L1 triage without direct console access
Work with each pillar's engineering SMEs to validate that routing logic matches each pillar's defined resolve-vs-escalate boundary
Own the on-call/notification wiring so critical alerts reach GIOC within agreed acknowledgment SLAs, with no manual relay step
Design and build a unified operational dashboard spanning all five datacenter/provider platforms (Airtel, Nextra, Equinix, QTS, CoreSite), replacing today's tool-hopping model
Define dashboard v1 scope jointly with GIOC and pillar leads and ship it on a committed milestone
Run legacy tools and the new dashboard in parallel during transition so GIOC never loses visibility mid-cutover
Build in SLA/health visualizations (ack time, escalation accuracy, open-alert aging) that GIOC leadership can use directly in status reporting
Build correlation engines and integrate all sources of telemetry from various product teams
Document the routing architecture and dashboard so both can be maintained by GIOC or the Tools team after this assignment ends
Qualifications
5–8 years in Tools/Platform Engineering or a similar infrastructure role, including 3–5 years hands-on with monitoring/alerting tooling (Icinga or equivalent NMS, Grafana, and at least one SIEM platform — Splunk, QRadar, or Sentinel)
Experience rapidly mapping an unfamiliar, multi-team monitoring and alerting estate and consolidating it — ideally at a cloud, hosting, or multi-datacenter operator. Equivalent hands-on Vultr experience is a strong plus.
3+ years building integrations against APIs/webhooks, with scripting ability (Python or similar)
2–3 years building or substantially modifying operational dashboards (Grafana, Kibana, Power BI, or similar)
Understanding of ticketing/ITSM workflow tools (Jira, Confluence, Mattermost) and how alert priority/severity maps into ticket fields
Preferred: prior exposure to multi-provider/datacenter operational environments, and experience supporting or building for a NOC/SOC or similar 24x7 operations function
Compensation
$125,000 - $135,000
Final compensation will vary depending on years of experience, background/skill set, location, and applicable laws.
Inclusion & Privacy
We are an equal opportunity employer and are committed to creating an inclusive environment for all employees. We welcome applications from individuals of all backgrounds and experiences, and we prohibit discrimination based on race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected status under applicable laws. Vultr will consider qualified applicants with arrest or conviction records in accordance with applicable laws and will not conduct a background check until after an offer of employment has been extended and accepted.
We also take your privacy seriously. We handle personal information responsibly and follow applicable laws, including U.S. privacy rules and India’s Digital Personal Data Protection Act, 2023. Your data is used only for legitimate business purposes and is protected with proper security measures.
Where allowed by law, applicants may request details about the data we collect, access or delete their information, withdraw consent for its use, and opt out of nonessential communications. For more details, please see our Privacy Policy.
Description as published by Vultr.