Observability & Enterprise Monitoring Architect with specialized expertise in SolarWinds platform architecture, design, and broader multi-tool observability ecosystems. This role will be responsible for the end-to-end architecture, deployment, implementation, optimization, integration, and operational governance of enterprise-scale implementations of monitoring solutions (SolarWinds). Roles & Responsibilities:Platform Architecture, Deployment & Lifecycle Management (SolarWinds)Core Module Architecture & Deployment: Design, deploy, configure, and optimize SolarWinds modules including NPM, NCM, NTA, SAM, and the broader Orion / SWOSH (Hybrid Cloud Observability) platform ecosystem.Deployment & Upgrade Strategy: Lead new platform rollouts, migrations, and routine/major version updates across platform components; establish standards for platform health governance using Active Diagnostics and My Deployment health checks.Polling Infrastructure Deployment: Architect, deploy, scale, and load-balance Additional Polling Engines (APEs) to ensure optimal performance, redundancy, and capacity across enterprise environments.Database & Storage Strategy: Oversee architectural strategy for the underlying MS SQL Database, ensuring high availability, performance tuning, and robust configuration and database backup governance. Alert Architecture, Dashboarding & ITSM IntegrationSignal & Alert Optimization: Design, implement, and tune custom Alert Triggers, Actions, and Threshold frameworks to eliminate alert noise and establish high-signal, actionable alerting.ITSM & Workflow Deployment: Deploy bi-directional ITSM/ticketing integrations to enable automated ticket creation, enrichment, routing, and lifecycle tracking.Reporting & Visibility Frameworks: Build enterprise operational and executive Dashboards, Views, and Reports tailored to multi-level stakeholder requirements.Incident & Deployment Support: Lead technical reviews for complex operational anomalies, troubleshoot systemic telemetry or deployment issues, and collaborate with domain teams on root cause analysis (RCA). at global scale.Protocol & Telemetry Mastery: Advanced understanding of SNMP (v2c/v3), WMI, WinRM, Syslog, NetFlow/sFlow, and core Observability pillars (Metrics, Logs, Traces).Automation & API Design: Intermediate skills in PowerShell/Python, REST APIs, and building API-driven automation for enterprise monitoring workflows.AIOps & Intelligent Automation: Strong grasp of AIOps concepts, machine learning algorithms for anomaly detection, automated event correlation, and predictive analytics within modern observability frameworks.Cloud & Hybrid Deployment: Hands-on experience architecting and deploying enterprise platform monitoring into AWS, Azure, or Google Cloud Platform environments.Infrastructure Foundations:System Administration: Advanced knowledge of Windows and Linux platform architecture and administration.Database Architecture: In-depth understanding of MS SQL/Database architecture, performance tuning, and query execution.Networking: Comprehensive understanding of enterprise networking architectures including TCP/IP, DNS, DHCP, Routing, and Switching.ITSM: Deep experience in enterprise ITSM processes, ITIL frameworks, and operational governance.
Create an account to see the full posting, access our search engine, and more.You're just 60 seconds away from your new Creativeloft account.