Senior Automation & Observability Engineer

🏢 Ensono · all 32 jobs
📍 United States
💰 USD 113,000 - 147,000 / annual
📅 Posted Sep 19, 2026 · via Himalayas
🏷 Observability Engineering, Infrastructure Automation, Site Reliability Engineering, Operational Governance, DevOps Engineer, Senior Security Automation Engineer +4 more
Apply on original site ↗

At Ensono , our Purpose is to be a relentless ally, disrupting the status quo and unleashing our clients to Do Great Things ! We enable our clients to achieve key business outcomes that reshape how our world runs. As an expert technology adviser and managed service provider with cross-platform certifications, Ensono empowers our clients to keep up with continuous change and embrace innovation.

We can Do Great Things because we have great Associates. The Ensono Core Values unify our diverse talents and are woven into how we do business. These five traits are the key to achieving our purpose:

Honesty, Reliability, Curiosity, Collaboration, and Passion.
About the role and what you'll be doing:

We are seeking an experienced IoT / Observability Engineer responsible for monitoring, managing, automating, and optimizing enterprise infrastructure, applications, IoT platforms, and enterprise telemetry ecosystems. The ideal candidate will possess strong expertise in observability platforms, monitoring technologies, automation frameworks, and operational support processes to ensure high availability, reliability, and performance of business-critical systems.

The engineer will support enterprise monitoring operations, FOAK services, ELT platforms, incident management, and automation initiatives while collaborating with Infrastructure, Cloud, Network, Application, and Service Delivery teams.

We want all new Associates to succeed in their roles at Ensono . That's why we've outlined the job requirements below. To be considered for this role, it's important that you meet all Required Qualifications. If you do not meet all of the Preferred Qualifications, we still encourage you to apply.
Key Responsibilities
Monitoring & Observability

- Design, implement, and maintain enterprise monitoring and observability solutions.

- Develop and maintain dashboards, alerts, and visualizations using Grafana.

- Monitor infrastructure, applications, middleware, and IoT services using IBM Instana, Grafana, SolarWinds, and related observability tools.

- Configure and manage data collection using Telegraf, Prometheus, and monitoring agents.

- Analyze metrics, logs, traces, events, and telemetry data to identify performance bottlenecks and service degradation.

- Support SLO, SLA, and operational health monitoring initiatives.

- Perform Root Cause Analysis (RCA) and troubleshooting for infrastructure and application issues.

Foak & Enterprise Logging/Telemetry

- Support onboarding, monitoring, and operational management of FOAK (First Office Application Kit) services and enterprise applications.

- Configure, validate, and troubleshoot Enterprise Logging & Telemetry (ELT) integrations across infrastructure, middleware, applications, and cloud platforms.

- Monitor telemetry pipelines, log ingestion, event correlation, and data quality to ensure complete observability coverage.

- Collaborate with engineering teams to improve telemetry standards, monitoring effectiveness, and proactive incident detection through ELT and observability frameworks.

- Support FOAK application integrations with Grafana, Instana, Prometheus, and enterprise monitoring platforms.

Infrastructure & Platform Monitoring
Monitor and support:

- Linux Servers

- Windows Servers

- VMware Infrastructure

- Citrix VDI Platforms

- DNS Services

- Proxy Services

- Middleware Platforms

- Integration Services

- Enterprise Applications

- IoT Platforms

Additional Responsibilities:

- Investigate performance issues, recurring alerts, and infrastructure anomalies.

- Validate monitoring platform health and monitoring coverage.

- Monitor capacity, availability, CPU, memory, storage, and service health metrics.

- Support platform upgrades, maintenance, and operational readiness reviews.

Database & Data Management

- Configure and maintain InfluxDB time-series databases.

- Manage data retention policies, performance tuning, and capacity planning.

- Develop operational dashb

Flights + hotels

This role requires you to be in the United States. If that means relocating or flying in, it is worth checking fares before you commit to a start date.

Compare flights and hotels →

← All remote jobs

Want more like this? Browse every live remote developer role.All remote developer jobs →
Get new developer jobs by email
Daily email, only when there's something new. One click to stop.

Get remote developer jobs like this by email

10 hand-picked jobs, one email a day. No spam, unsubscribe anytime.

Similar for you