Jobs / Software Development Engineer (SDE) / Red Hat Hiring Senior Information Systems Engineer in Pune | Hybrid
Red Hat logo

Red Hat

Red Hat Hiring Senior Information Systems Engineer in Pune | Hybrid

location_on Pune | Hybrid
work 3+ years Experience
payments Competitive Salary (Not disclosed)
schedule Full Time

Role Overview

Job Overview

Red Hat is hiring a Senior Information Systems Engineer for its enterprise monitoring, logging, and observability environment in Pune. This is a full-time hybrid opportunity focused on platform engineering, cloud-native observability, telemetry management, incident response, and reliability practices.

The role is centered on building and maintaining an enterprise-wide observability ecosystem that handles large volumes of logs, metrics, and traces. You will work with platforms such as Datadog, Sumo Logic, Cribl, PagerDuty, Ansible, and OpenTelemetry to help engineering and product teams gain reliable visibility into their applications and infrastructure.

A major part of the position involves improving telemetry quality and controlling ingestion costs. You will configure data pipelines, automate monitoring-agent management, establish consistent observability standards, and support teams in creating useful alerts and operational dashboards.


Key Responsibilities

  • Maintain and extend monitoring dashboards, metrics, application performance monitoring, and tracing capabilities across Datadog and Sumo Logic.
  • Manage multi-tenant observability workspaces and help maintain consistent tagging and telemetry standards.
  • Configure PagerDuty services, event routing, service orchestration, on-call schedules, alert intelligence, and integrations with collaboration tools such as Slack.
  • Optimize telemetry pipelines through Cribl by filtering unnecessary data, removing duplicate fields, reducing oversized payloads, and directing useful signals to the appropriate destinations.
  • Support the separation of high-value operational telemetry from lower-value compliance information that may be retained in archival storage.
  • Automate deployment, onboarding, patching, and configuration consistency for monitoring agents and telemetry components using Ansible Playbooks and Roles.
  • Help establish and enforce telemetry schemas, log formats, operational signals, and enterprise observability standards.
  • Contribute to OpenTelemetry collector implementation and configuration.
  • Work with infrastructure, application, DevOps, and product engineering teams to identify meaningful metrics and improve system visibility.
  • Create technical documentation, runbooks, and enablement material for modern monitoring, logging, and alerting practices.
  • Share knowledge with internal teams and support adoption of standardized observability approaches.


Required Skills

Candidates should have at least 3 years of Python development experience and practical knowledge of enterprise monitoring and observability environments.

Strong Datadog experience is expected, including AWS integrations and dashboard templating. Candidates should also understand monitoring of distributed systems and have experience working with infrastructure, application, and DevOps teams to define useful operational metrics.

Knowledge of Site Reliability Engineering concepts is important because the position supports reliability, incident response, platform maintenance, and production visibility. Strong AWS architecture knowledge and familiarity with cloud-native observability are also required.

The role expects familiarity with OpenShift or Kubernetes, Ansible, Infrastructure-as-Code concepts, and OpenTelemetry. Experience with SignalFX or Splunk Observability Cloud and older monitoring approaches is also relevant.


Preferred Skills

Certifications related to Datadog, AWS, or other observability platforms are preferred. Experience participating in enterprise-scale monitoring or observability transformations can further strengthen a candidate's profile.

Candidates who have worked with large telemetry environments, automated monitoring-agent deployment, cloud infrastructure, centralized logging, incident management, or SRE practices may be particularly well suited to the position.


Education

The supplied job description does not specify a required educational qualification. Candidates should therefore rely on their relevant engineering, cloud, monitoring, observability, automation, and platform experience when assessing their suitability.


Experience

The job posting specifically asks for 3 years of Python development experience. It does not provide a separate maximum experience requirement or a clearly defined total years-of-experience range. The position is described as a Senior Information Systems Engineer, so practical experience with enterprise observability and platform engineering is highly relevant.


Required Technologies

  • Python
  • Datadog
  • AWS
  • Sumo Logic
  • SignalFX
  • Splunk Observability Cloud
  • Cribl
  • PagerDuty
  • Slack
  • Ansible
  • OpenTelemetry
  • OpenTelemetry Collectors
  • Kubernetes
  • OpenShift
  • Infrastructure-as-Code
  • Site Reliability Engineering
  • Application Performance Monitoring
  • Monitoring
  • Logging
  • Metrics
  • Distributed Systems
  • Cloud-Native Observability


Soft Skills

Strong communication and stakeholder management skills are important because the engineer will work across infrastructure, application, DevOps, product, and internal customer teams. The ability to explain observability concepts clearly and guide teams toward consistent operational practices is valuable.

Candidates should also be comfortable mentoring engineers, documenting technical procedures, leading enablement sessions, troubleshooting production concerns, and collaborating on platform improvements. Attention to detail is important when managing telemetry configurations, alert rules, data routing, and cost controls.


Benefits of Working in this Role

The role provides practical exposure to enterprise-scale observability, cloud infrastructure, SRE practices, telemetry pipelines, monitoring automation, and modern logging technologies. It also offers opportunities to work across multiple engineering teams and contribute to standardization and platform transformation.

The position can help an experienced engineer deepen skills in cloud-native monitoring, automation, reliability engineering, incident management, data pipeline optimization, and enterprise platform engineering.


Work Mode

The job is listed as hybrid. It is a full-time position based in Pune.


Location

The role is located in Pune, India. Candidates should confirm the applicable office attendance expectations and current work arrangement during the application process.


Who Should Apply

This opportunity is suitable for Senior Information Systems Engineers, Observability Engineers, Platform Engineers, Site Reliability Engineers, DevOps Engineers, Cloud Engineers, Monitoring Engineers, and infrastructure professionals with relevant experience.

It is especially relevant for candidates who have worked with Datadog, AWS, Sumo Logic, Cribl, PagerDuty, Ansible, OpenTelemetry, Kubernetes, or enterprise monitoring environments. Engineers interested in improving reliability, telemetry quality, operational visibility, and automation should consider this role.


Career Growth

This position can support career growth toward senior platform engineering, observability architecture, SRE, cloud engineering, DevOps, and enterprise monitoring leadership. Working with high-volume telemetry and multiple observability platforms can strengthen expertise in distributed systems, cloud operations, automation, reliability, and platform governance.

The mentoring and enablement responsibilities can also help experienced engineers develop stronger technical communication, stakeholder management, documentation, and engineering leadership skills.


Application Advice

Applicants should highlight hands-on experience with Datadog, AWS, Python, Ansible, OpenTelemetry, Sumo Logic, Cribl, PagerDuty, Kubernetes, or related technologies. Clearly describe experience with monitoring distributed systems, building dashboards, configuring alerts, automating agent deployment, managing telemetry pipelines, and applying SRE practices.

If you hold Datadog, AWS, or related observability certifications, include them prominently. Experience with enterprise monitoring transformations, cloud-native observability, Infrastructure-as-Code, and cross-functional engineering collaboration should also be clearly presented.

Technical Ecosystem

Eligibility Criteria

school

Education

B.Tech

work_history

Experience

3+ years

Top Picks for You

smart_toy

SoftoBot

Beta

Your job search assistant

Ask about roles, locations, remote work, or fresher jobs.