Software Development Senior Specialist
<p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">Job Description</span></p><p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">We are seeking a Lead Integration & Observability Specialist to design, implement, and lead enterprise observability and reliability solutions, while supporting cloud-based integration platforms on AWS/Azure. The role focuses on monitoring, automation, and operational readiness of applications, APIs, data pipelines, and messaging systems.</span></p><p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">This is a hands-on technical leadership role with mentoring and solution ownership responsibilities. The working environment includes Windows-based servers and .NET-based applications. Prior experience in Windows/.NET environments is preferred but not mandatory. The candidate should be a fast learner and willing to work across different technologies, platforms, and application environments.</span></p><p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">Key Responsibilities</span></p><p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">• Lead the implementation of enterprise observability for applications, APIs, services, batch jobs, and data pipelines.<br>• Design and standardize monitoring, alerting, logging, metrics, and health checks across distributed systems.<br>• Integrate observability platforms with incident management and automation tools to support proactive issue detection and remediation.<br>• Support reliability and availability of integration platforms built on AWS/Azure.<br>• Perform advanced troubleshooting using logs, metrics, and traces to resolve production issues.<br>• Define operational readiness standards and non-functional requirements.<br>• Mentor engineers on observability best practices and platform usage.<br>• Collaborate with product, support, and operations teams to improve service stability and delivery.<br>• Work across different application environments, including Windows servers, .NET applications, cloud platforms, and integration/messaging systems.</span></p><p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">Required Skills Mandatory</span></p><p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">• 7+ years of overall IT experience.<br>• 5+ years of relevant experience in Observability / Monitoring / Reliability Engineering.<br>• Strong hands-on experience with enterprise observability tools, such as IBM Instana, Dynatrace, AppDynamics, Prometheus, or Grafana.<br>• Expertise in monitoring and alerting design.<br>• Log management and analysis.<br>• Metrics and distributed tracing.<br>• Health checks and SLO/SLI concepts.<br>• Experience monitoring AWS/Azure workloads.<br>• Strong troubleshooting and incident analysis skills.<br>• Experience defining operational and non-functional requirements.<br>• Technical leadership and mentoring experience.<br>• Automation and ITSM integration, including ServiceNow workflows and incident automation.<br>• CI/CD and release management exposure.<br>• Cloud integration and messaging exposure.</span></p><p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">Preferred / Good-to-Have Skills</span></p><p style="margin:0.0cm;font-size:12.0pt;font-family:Aptos, sans-serif"><span style="font-size:11.0pt">• Experience working in Windows server environments.<br>• Experience supporting or monitoring .NET-based applications, IIS, Windows services, or related Microsoft technology platforms.<br>• Ability to quickly learn and adapt to new tools, platforms, and application environments.<br>• Willingness to work across various technology stacks, including Windows/.NET, cloud platforms, integration tools, messaging systems, and observability platforms.<br>• Exposure to enterprise production support, application reliability, and operational readiness practices.</span></p>