Mumbai, Maharashtra, India
• Monitored the performance of services or processes running on Linux systems and tuned them as needed. • Managed user accounts on the Linux operating system, including creating new user accounts. • Installed and configured software packages on Linux systems according to customer requirements. • Performed troubleshooting in Linux environments, including log analysis, and storage checks. • Manage and maintain a fleet of 200+ Linux servers across on-premise and AWS environments. • Reduced manual configuration time by 40% by developing and implementing Ansible playbooks for patch management and software deployment. • Automated daily backup routines and log rotations using Cron jobs and Bash scripts, saving manual work weekly. • Monitored system performance and resource utilization, performing root cause analysis (RCA) on system outages. • Developed custom Dynatrce Dashboards and Management Zones to provide tailored performance insights for application owners. • Monitored Host Health metrics in Dynatrce (Disk I/O, Network Latency, Memory swap) to proactively scale resources before system degradation occurred. • Managed and resolved ServiceNow incidents and service requests, prioritizing issues to avoid SLA breaches, and following escalation procedures. • Handled infrastructure and applicative alerts using IBM Netcool. • Performed Autosys agent installation and upgrades on Linux servers maintaining compliance with scheduling interests. • Executes Autosys commands to manage jobs, check job logs, and restart failed jobs.
• Managed and supported the L1 Network. • Upgraded the firmware of Cisco devices. • Responsible for configuring new routers and switches. • Coordinating with Internet Service providers for link down, link flapping, latency, and routing-related issues. • Installation and replacement of routers and switches in Data centers. • Analyzing the ping response of IPs.