London Area, United Kingdom
- Modernized core network infrastructure by migrating legacy IPAM systems to a scalable platform with an event-driven DNS backend processing millions of events per day, improving scalability and reliability for Cisco's global network. - Managed on-prem Kubernetes clusters across multiple regions running production network services, maintaining cluster health and resolving pod/node-level failures to ensure high service availability. - Eliminated cascading failures and improved reliability of the DNS service handling 2M QPS, by implementing client attributions and rate limiting, reducing service disruptions under high load. - Bridged data consistency gaps during migration of 3.5M+ IPAM and DNS records by building synchronization and validation mechanisms, ensuring seamless cutover with no service disruption. - Cut MTTR by 60% by automating incident remediation workflows and creating operational runbooks that enabled efficient incident delegation and faster resolution.
• Developed visually compelling KPI dashboards aligned with business requirements empowering users with data-driven decision-making using real-time insights and actionable information. • Initiated and drove end-to-end data pipeline, culminating in dynamic dashboards that enabled high-level monitoring of a critical initiative to reduce network ACLs.
Revamped the front-end of a centralized vulnerability management system to expedite issue resolution and establish transparent SLAs.
• Implemented a Statistical Intrusion Detection Algorithm to effectively detect anomalous network traffic, especially DDoS attacks. • Introduced the algorithm using DPDK enabled Open vSwitch connected to the SDN controller.
• Accelerated lead generation process by implementing an end-to-end machine learning pipeline to automatically generate leads from webpage inputs, resulting in a 60% increase in efficiency over previous manual methods.