
SRE with Network
Job Summary
We are looking for an experienced Site Reliability Engineer (SRE) with strong expertise in networking, DNS, load balancing, and hybrid cloud environments. The candidate will be responsible for maintaining the reliability, availability, performance, and scalability of production infrastructure and services.
The ideal candidate should have strong troubleshooting skills and experience working across network, cloud, infrastructure, and application environments.
Key Responsibilities
Monitor and maintain the availability and reliability of production systems and services.
Troubleshoot complex network, infrastructure, and application connectivity issues.
Manage and troubleshoot DNS services, DNS resolution, records, and configuration issues.
Configure, manage, and troubleshoot Load Balancers and traffic routing.
Work with Layer 4 and Layer 7 networking and understand TCP/IP, HTTP/HTTPS, routing, and network connectivity.
Support hybrid cloud environments involving on-premises infrastructure and public cloud platforms.
Troubleshoot connectivity between on-premises data centers and cloud environments.
Participate in production incidents, troubleshooting, root cause analysis (RCA), and problem management.
Develop automation scripts and tools to reduce manual operational activities.
Configure and maintain monitoring, alerting, and observability solutions.
Work closely with Network, Cloud, DevOps, Security, and Application teams.
Participate in on-call support and resolve production issues within defined SLAs.
Document infrastructure, troubleshooting procedures, incident reports, and operational processes.
Identify opportunities to improve system reliability, performance, and scalability.
Never pay to get work. If a listing asks for a fee, it is a scam. The ten signs →