Site Reliability Engineer 3

Granicus India All jobs
12 day(s) ago
Job Overview
Company Granicus India
Job Type fulltime
Category Engineering and Information Technology
Last Seen 12 day(s) ago

Job Description

The Company Serving the People Who Serve the People Granicus is driven by the excitement of building implementing and maintaining technology that is transforming the Govtech industry by bringing governments and its constituents together. We are on a mission to support our customers with meeting the needs of their communities and implementing our technology in ways that are equitable and inclusive. Granicus has consistently appeared on the GovTech 100 list over the past 5 years and has been recognized as the best companies to work on BuiltIn. Over the last 25 years we have served 5500 federal state and local government agencies and more than 300 million citizen subscribers power an unmatched Subscriber Network that use our digital solutions to make the world a better place. With comprehensive cloud-based solutions for communications government website design meeting and agenda management software records management and digital services Granicus empowers stronger relationships between government and residents across the U.S. U.K. Australia New Zealand and Canada. By simplifying interactions with residents while disseminating critical information Granicus brings governments closer to the people they serve—driving meaningful change for communities around the globe. Want to know more? See more of what we do here .

Job Summary

Granicus is the leading provider of citizen engagement technologies and services for the public sector bringing governments closer to the people they serve with the first-and-only Civic Engagement Platform. Granicus works with more than 5500 government organizations and connects more than 280 million people in the largest Citizen Subscriber Network of its kind. What Your Impact Will Look Like Granicus is seeking an experienced and highly skilled Site Reliability Engineer 3 (SRE) to join our SRE team. As a SRE you will play a pivotal role in ensuring the reliability scalability and performance of our services. You will lead efforts in building and maintaining a robust infrastructure automating processes and guiding the team to implement best practices in site reliability. Essential Function On-call Production Support Provide production support on a shift according to the team on-call roster. Work on the customer and internal engineering/implementation team raised tickets while not on-call for production support. For example a client may request to correct some data on the database server which cannot be done through the web interface. Work on SREs backlog items. Monitor and Maintain Systems Continuously monitor the health and performance of our services systems and infrastructure. Respond to alerts and incidents promptly to ensure high availability. Automate Processes Develop and maintain automation scripts and tools to streamline operations and reduce manual intervention. Incident Management Assist in troubleshooting and resolving incidents performing root cause analysis and implementing long-term fixes to prevent recurrence. System Improvements Participate in designing and implementing system improvements to enhance reliability scalability and performance. Collaboration Work closely with software engineers to understand application requirements provide feedback on design and architecture and support deployment and release processes. Documentation Create and maintain documentation for processes procedures and troubleshooting guides to ensure knowledge sharing within the team. Capacity Planning Assist in capacity planning activities to anticipate future needs and ensure that our infrastructure can handle growth. Security Implement and adhere to security best practices to protect our systems and data. You Will Love This Job If You Have

Technical Skills

Good understanding of Linux/Unix systems networking and cloud services (AWS Azure or Google Cloud). Experience with scripting languages such as Python Bash or Ruby. Education Bachelor’s or master’s degree in computer science Information Technology or a related field or equivalent practical experience. Experience 5+ years of experience in site reliability engineering system administration or a similar role with a proven track record of managing large-scale high-availability systems AI skills Familiarity with AI/ML operations including model lifecycle management vector databases and inference performance tuning.

Technical Skills

Expertise in Linux/Unix systems networking and cloud services (AWS Azure or Google Cloud). Proficiency in scripting languages (Python Bash Ruby) and programming languages (Go Java C++). Tools and Technologies Advanced knowledge of monitoring and logging tools like Elastic (Prometheus Grafana Splunk) configuration management (Ansible Chef Puppet) and CI/CD pipelines. Problem-Solving Strong analytical and problem-solving skills with the ability to diagnose and resolve complex issues efficiently. Communication Excellent verbal and written communication skills with the ability to convey complex technical concepts to non-technical stakeholders. Leadership Demonstrated ability to lead and mentor a team drive projects to completion and manage cross-functional initiatives. Certifications Relevant certifications such as AWS Certified DevOps Engineer AWS Certified Machine Learning – Specialty Google Cloud Professional DevOps Engineer or similar are a plus.

About Us

Don’t have all the skills/experience mentioned above? At Granicus we are trying to build diverse inclusive teams. We do not have degree requirements for most of our roles. If you don’t meet every requirement above but are excited to learn more we encourage you to apply. We might just be able to find another role that could be a perfect fit! Security And Privacy

Requirements

Responsible for Granicus information security by appropriately preserving the Confidentiality Integrity and Availability (CIA) of Granicus information assets in accordance with the company's information security program. Responsible for ensuring the data privacy of our employees and customers their data as well as taking all required privacy training in a timely manner in accordance with company policies. The Team We are a remote-first company with a globally distributed workforce across the United States Canada United Kingdom India Armenia Australia and New Zealand. The Culture At Granicus we are building a transparent inclusive and safe space for everyone who wants to be a part of our journey. A few culture highlights include – Employee Resource Groups to encourage diverse voices Coffee with Mark sessions – Our employees get to interact with our CEO on very important and sometimes difficult issues ranging from mental health to work-life balance and current affairs. Microsoft Teams communities focused on wellness art furbabies family parenting and more. We bring in special guests from time to time to discuss issues that impact our employee population The Impact We are proud to serve dynamic organizations around the globe that use our digital solutions to make the world a better place — quite literally. We have so many powerful success stories that illustrate how our solutions are impacting the world. See more of our impact here .

Unlock: Sign Up for free / Sign In and use the searches from your home page or the links in the footer.