Site Reliability Engineer I (Payzapp)

Posted 5 Days Ago
Be an Early Applicant
Bangalore, Bengaluru, Karnataka
Entry level
Cloud • Fintech • Financial Services
The Role
The Site Reliability Engineer I is responsible for ensuring the reliability of software systems, developing automation tools, responding to incidents, and optimizing performance. This role involves implementing infrastructure as code, monitoring system behavior, collaborating on security best practices, and maintaining disaster recovery plans.
Summary Generated by Built In

All about Zeta Suite :

Zeta is the world’s first and only Omni Stack for banks and fintechs. We are rethinking payments from core to the edge, led by the vision to augment the purpose of money and banking with technology. A single, modern software stack comprising processing, loans, customizable mobile and web apps, fraud engine, and rewards for retail banking. We are a new-age, high-growth startup (& a unicorn!) founded in 2015 by two visionary leaders Bhavin Turakhia & Ramki Gaddipati, whose entrepreneurial legacy & excellence has put us on top of the global fintech ecosystem. Zeta counts amongst its customers, over 10 banks and 25 fintechs, across 8 countries - some of our notable clients include Sodexo - a leading issuer of employee benefits & rewards with over 30 million global users, and HDFC Bank - the 14th largest bank by market cap in the world. Learn more about our manifesto & beyond.

Responsibilities

  • System Reliability: Ensuring the reliability of software systems by designing, implementing, and maintaining scalable and reliable infrastructure.
  • Automation: Developing automation tools and scripts to streamline operational tasks, reduce manual intervention, and improve overall system efficiency.
  • Incident Response and Resolution: Monitoring system performance and responding to incidents promptly to minimize downtime and ensure high availability.
  • Capacity Planning: Analyzing system usage patterns and forecasting future capacity needs to ensure that the infrastructure can handle current and future demands.
  • Performance Optimization: Identifying and addressing performance bottlenecks in software systems through optimization and tuning.
  • Infrastructure as Code (IaC): Implementing infrastructure as code practices, using tools like Terraform or Ansible, to define and manage infrastructure in a version-controlled and automated manner.
  • Monitoring and Logging: Implementing and maintaining monitoring and logging solutions to gain insights into system behavior, troubleshoot issues, and proactively address potential problems.
  • On-Call Support: Participating in an on-call rotation to respond to incidents outside of regular working hours and ensure 24/7 system availability
  • Security: Collaborating with security teams to implement and maintain security best practices in infrastructure and application
  • Disaster Recovery Planning: Developing and maintaining disaster recovery plans to ensure that systems can quickly recover from major outages or failures
  • Continuous Improvement: Continuously analyzing system performance, reliability, and incidents to identify areas for improvement and implementing changes to enhance overall system resilience.

Skills

  • Programming Languages: Proficiency in one or more programming languages, commonly Python, Go, Shell, Bash.
  • Automation and Scripting: Strong automation skills using tools like Ansible, Puppet, Chef, or custom scripts. Knowledge of Infrastructure as Code (IaC) tools like Terraform
  • Containerization and Orchestration: Experience with containerization technologies like Docker and container orchestration platforms like Kubernetes.
  • Cloud Computing: Proficiency in any of the cloud platforms such as AWS, Azure, or Google Cloud Platform, and knowledge of managing infrastructure in the cloud.
  • Monitoring and Logging: Familiarity with monitoring tools (e.g., Prometheus, Grafana, ELK stack) and logging frameworks to track system performance and troubleshoot issues.
  • Networking: Understanding of networking concepts, protocols, and troubleshooting skills.
  • Security: Knowledge of security best practices, including encryption, access controls, and vulnerability management.
  • Continuous Integration/Continuous Deployment (CI/CD): Understanding and implementation of CI/CD pipelines for automated testing and deployment.
  • Load Balancing: Experience in incident response, troubleshooting, and resolution.
  • Version Control: Proficient use of version control systems like Git.

Experience and Qualifications

  • 1-2 year of experience in site reliability engineering.
  • B.Tech/M.Tech in computer science, information technology or a related field.
  • Having experience working for a product organization is a plus.

Equal Opportunity

  • Zeta is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. We encourage applicants from all backgrounds, cultures, and communities to apply and believe that a diverse workforce is key to our success

Top Skills

Bash
Go
Python
Shell
The Company
HQ: San Francisco, California
1,834 Employees
On-site Workplace
Year Founded: 2015

What We Do

Founded in 2015, Zeta is a provider of next-gen credit card processing platform. Zeta’s cloud-native and fully API-enabled stack offers a comprehensive range of capabilities, including processing, issuing, lending, core banking, fraud detection, and loyalty programs. With a strong focus on technology, Zeta has over 1700+ employees and contractors, with more than 70% dedicated to technology roles. Operating across the US, UK, Middle East, and Asia, Zeta has served a global customer base of 35+ clients who have issued over 15 million cards on Zeta's platform to date. Backed by prominent investors such as Softbank Vision Fund 2 and Mastercard, Zeta has raised $280 million, at a valuation of $1.5 billion.

Similar Jobs

BlackLine Logo BlackLine

Senior Site Reliability Engineer

Cloud • Fintech • Information Technology • Machine Learning • Software • App development • Generative AI
Remote
Hybrid
Bengaluru, Karnataka, IND
1810 Employees

Exabeam Logo Exabeam

Senior Site Reliability Engineer

Artificial Intelligence • Information Technology • Machine Learning • Security • Software • Cybersecurity • Generative AI
Hybrid
Bangalore, Bengaluru, Karnataka, IND
850 Employees
Bangalore, Bengaluru, Karnataka, IND
913 Employees

Zoom Logo Zoom

Site Reliability Engineer

Artificial Intelligence • Information Technology • Software
Industrial Area SSI, Rajaji Nagar, Bangalore, Karnataka, IND
11053 Employees

Similar Companies Hiring

Energy CX Thumbnail
Utilities • Professional Services • Greentech • Financial Services • Energy • Consulting • Business Intelligence
Chicago, IL
55 Employees
MassMutual India Thumbnail
Insurance • Information Technology • Fintech • Financial Services • Big Data
Hyderabad, Telangana
Jobba Trade Technologies, Inc. Thumbnail
Software • Professional Services • Productivity • Information Technology • Cloud
Chicago, IL
45 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account