OOpportivia
All jobs
Full-time

Software Engineer: Site Reliability Engineering

Cellulant

Organization
Cellulant
Work mode
Hybrid
City
Accra, Accra
Country
Ghana
Job type
Full-time
Category
Technology

About Cellulant:

Cellulant is Africa’s leading payments company, providing seamless, secure and innovative solutions that empower businesses, banks, and global brands to thrive in a fast-changing global economy. 

With a presence in over 24 countries and 200+ payment methods, including cards, bank transfers, and mobile money, our single API payment platform, Tingg, simplifies collections, disbursements, and reconciliations. It processes over 4.5 million transactions daily for market leaders across sectors such as Airlines, Telecoms, E-commerce, Ride-Hailing, Retail, and Remittances. By simplifying how people pay and get paid, we drive trust, commerce and scale – and connect companies to their ambitions.  

Our Story:

Across Africa, payments are more than transactions. They are gateways to prosperity, connecting people, businesses and communities to opportunities and growth.

From enabling a logistics company in Lusaka to pay suppliers across borders, to enabling a hospitality brand in Lagos to scale effortlessly, to supporting an airline in Nairobi to reconcile payments from multiple platforms, Cellulant is the bridge that makes it all possible. 

Through trusted technology and customer-centric innovation, we build connections that inspire progress, strengthen economies and transform payments into a tool for progress.

Since our founding in 2003, we've continuously adapted and grown, leveraging our experiences to simplify payments for businesses. We are driven by an unshakable belief that seamless people-centred payments are the key to unlocking prosperity. 

Today, Cellulant powers online and offline payment processing, allowing businesses to collect payments, send payouts, and accelerate business growth.

Our Mission: 

To deliver seamless, secure and innovative payment solutions for businesses. 

Our Vision: 

To create a connected world where businesses move money as easily as they share ideas.

Role Overview: 

We are seeking a highly skilled and motivated Senior Software Engineer to lead our Site Reliability Engineering (SRE) team. As a Senior Software Engineer - SRE, you will play a crucial role in ensuring the seamless operation and performance of our systems, contributing to the overall reliability and efficiency of our services.

This role demands a blend of technical expertise, communication skills, and problem-solving capabilities to ensure seamless customer experiences.

What You'll do:

• Ticket Management: Collaborate with your team to ensure the timely recording of tickets on Sysaid, and drive the resolution process within specified Service Level Agreements (SLA).

• Proactive Monitoring: Take the lead in proactively monitoring our services to maintain continuous high-quality service for our customers. Identify potential issues before they impact the user experience and implement preventive measures.

• Issue Escalation and Follow-up: Effectively escalate service issues to the appropriate stakeholders and follow up diligently to ensure timely resolution within specified SLA. Keep stakeholders informed of the progress and resolution status.

• Data Management: Oversee the continuous archiving of data and maintain databases in a lightweight state to maximize transactional efficiency. Implement best practices for data management and contribute to database optimization efforts. 

• On Call Support: Be available for on-call support, especially during critical periods or in the event of system outages.

• Change Management Support: Collaborate with the development teams to communicate and apply changes effectively. 

What we are looking for:

Must have experience:

• Minimum of 4-5 Years  of experience in a similar role, with a proven track record of leading or having a senior role in  SRE teams and maintaining system reliability.

• Technical Skills in the following:-

• Proficient in Kubernetes administration, troubleshooting, and optimization.

• Proficient in Sysaid or similar ticketing systems.

• Strong expertise in proactive monitoring tools and practices.

• Experience with relational database management and optimization.

• SLA Management: A solid understanding of Service Level Agreements and experience in managing and meeting SLA commitments.

Experience that will count in your favor: 

• Substantial hands-on experience in administering Kubernetes clusters, including deployment, scaling, monitoring, and troubleshooting.

• Experience working with cloud platforms (e.g., AWS, Azure, Google Cloud) and deploying applications on cloud infrastructure, especially in conjunction with Kubernetes.

• Strong scripting skills (e.g., Python, Bash) and experience in automating routine tasks and processes, contributing to efficient Kubernetes cluster management.

Nice to have experience: 

• In-depth knowledge and experience with containerization technologies such as Docker, and understanding of container orchestration tools beyond Kubernetes.

• Proficiency in using monitoring tools like Prometh

Similar jobs