Current jobs related to Site Reliability Engineer - Jakarta - Paymentology
-
Site Reliability Engineer
2 weeks ago
Jakarta, Indonesia Abhidi Solution Private Limited Full time**Responsibilities**: - Administer production related jobs - Address production issue - Improve system reliability through configuration or code changes - System monitoring and improve system observability - Remove toil and automate whenever possible - Problem solving, including troubleshoot a production issue **Skills**: - Experience with cloud...
-
Site Reliability Engineer
1 week ago
Jakarta, Indonesia PT Midas Daya Teknologi Full timeJob Description: Works with project engineering to ensure the reliability and maintainability of new and modified software. - The reliability engineer is responsible for adhering to the life cycle software management process throughout the entire life cycle. - Also responsible for end-to-end site reliability including service offerings, in particular...
-
Site Reliability Engineer
7 months ago
Jakarta, Indonesia Pro Sigmaka Full timeWe established at 2012. With experience in several industry sectors, a broad portfolio and technology platform as well as bringing a dedicated and highly qualified team, enabling the talent we provide to provide fast and responsive services, making it the best choice for companies that want to increase the usability of their businesses. OUR SERVICES -...
-
Site Reliability Engineer
2 weeks ago
Jakarta, Indonesia Ajaib Full timeCompany Description **Job Description**: - Perform day-to-day operations to support developers and DevOps. - Create end-to-end monitoring, logging, and alerting system. - Provide technical assistance to improve system performance, capacity, reliability and scalability - Perform root cause analysis of reliability issues. - Document every action so your...
-
Site Reliability Engineer
3 months ago
Jakarta, Indonesia PT. Amalura Multi Dimensi Full timeManage and optimize cloud infrastructure (AWS, GCP, Azure). - Administer Linux system, ensuring stability and security. - Implement observability (e. g, OpenTelemtry, HoneyComb, Sentry) to monitor performance. - Optimize content delivery networks (e. g., Akamai) to enhance user experience. - Design monitoring, alerting, and incident response procedure for...
-
Site Reliability Engineer Remote
1 day ago
Jakarta, Indonesia Kalibrr Full timeYour main responsibilities as a Site Reliability Engineer at Kalibrr are: Engage in and improve the whole lifecycle of the Kalibrr services-design, deployment, operation, and refinement. Practice incident response and blameless postmortems. Participate in an on-call rotation Scale systems and operations through automation. Maintain services by monitoring...
-
Site Reliability Engineer
9 months ago
Jakarta, Indonesia PT Tiga Daya Digital Indonesia (Eksad Technology) Full timeTiga Daya Digital Indonesia, a susidiary company of Triputra Group and DCI Group To be IT partner to enable client growth rapidly. Eksad Providing Services High Quality Based on Strong Experience in the industry and technology. Building the right IT Service Solution to enable it Partners in speeding up business development based on digital technology by...
-
Site Reliability Engineer
3 days ago
Jakarta, Indonesia Global Tiket Network Full timeWe think you also hate when travel app is giving you a headache, right? A slight misinformation can ruin the trip. - That is exactly what we are tackling as t-fam! Making sure that our 17+ million users have the best experience in crafting their own adventure. LI-Hybrid Catch the sunrise on the top of Padar Island and see fascinating views of the boundless...
-
Site Reliability Engineer
7 months ago
Jakarta, Indonesia Digital Muda Solutions Full timeDeskripsi: - Menjaga ketersediaan, kehandalan, dan performa sistem dengan fokus pada infrastruktur teknis, keamanan, dan skala pengguna. - Berkolaborasi dengan tim pengembangan dan operasi untuk merancang, menguji,dan menerapkan praktik terbaik dalam infrastruktur teknologi, serta melakukan perbaikan dan peningkatan sesuai kebutuhan. - Memastikan integrasi...
-
Site Reliability Engineer
2 weeks ago
Jakarta, Indonesia Catalyst Tech Full timeAt Catalyst, People are the heartbeat for our company. We believe that good quality people will have a positive impact to our business. We are looking for a **Site Reliability Engineer / DevOps** to join our growing team. If you are passionate about being part of the team, building some of the most critical products, Working alongside teams in the industry...
-
Site Reliability Engineer
9 months ago
Jakarta, Indonesia PT Salva Teknologi Digital Full timeSite Reliability Engineer (Junior) - Applicants should have sufficient qualification and relevant experiences in the respective fields "Waspada terhadap Modus Penipuan pada saat proses interview. Perusahaan tidak akan memungut biaya apapun dalam melakukan proses interview. Mohon segera melaporkan ke kami, jika pada saat Anda diundang untuk interview dan...
-
Site Reliability Engineer
4 months ago
Jakarta, Indonesia Hukumonline.com Full timeManage and optimize cloud infrastructure on AWS, GCP, and Azure. - Administer and maintain Linux-based systems, ensuring their stability and security. - Implement and maintain observability solutions, including OpenTelemetry, HoneyComb, and Sentry, to monitor system performance and diagnose issues. - Configure and optimize content delivery networks, with a...
-
Site Reliability Engineer
7 months ago
Jakarta, Indonesia AccelByte Full timeAt AccelByte, our mission is to empower game creators by providing them with the backend platform and tools required to make scalable, reliable AAA-quality games. The company was founded in 2016 by industry veterans who have engineered online systems for some of the largest game and distribution platforms in the world including Fortnite, Epic Store, Xbox...
-
Site Reliability Engineer
7 months ago
Jakarta, Indonesia AccelByte Full timeAt AccelByte, our mission is to empower game creators by providing them with the backend platform and tools required to make scalable, reliable AAA-quality games. The company was founded in 2016 by industry veterans who have engineered online systems for some of the largest game and distribution platforms in the world including Fortnite, Epic Store, Xbox...
-
Site Reliability Engineer
7 months ago
Jakarta, Indonesia AccelByte Full timeAt AccelByte, our mission is to empower game creators by providing them with the backend platform and tools required to make scalable, reliable AAA-quality games. The company was founded in 2016 by industry veterans who have engineered online systems for some of the largest game and distribution platforms in the world including Fortnite, Epic Store, Xbox...
-
Senior Site Reliability Engineer
8 months ago
Jakarta, Indonesia DKatalis Full time**Site Reliability Engineer**: **About DKatalis** DKatalis is a financial technology company with multiple offices in the APAC region. In our quest to build a better financial world, one of our key goals is to create an ecosystem linked financial services business. DKatalis is built and backed by experienced and successful entrepreneurs, bankers, and...
-
Site Reliability Engineer(DevOps)
7 months ago
Jakarta, Indonesia Digital Muda Solutions Full timeDeskripsi: - Menjaga ketersediaan, kehandalan, dan performa sistem dengan fokus pada infrastruktur teknis, keamanan, dan skala pengguna. - Berkolaborasi dengan tim pengembangan dan operasi untuk merancang, menguji,dan menerapkan praktik terbaik dalam infrastruktur teknologi, serta melakukan perbaikan dan peningkatan sesuai kebutuhan. - Memastikan integrasi...
-
Site Reliability Engineer
7 months ago
Jakarta, Indonesia PT Astra Digital Mobil (mobbi) Full timeJob Description: - Maintain system availability, reliability and performance by focusing on technical infrastructure, security and user scale. - Collaborate with development and operations teams to design, test, and implement best practices in technology infrastructure, and make fixes and improvements as needed. - Conduct in-depth analysis of incidents and...
-
Site Reliability Engineer
3 weeks ago
Jakarta, Indonesia Zenius Education Full timeDesign and implement the architecture of the next generation of automated infrastructure following Infrastructure as a Code model. Build and maintain container native CI/CD pipelines. Build tools and automation to improve system’s observability, availability,reliability. Design & Implement observability stack for the infrastructure - System/Application...
-
Senior Site Reliability Engineer
7 months ago
Jakarta, Indonesia AccelByte Full timeAt AccelByte, our mission is to empower game creators by providing them with the backend platform and tools required to make scalable, reliable AAA-quality games. The company was founded in 2016 by industry veterans who have engineered online systems for some of the largest game and distribution platforms in the world including Fortnite, Epic Store, Xbox...
Site Reliability Engineer
3 months ago
Paymentology is the first truly global issuer-processor, giving banks and fintechs the technology, team and experience to rapidly issue and process Mastercard, Visa and UnionPay cards across more than 50 countries, at scale.
Our advanced, multi-cloud platform, offering both shared and dedicated processing instances, vast global presence and richer, real-time data, set us apart as the leader in payments.
We're on the hunt for an exceptional **Site Reliability Engineer (SRE)** to join our dedicated team. As an SRE at Paymentology, you'll be the superhero responsible for maintaining, improving, and ensuring the high availability, scalability, and performance of our platform.
**Requirements**:
**What it takes to succeed**:
- Bachelor’s Degree in Computer Science, Information Technology, or related field.
- A minimum of 3 years in a dedicated SRE role, as well as 5+ years of prior software development experience.
- Comprehensive understanding of large-scale distributed platform architecture.
- Extensive hands-on cloud experience, particularly with AWS.
- Proven experience developing scalable, modular infrastructure-as-code projects using tools such as Terraform, CloudFormation, Puppet, and Ansible.
- Practical experience with Docker and container orchestrators, including AWS ECS & EKS, and Kubernetes.
- Experience in administering or integrating identity management systems for SSO, including AWS IAM, Okta, and Active Directory.
- Experience with disaster recovery and redundancy strategies in both cloud and on-premises environments.
- Proficiency with leading monitoring tools, such as Datadog, Honeycomb.io, Splunk, Prometheus, Grafana, ELK Stack, and New Relic.
- Programming expertise, especially in systems programming languages (e.g., Java, Kotlin, Scala) and databases (e.g., SQL Server, PostgreSQL).
- Familiarity with industry-leading CI/CD tools such as Jenkins, GitHub Actions, Gitlab CI, CodePipelines, CircleCI, and ArgoCD.
- Track record of achieving platform-level and end-to-end SLIs, SLOs, and SLAs, and fostering accountability.
- Ability to navigate complex situations and lead effective post-incident reviews (PIRs).
- Knowledge of implementing solutions to reduce Mean Time to Identify (MTTI) and Mean Time to Resolve (MTTR).
- Expertise in implementing best practices for load balancing, fault tolerance, and resource allocation to maintain service quality and efficiency at scale.
- Understanding of security best practices within cloud environments.
****You'll also need to bring a collaborative mindset, working seamlessly across teams to drive innovative solutions. And of course, your exceptional communication skills in English will allow you to clearly convey your ideas and recommendations.
As a key member of our technical team, you will be expected to maintain high availability and be ready to address critical incidents, ensuring the continuous performance of our systems. This includes being part of an on-call schedule to support 24/7 operations.
**Why Paymentology?**
- Full-time remote position with flexible hours.
- An inclusive and supportive work environment that values diversity.
- A chance to work on cutting-edge technology projects that make a difference.
- Opportunities for continuous learning and development.
What you get to do:
**Platform Reliability and Scalability**:
- Build software that enhances Paymentology services' scalability and reliability.
- Ensure platform services meet required uptime and service quality levels.
- Contribute to the design of reliable cloud infrastructure and implement reusable cloud-uptime components as code.
- Regularly review and optimise SRE practices, tools, and methodologies to enhance overall system reliability and team efficiency.
**Observability and Automation**:
- Contribute to the design, implementation, and maintenance of observability and monitoring solutions to track the platform health, its cost-effectiveness, the reliability, and scalability, and identify potential issues which can be fed back to product and platform engineering in a continuous improvement loop.
- Develop and implement automation scripts and tools to streamline operations and reduce manual interventions.
- Enable product teams to self-serve by participating in the development of a developer platform.
**Production Issue Resolution**:
- Play an active role with the incident response teams, diagnosing and resolving production issues quickly to minimise downtime.
**Standards Compliance**:
- Support product teams in building services that adhere to our security and quality standards.
**Cross-team Collaboration**:
- Work closely with engineering, operations, and product teams to ensure reliability is considered throughout the end-to-end software development lifecycle. We seek to achieve this through advocacy and developing a culture of reliability.
What you can look forward to:
At Paymentology we value making a difference to the lives of the people who work for us and