@yurynino Ingeniera de Sistemas. Especialista en Ingeniería de Software Coorganizadora de GDG Bogotá, GDG Cloud Bogotá y Women Techmakers Bogotá “A single point of failure can trigger a chain reaction across the value chain, including suppliers to the final customer, and cause a severe business interruption” Dimitar Pachov
High Availability? Apache Spark is Highly Available The promise: Resilience Patterns Chaos Engineering Chaos Principles … Testing High Availability with Chaos!
… 38% of customers at traditional banks experienced disruption to their service every year, compared with 21% of challenger bank customers. The FCA says bank outages have risen 138% in the past year in the world!
data processing. • Spark is available for both batch and streaming data. • Spark allows to write applications in Java, Scala, Python, and SQL. • Spark makes easy to build parallel apps. • Spark combines SQL, streaming, and complex analytics. • Spark runs on Hadoop, Apache Mesos, Kubernetes, standalone, or in the cloud.
able to restart the Master if it goes down, Filesystem mode can take care of it. When Applications and Workers register, they have enough state written to the provided directory so that they can be recovered upon a restart of the Master.
storage, you can launch multiple Masters in your cluster connected to the same ZooKeeper instance. One will be elected “leader” and the others will remain in standby mode. If the current leader dies, another Master will be elected, recover the old Master’s state, and then resume scheduling.
I do What software engineers think I do What I really do Who is a Chaos Engineer? Help service owners to increase their resilience through education, tools and encouragement.
metrics when the different Spark components fail. To simulate such failures, we employed a whack-a-mole approach and killed the various Spark components.