Articles about distributed processing DAG on waitingforcode.com

December 31, 2017 • Apache Beam

TransformHierarchy in Apache Beam

Apache Beam has some similarities with Apache Spark. One of them is the definition of processing pipeline as a Directed Acyclic Graph.

Continue Reading →

October 23, 2016 • Apache Spark

Directed Acyclic Graph in Spark

As we already know, RDD is the main data concept of Spark. It's created either explicitly or implicitly, through computations called transformations and actions. But these computations are all organized as a graph and scheduled by Spark's components. This graph is called DAG and it's the main topic of this post.

Continue Reading →

distributed processing DAG articles

TransformHierarchy in Apache Beam

Directed Acyclic Graph in Spark