This video demonstrates how to build your first Spark ETL pipeline with StreamSets' Transformer Engine. Design your pipeline in a simple, visual canvas and add any of 100+ pre-built processors in your pipeline. This specific pipeline starts with two origins ingesting data and includes an inner join and partition processor. Then, the data is stored in its destination in Parquet format.
Try StreamSets now: https://streamsets.com/try-dataops/?u...
Learn more about StreamSets: https://streamsets.com/products/datao...