Optimizing Apache Druid for query performance involves creating segment files of the right size and organizing the data in preparation for efficient read access. Secondary Partitioning is one of the tools on your belt to achieve this.
This video by @sergioferragut7232 explains how Apache Druid partitions data at ingestion time using a ranged secondary partitioning strategy.
00:00 - Intro,
00:23 - Ranged Partitioning Benefits
02:35 - How it Works
09:13 - Demo
14:04 - Summary of Ranged Partitioning
Learn More:
Check out the Apache Druid docs on Partitioning here: https://druid.apache.org/docs/latest/...
Get hands-on training for free at: https://learn.imply.io/
Join the Apache Druid workspace on Slack:
https://druid.apache.org/community/jo...
Connect:
Subscribe: / implydata
Imply GitHub: https://github.com/implydata
Apache Druid GitHub: https://github.com/apache/druid
Twitter: / implydata
LinkedIn: / imply
About Imply
Developers are in the driver’s seat when it comes to analytics, building applications that serve real-time insights on terabytes to petabytes of streaming and batch data at hundreds to thousands of queries per second.
With Imply, developers have a database that is uniquely built for these analytics applications, delivering sub-second queries at scale and under load. The result? No spinning wheel and no limit to the analytics in their applications.
Check us out at https://imply.io/