Do you monitor your Postgres error logs for gold? Lukas Fittl and Rob Treat join Claire and Pino on the Talking Postgres podcast (formerly called Path To Citus Con*)—to discuss their respective journeys into Postgres monitoring. Have you ever asked yourself: “Why is my query so slow?” Or had to figure out which query is slowing things down? Or why your database server is at 90% CPU? There are so many ways to monitor Postgres: pganalyze, pgMustard, pgBadger, pgDash, your cloud provider’s Query Performance Insights, pg_stat_statements, pg_stat_io, & more. If you’re running Postgres on a managed service, what kinds of things do you need to monitor & optimize for (vs. what will your cloud service provider do)? There’s also a segue on monitoring vs. observability: what’s the difference?
*[Update: July 2024] Path To Citus Con has been renamed to Talking Postgres. All of the past podcast episodes from Path To Citus Con—now called Talking Postgres with Claire Giordano—can be found here: https://talkingpostgres.com and on this YouTube playlist: https://aka.ms/TalkingPostgres-playlist
Guests Lukas Fittl and Rob Treat:
Lukas Fittl is a serial entrepreneur and founder of pganalyze, where he empowers developers at companies like Atlassian and DoorDash to do their best work with a powerful product that enables them to deliver consistent Postgres performance and availability through intelligent tuning advisors and continuous database profiling.
Rob Treat is a former scalability practitioner and devopsdays organizer—his consulting clients included Etsy, Gilt, Doordash, National Geographic, 3/5 of FAANG, several monitoring companies, and many more. He is currently semi-retired but enjoys helping people with Postgres.
Chapters:
⏩ 00:00 Intro & origin of pganalyze
⏩ 6:56 Circonus and monitoring
⏩ 15:10 Monitoring vs. observability
⏩ 21:12 Monitoring is for known unknowns?
⏩ 24:31 Slow queries: major pain point
⏩ 30.23 How should people think about monitoring?
⏩ 31:52 Top 5 monitoring tools
⏩ 52:17 pg_stat_io on Postgres16
⏩ 1:01:07 Cloud vs. on-prem monitoring
⏩ 1:07:05 Is AI going to monitor everything?
⏩ 1:13:31 Defining pg_hint_plan
⏩ 1:16:40 Postgres mailing lists
📜 Full transcript of this podcast episode available at:
https://talkingpostgres.com/episodes/...
✅ Listen to more episodes of Talking Postgres:
https://talkingpostgres.com
💥 Subscribe to Talking Postgres, so you never miss an episode:
https://talkingpostgres.com/subscribel
Links mentioned in this episode:
🔹OpenTelemetry: https://opentelemetry.io/
🔹pganalyze: https://pganalyze.com/
🔹pgDash: https://pgdash.io/
🔹pgMustard: https://www.pgmustard.com/
🔹pg_stat_statements docs: https://www.postgresql.org/docs/curre...
🔹pg_hint_plan: https://github.com/ossc-db/pg_hint_plan
🔹pg_hint_plan hint list: https://github.com/ossc-db/pg_hint_pl...
🔹Example for PostgreSQL with pg_hint_plan: https://api.rubyonrails.org/classes/A...
🔹5mins of Postgres by pganalyze: • 5mins of Postgres
🔹Monitoring page on PostgreSQL wiki: https://wiki.postgresql.org/wiki/Moni...
🔹PgHero GitHub repo: https://github.com/ankane/pghero
🔹Insights on pgBadger: A PGSQL Phriday #010 Recap: https://techcommunity.microsoft.com/t...
🔹Get PostgreSQL Logs Into Honeycomb: https://docs.honeycomb.io/getting-dat...
🔹Blog post by Lukas Fittl about pg_stat_io by Lukas: https://pganalyze.com/blog/pg-stat-io\
🔹Blog post by Andrew Atkinson about pg_stat_io: https://andyatkinson.com/blog/2023/11...
🔹BPFtrace by iovisor GitHub repo: https://github.com/iovisor/bpftrace
🔹Trace PostgreSQL locks with pg_lock_tracer: https://jnidzwetzki.github.io/2023/01...
🔹sysdig by draios GitHub repo: https://github.com/draios/sysdig
🔹Using BPFtrace to trace PostgreSQL vacuum operations: https://www.timescale.com/blog/using-...
🔹PostgreSQL Mailing Lists: https://www.postgresql.org/list/
#podcast #TalkingPostgres #PostgreSQL