Home > mailing lists

Re: Number of buckets in a hash join - Mailing list pgsql-hackers

From	Simon Riggs
Subject	Re: Number of buckets in a hash join
Date	January 28, 2013 17:30:43
Msg-id	CA+U5nM+WJaQh3HBEVi-4n+rxu-NUqn0RGZ-gjy9S-zYMvhaWHw@mail.gmail.com Whole thread Raw
In response to	Number of buckets in a hash join (Heikki Linnakangas <hlinnakangas@vmware.com>)
List	pgsql-hackers

Tree view

On 28 January 2013 10:47, Heikki Linnakangas <hlinnakangas@vmware.com> wrote:

> There's also some overhead from empty
> buckets when scanning the hash table

Seems like we should measure that overhead. That way we can plot the
cost against number per bucket, which sounds like it has a minima at
1.0, but that doesn't mean its symmetrical about that point. We can
then see where the optimal setting should be.

Having said that the hash bucket estimate is based on ndistinct, which
we know is frequently under-estimated, so it would be useful to err on
the low side.

-- Simon Riggs                   http://www.2ndQuadrant.com/PostgreSQL Development, 24x7 Support, Training & Services

pgsql-hackers by date:

From: Kevin Grittner
Date: 28 January 2013, 17:28:49
Subject: Re: "pg_ctl promote" exit status

From: Peter Eisentraut
Date: 28 January 2013, 17:46:37
Subject: Re: "pg_ctl promote" exit status

Re: Number of buckets in a hash join - Mailing list pgsql-hackers

Previous

Next