MurmurHash - what is it?

前端 未结 2 874
我寻月下人不归
我寻月下人不归 2020-12-04 05:54

I\'ve been trying to get a high level understanding of what MurmurHash does.

I\'ve read a basic description but have yet to find a good explanation of when to use i

相关标签:
2条回答
  • 2020-12-04 06:11

    Murmur is a family of good general purpose hashing functions, suitable for non-cryptographic usage. As stated by Austin Appleby, MurmurHash provides the following benefits:

    • simple (in term of number of generated assembly instructions).
    • good distribution (passing chi-squared tests for practically all keysets & bucket sizes.
    • good avalanche behavior (max bias of 0.5%).
    • good collision resistance (passes Bob Jenkin's frog.c torture-test. No collisions possible for 4-byte keys, no small (1- to 7-bit) differentials).
    • great performance on Intel/AMD hardware, good tradeoff between hash quality and CPU consumption.

    You can certainly use it to hash UUIDs (like any other advanced hashing functions: CityHash, Jenkins, Paul Hsieh's, etc ...). Now, a Redis bitset is limited to 4 GB bits (512 MB). So you need to reduce 128 bits of data (UUID) to 32 bits (hashed value). Whatever the quality of the hashing function, there will be collisions.

    Using an engineered hash function like Murmur will maximize the quality of the distribution, and minimize the number of collisions, but it offers no other guarantee.

    Here are some links comparing the quality of general purpose hash functions:

    http://www.azillionmonkeys.com/qed/hash.html

    http://www.strchr.com/hash_functions

    http://blog.aggregateknowledge.com/2011/12/05/choosing-a-good-hash-function-part-1/

    http://blog.aggregateknowledge.com/2011/12/29/choosing-a-good-hash-function-part-2/

    http://blog.aggregateknowledge.com/2012/02/02/choosing-a-good-hash-function-part-3/

    0 讨论(0)
  • 2020-12-04 06:15

    MurmurHash can be return negtive value, original value bit AND against 0x7fffffff。that is value & 0x7fffffff .When the input is positive, the original value is returned. When the input number is negative, the returned positive value is the original value bit AND against 0x7fffffff which is not its absolutely value. note:MurmurHash's return value can not be fix length.

    0 讨论(0)
提交回复
热议问题