data-structures - 哈希表 : why size should be prime?-6ren

data-structures - 哈希表 : why size should be prime?

转载作者：行者123 更新时间：2023-12-03 10:41:47

25

4

这个问题在这里已经有了答案:

10年前关闭。

Possible Duplicate:
Why should hash functions use a prime number modulus?

为什么哈希表(数据结构)的大小必须是素数？

据我了解，它确保了更均匀的分布，但还有其他原因吗？

最佳答案

唯一的原因是避免将值聚集到少量桶中(是的，分布)。更均匀分布的哈希表将更一致地执行。
来自 http://srinvis.blogspot.com/2006/07/hash-table-lengths-and-prime-numbers.html

If suppose your hashCode function results in the following hashCodes among others {x , 2x, 3x, 4x, 5x, 6x...}, then all these are going to be clustered in just m number of buckets, where m = table_length/GreatestCommonFactor(table_length, x). (It is trivial to verify/derive this). Now you can do one of the following to avoid clustering

Make sure that you don't generate too many hashCodes that are multiples of another hashCode like in {x, 2x, 3x, 4x, 5x, 6x...}.But this may be kind of difficult if your hashTable is supposed to have millions of entries.

Or simply make m equal to the table_length by making GreatestCommonFactor(table_length, x) equal to 1, i.e by making table_length coprime with x. And if x can be just about any number then make sure that table_length is a prime number.

更新: (来自原始答案作者)
这个答案对于哈希表的常见实现是正确的，包括原始 Hashtable 的 Java 实现。以及 .NET 的当前实现 Dictionary .
Java 的 HashMap 的答案和容量应该是素数的假设都不准确。尽管。 HashMap的实现非常不同，它使用一个大小为 2 的表来存储桶并使用 n-1 & hash计算使用哪个桶而不是更传统的 hash % n公式。
Java的 HashMap将强制实际使用的容量为请求容量之上的下一个最大基数 2 数。
比较 Hashtable :

int index = (hash & 0x7FFFFFFF) % tab.length

https://github.com/openjdk/jdk/blob/jdk8-b120/jdk/src/share/classes/java/util/Hashtable.java#L364
致 HashMap :

first = tab[(n - 1) & hash]

https://github.com/openjdk/jdk/blob/jdk8-b120/jdk/src/share/classes/java/util/HashMap.java#L569

关于data-structures - 哈希表 : why size should be prime?，我们在Stack Overflow上找到一个类似的问题： https://stackoverflow.com/questions/3980117/

25

4

0

文章推荐： 64-bit - 什么是 16、32 和 64 位架构？

文章推荐： django 拒绝识别静态文件夹？

文章推荐： sha256 - (比特币)从 getwork 函数计算哈希 - 怎么做？

文章推荐： visual-studio-2008 - 选择哪一个？ DXCore、Resharper 还是 VSX？

size - ValueError : Target size (torch. Size([16])) 必须与输入大小相同 (torch.Size([16, 1]))
ValueError Traceback (most recent call last) in 23 out
CSS Percent size specifier sizing element to more than specified size
在 CSS 中，我从来没有真正理解为什么会发生这种情况，但每当我为某物分配 margin-top:50% 时，该元素就会被推到页面底部，几乎完全消失这一页。我假设 50% 时，该元素将位于页面的中间位
neural-network - ValueError : Target size (torch. Size([1000])) must be the same as input size (torch.Size([1000, 1]))
我正在尝试在 pyTorch 中训练我的第一个神经网络(我不是程序员，只是一个困惑的化学家)。网络本身应该采用 1064 个元素向量并用 float 对它们进行评级。到目前为止，我遇到了各种各样的
c# - 数组移位/错误索引/i = [x+y*size+z*size*size]
我有一个简单的问题。如何在 3 个维度上移动线性阵列？这似乎太有效了，但在 X 和 Y 轴上我遇到了索引问题。我想这样做的原因很简单。我想创建一个带有 block 缓冲区的体积地形，所以我只需要在视口
python - 如何解决与输入大小 (torch.Size([1])) 不同的 UserWarning : Using a target size (torch. Size([]))？
我正在尝试运行我购买的一本关于 Pytorch 强化学习的书中的代码。代码应该按照本书工作，但对我来说，模型没有收敛，奖励仍然为负。它还会收到以下用户警告: /home/user/.local/li
python - PyTorch ValueError : Target size (torch. Size([64])) 必须与输入大小相同 (torch.Size([15]))
我目前正在使用 this repo使用我自己的数据集执行 NLP 并了解有关 CNN 的更多信息，但我一直遇到有关形状不匹配的错误: ValueError: Target size (torch.Si
objective-c - UIScrollView.size = view.size - allAdditionalBars.size(如 TabBar 或 NavigationBar)以编程方式
UIScrollView 以编程方式设置，请不要使用 .xib 文件发布答案。我的 UIScrollView 位于我的模型类中，所以我希望代码能够轻松导入到另一个项目中，例如。适用于 iPad 或旋
css - Bootstrap 4 : How Can I Set $font-size-base for Different Monitor Sizes using Responsive Font Sizing?
我在我的 Ruby on Rails 应用程序(版本 4.3.1)中使用 Bootstrap gem。我最近发现了响应式字体大小功能 (rfs)。根据 Bootstrap 文档，它刚刚在 4.3 版中
Android App开发错误: "Bad XML block: header size 60 or total size 3932356 is larger than data size 0"
这个问题不太可能帮助任何 future 的访客；它仅与一个小地理区域、一个特定时刻或一个非常狭窄的情况相关，而这些情况通常不适用于互联网的全局受众。如需帮助使这个问题更广泛地适用，visit the
scala - size 和 size 的区别是
size 之间的语义区别是什么？和 sizeIs ?例如， List(1,2,3).sizeIs > 1 // true List(1,2,3).size > 1 // true Luis 在 c
javascript - 从子元素中删除 Size 和 font-size
我想从 div 中删除一些元素属性。我的 div 是自动生成的。我想遍历每个 div 和子 div，并想删除所有 font-size (font-size: Xpx)和 size里面font tag
python - 使用 self.size = size 时语法无效
super ，对 Python 和一般编程 super 新手。我有一个问题应该很简单。我正在使用一本使用 Python 3.1 版的 python 初学者编程书。我目前正在写书中的一个程序，我正在学
size - native 库 : change thumbnail default size
我无法从 NativeBase 更改缩略图的默认大小。我可以显示默认圆圈，即小圆圈和大圆圈，但我想显示比默认大小更大的圆圈。这是我的缩略图代码: Prop 大小不起作用，缩略图仍然很小。我的 Na
pytorch - pytorch中张量torch.Size([])和torch.Size([1])的形状差异
我是pytorch的新手。在玩张量时，我观察到了两种类型的张量- tensor(58) tensor([57.3895]) 我打印了它们的形状，输出分别是 - torch.Size([]) torch
Docker 镜像 : virtual size vs real size
这是我的 docker images 命令的输出: $ docker images REPOSITORY TAG IMAGE ID CREATED
java - 为什么使用 "s = --size"而不是 "s = size"？
来自 PriorityQueue 的代码: private E removeAt(int i) { assert i >= 0 && i < size; modCount++;
c++ - sizeof() : the size of a class isn't the same as the size of it's members together?
首先，在我的系统上保留以下内容:sizeof(char) == 1 和 sizeof(char*) == 4。很简单，当我们计算下面类的总大小时: class SampleClass { char c
iphone - cocos2d content.size、boundingBox 和 size
我正在编写一个游戏来查找 2 个图像之间的差异。我创建了 CCSprite 的子类 Spot。首先我尝试创建小图像并根据其位置添加自身，但后来我发现位置很难确定，因为很难避免 1 或 2 个像素的偏移
javascript - Tumblr:photoUrl-(size) - size depending on class？
我有一个 Tumblr Site每个帖子的宽度由标签决定。如果一篇文章被标记为 #width200，CSS 类 .width200 被分配。问题是，虽然帖子的宽度不同，但它们都使用主题运算符加载相
c++ - 为什么动态分配的数组大小在插入时是初始数组的 2*size，而不是 size+1？
这个问题在这里已经有了答案: What is the ideal growth rate for a dynamically allocated array? (12 个答案) 关闭 8 年前。我

首页

博学

6Ren·AI

商城

data-structures - 哈希表 : why size should be prime?