multithreading - Python : multithreaded learning neural networks using PyBrain and Multiprocessing-6ren

multithreading - Python : multithreaded learning neural networks using PyBrain and Multiprocessing

转载作者：行者123 更新时间：2023-12-03 13:10:06

32

4

我正在尝试使用PyBrain和Python的multiprocessing软件包在Python中训练神经网络。

这是我的代码(它训练了一个简单的神经网络来学习XOR逻辑)。

import pybrain.tools.shortcuts as pybrain_tools
import pybrain.datasets
import pybrain.supervised.trainers.rprop as pybrain_rprop
import multiprocessing
import timeit


def init_XOR_dataset():
    dataset = pybrain.datasets.SupervisedDataSet(2, 1)
    dataset.addSample([0, 0], [0])
    dataset.addSample([0, 1], [1])
    dataset.addSample([1, 0], [1])
    dataset.addSample([1, 1], [0])
    return dataset


def standard_train():
    net = pybrain_tools.buildNetwork(2, 2, 1)
    net.randomize()
    trainer = pybrain_rprop.RPropMinusTrainer(net, dataset=init_XOR_dataset())
    trainer.trainEpochs(50)


def multithreaded_train(threads=8):
    nets = []
    trainers = []
    processes = []
    data = init_XOR_dataset()

    for n in range(threads):
        nets.append(pybrain_tools.buildNetwork(2, 2, 1))
        nets[n].randomize()
        trainers.append(pybrain_rprop.RPropMinusTrainer(nets[n], dataset=data))
        processes.append(multiprocessing.Process(target=trainers[n].trainEpochs(50)))
        processes[n].start()

    # Wait for all processes to finish
    for p in processes:
        p.join()


if __name__ == '__main__':
    threads = 4
    iterations = 16

    t1 = timeit.timeit("standard_train()",
                       setup="from __main__ import standard_train",
                       number=iterations)
    tn = timeit.timeit("multithreaded_train({})".format(threads),
                       setup="from __main__ import multithreaded_train",
                       number=iterations)

    print "Execution time for single threaded training: {} seconds.".format(t1)
    print "Execution time for multi threaded training: {} seconds.".format(tn)

在我的代码中，有两个功能:一个运行单线程，另一个(据说)运行使用多处理程序包的多线程。

据我判断，我的多处理代码是正确的。但是，当我运行它时，多处理代码不会在一个以上的内核上运行。我通过检查运行时间来验证了这一点(使用线程= 4和4个内核时，它花费的时间是原来的4倍，而执行一次线程所需的时间大约是它的4倍)。我通过查看 htop/ atop对其进行了仔细检查。

我知道 Global Interpreter Lock (GIL)，但是应该由多处理程序包来处理。

我也知道 issue that scipy causes the cpu affinity to be set in such a way that only one core is used。但是，如果在scipy导入PyBrain包( print psutil.Process(os.getpid()).cpu_affinity())后立即打印过程亲和力，我可以看到亲和力还可以:

$ python ./XOR_PyBrain.py
[0, 1, 2, 3]
Execution time for single threaded training: 14.2865240574 seconds.
Execution time for multi threaded training: 46.0955679417 seconds.

我在Debian NAS，Debian Desktop和Mac上都观察到了这种现象。

Debian NAS的版本信息:

CPU:Intel(R)Atom(TM)CPU D2701 @ 2.13GHz

Debian 8.4。内核:3.2.68-1 + deb7u1 x86_64

Python 2.7.9

PyBrain 0.3

Scipy 0.14.0-2

因此，我的问题是:如何让PyBrain在多个内核上训练？

最佳答案

一切似乎都是正确的。我已经测试了您的代码，并且所有内核都在工作。时序问题来自于以下事实:“当您尝试处理少量数据时，CPU内核无法实现其填充潜力。”

您建立的网络非常简单，因此在每个内核获得100％使用率之前，该过程将结束。就像想处理数百万个5x5像素的图像一样。您的计算机等待硬盘驱动器获取数据的时间远远超过CPU处理数据所花费的时间。

一些去这里。但是在更快的区域(RAM)和更少的数据量(十进制数)中。如果您的计算机有DDR4 RAM，情况可能会发生变化。 (但我真的不这么认为)

尝试更大的过程，更大的网络，更多数据等，您将看到期望的结果。

关于multithreading - Python : multithreaded learning neural networks using PyBrain and Multiprocessing，我们在Stack Overflow上找到一个类似的问题： https://stackoverflow.com/questions/37706743/

32

4

0

文章推荐： reactjs - 在React中，如何用逗号格式化数字？

文章推荐： twitter-bootstrap - Bootstrap : fluid table too wide for window

文章推荐： java - Java生产者使用者ArrayBlockingQueue在take()上发生死锁

MySQL相似表的索引使用不一致: 'Using index' and 'Using where; Using index'
我在优化 JOIN 以使用复合索引时遇到问题。我的查询是: SELECT p1.id, p1.category_id, p1.tag_id, i.rating FROM products p1
sql - 优化查询以删除 "Using where; Using temporary; Using filesort"
我有一个简单的 SQL 查询，我正在尝试对其进行优化以删除“使用位置；使用临时；使用文件排序”。这是表格: CREATE TABLE `special_offers` ( `so_id` int
mysql - EXPLAIN 语句说 'Using where; Using index' 如果在查询中设置了 USE INDEX() 否则就说 'Using where'
我有一个具有以下结构的应用程序表 app_id VARCHAR(32) NOT NULL, dormant VARCHAR(6) NOT NULL, user_id INT(10) NOT NULL
mysql - Extra :-Using where; Using temporary; Using filesort如何优化MYSQL
此查询的正确索引是什么。我尝试为此查询提供不同的索引组合，但它仍在使用临时文件、文件排序等。总表数据 - 7,60,346 产品= '连衣裙' - 总行数 = 122 554 CREATE TAB
mysql - 为什么额外的是 "using where;using index"而不是 "using index"
为什么额外的是“使用where;使用索引”而不是“使用索引”。 CREATE TABLE `pre_count` ( `count_id`
按日期排序时，MySQL 数据库使用 "Using where; Using temporary; Using filesort"
我有一个包含大量记录的数据库，当我使用以下 SQL 加载页面时，速度非常慢。 SELECT goal.title, max(updates.date_updated) as update_sort F
MySQL - 'Using index condition' 与 'Using where; Using index'
我想知道 Using index condition 和 Using where 之间的区别；使用索引。我认为这两种方法都使用索引来获取第一个结果记录集，并使用 WHERE 条件进行过滤。 Q1。有什
Cannot setup TypeScript to use `using` keyword(无法将TypeScript设置为使用“using”关键字)
I am using TypeScript 5.2 version, I have following setup:我使用的是TypeScript 5.2版本，我有以下设置： { "
Cannot setup TypeScript to use `using` keyword(无法将TypeScript设置为使用“using”关键字)
I am using TypeScript 5.2 version, I have following setup:我使用的是TypeScript 5.2版本，我有以下设置： { "
Cannot setup TypeScript to use `using` keyword(无法将TypeScript设置为使用“using”关键字)
I am using TypeScript 5.2 version, I have following setup:我使用的是TypeScript 5.2版本，我有以下设置： { "
mysql - 如何避免MySQL中的 "Using index; Using temporary; Using filesort "，21表JOIN
mysql Ver 14.14 Distrib 5.1.58，用于使用 readline 5.1 的 redhat-linux-gnu (x86_64) 我正在接手一个旧项目。我被要求加快速度。我通过
mysql - OrmLite(服务堆栈): Only use temporary db-connections (use 'using' ?)
在过去 10 多年左右的时间里，我一直打开数据库 (mysql) 的连接并保持打开状态，直到应用程序关闭。所有查询都在连接上执行。现在，当我在 Servicestack 网页上看到示例时，我总是看到
sql - 优化 MySQL 查询以避免 "Using where; Using temporary; Using filesort"
我使用 MySQL 为我的站点构建了一个自定义论坛。列表页面本质上是一个包含以下列的表格:主题、上次更新和# Replies。数据库表有以下列: id name body date topic_id
mysql - EXPLAIN中的 "Using index"和 "Using where; Using index"有什么区别
在mysql中解释的额外字段中你可以得到: 使用索引使用where;使用索引两者有什么区别？为了更好地解释我的问题，我将使用下表: CREATE TABLE `test` ( `id` bi
using - Haxe中的 `using`关键字是什么？
我经常看到人们在其Haxe代码中使用关键字using。它似乎在import语句之后。例如，我发现这是一个代码片段: import haxe.macro.Context; import haxe.ma
克洛尤尔 : how do I use use "and" in "reduce"?
这个问题在这里已经有了答案: "reduce" or "apply" using logical functions in Clojure (2 个答案) 关闭 8 年前。 “and”似乎是一个宏，
克洛尤尔 : how do I use use "and" in "reduce"?
这个问题在这里已经有了答案: "reduce" or "apply" using logical functions in Clojure (2 个答案) 关闭 8 年前。 “and”似乎是一个宏，
c++ - 注册表模式 : to use or not to use
我正在考虑在我的应用程序中使用注册表模式来存储指向某些应用程序窗口和 Pane 的弱指针。应用程序的一般结构如下所示。该应用程序有一个 MainFrame 顶层窗口，其中有几个子 Pane 。可以有
When to use == and when to use is?(什么时候使用==，什么时候使用IS？)
奇怪的是：。似乎a是b或多或少被定义为id(A)==id(B)。用这种方式制造错误很容易：。有些名字出人意料地出现在Else块中。解决方法很简单，我们应该使用ext==‘.mp3’，但是如果ext表面
mysql - 优化 'GROUP BY'-查询，消除 'Using where; Using temporary; Using filesort'
我遇到了一个我似乎无法解决的 MySQL 问题。为了能够快速执行用于报告目的的 GROUP BY 查询，我已经将几个表非规范化为以下内容(该表由其他表上的触发器维护，我已经同意了与此): DROP T

首页

博学

6Ren·AI

商城

multithreading - Python : multithreaded learning neural networks using PyBrain and Multiprocessing