python - Pandas 时间序列 : avg of a timestamp column-6ren

python - Pandas 时间序列 : avg of a timestamp column

转载作者：太空宇宙更新时间：2023-11-03 14:44:32

27

4

我有一个数据框，看起来像这样:

ID      Date
16911   2017-04-15
16911   2017-04-25
16911   2017-04-27
16911   2017-05-08
16911   2017-05-20
16911   2017-05-25
16911   2017-08-08
16911   2017-08-11
16911   2017-08-24
16912   2017-04-15
16912   2017-04-25
16812   2017-04-27
16812   2017-05-08
16812   2017-05-20
16812   2017-05-25
16812   2017-08-08
16812   2017-08-11

日期已排序，我想找到时间戳之间的差异并找到每个 ID 的平均值。

还有，

假设 ID - 16911，我想要例如 -> 列表 a 的日期差异列表；

16911   2017-04-15
16911   2017-04-25
difference between the above two dates is 10, so a is
a = [10]

16911   2017-04-25
16911   2017-04-27
difference between the above two dates is 2, so a is
a=[10,2]

16911   2017-04-27
16911   2017-05-08
difference between the above two dates is 11(assuming), so a is
a=[10,2,11]

所以最终的输出应该是:

ID      Average_Day Diff
16911   3 days      [10,2,11]

最佳答案

使用groupby与 diff和均值:

df = df.groupby('ID')['Date'].apply(lambda x: x.diff().mean()).reset_index()
print (df)
      ID             Date
0  16812 21 days 04:48:00
1  16911 16 days 09:00:00
2  16912 10 days 00:00:00

如果需要转换时间增量，例如到 天:

df = df.groupby('ID')['Date'].apply(lambda x: x.diff().mean().days).reset_index()
print (df)
      ID  Date
0  16812    21
1  16911    16
2  16912    10

编辑:

#create difference column per ID
df['new'] = df.groupby('ID')['Date'].diff().dt.days
#remove NaT rows (first for each group)
df = df.dropna(subset=['new'])
#convert to integers
df['new'] = df['new'].astype(int)
#aggreagte lists and mean
df = df.groupby('ID', sort=False)['new'].agg([('val', lambda x: x.tolist()),('avg', 'mean')])
print (df)

ID                                          
16911  [10, 2, 11, 12, 5, 75, 3, 13]  16.375
16912                           [10]  10.000
16812             [11, 12, 5, 75, 3]  21.200

关于python - Pandas 时间序列 : avg of a timestamp column，我们在Stack Overflow上找到一个类似的问题： https://stackoverflow.com/questions/50795491/

27

4

0

文章推荐： python - 如何从类方法外部访问类方法内部的变量？

文章推荐： jquery - $.ajax({}) 不是 django 中的函数

文章推荐： java - SSL SSLv2 客户端问候 - 握手失败

timestamp - KnexJS : How do you insert/update a timestamp field with current timestamp?
标题基本上说明了一切。我主要对更新案例感兴趣。假设我们正在尝试更新具有时间戳记字段的记录，并且我们希望将该字段设置为记录更新的时间戳记。有没有办法做到这一点？最佳答案经过一些实验，我找到了合适的
python - 'Timestamp' 对象没有属性 'timestamp'
我正在学习一门类(class)，其中我必须将日期转换为 unix 时间戳。 import pandas as pd df = pd.read_csv('file.csv') print type(df
sql - TIMESTAMP、TIMESTAMP with TIME ZONE 和 TIMESTAMP with LOCAL TIME ZONE 之间的区别
我在两个不同的数据库中运行了相同的语句:我的本地数据库和 Oracle Live SQL . CREATE TABLE test( timestamp TIMESTAMP DEFAULT SY
sql - TIMESTAMP、TIMESTAMP with TIME ZONE 和 TIMESTAMP with LOCAL TIME ZONE 之间的区别
我在两个不同的数据库中运行了相同的语句:我的本地数据库和 Oracle Live SQL . CREATE TABLE test( timestamp TIMESTAMP DEFAULT SY
python - bson.timestamp.Timestamp - 递增计数器是什么？
bson.timestamp.Timestamp需要两个参数:time 和 inc。 time 显然是存储在 Timestamp 中的时间值。什么是公司？它被描述为递增计数器，但它有什么用途呢？它应
php - 查询 where timestamp < timestamp 不起作用？
2016-08-18 04:52:14 是我从数据库中获取的时间戳，用于跟踪我想从哪里加载更多记录，这些记录小于该时间这是代码 foreach($explode as $stat){
timestamp - 如何转换Erlang :timestamp() to normal date format?
我想将 erlang:timestamp() 的结果转换为正常的日期类型，公历类型。普通日期类型表示“日-月-年，时:分:秒”。 ExampleTime = erlang:timeStamp(),
timestamp - 如何转换Erlang :timestamp() to normal date format?
我想将 erlang:timestamp() 的结果转换为正常的日期类型，公历类型。普通日期类型表示“日-月-年，时:分:秒”。 ExampleTime = erlang:timeStamp(),
java - 将 Timestamp 与另一个 Timestamp 对象进行比较
我是 Java 新手。我正在使用两个 Timestamp 对象 dateFrom和dateTo 。我想检查是否dateFrom比 dateTo早 45 天。我用这个代码片段来比较这个 if(dateF
python - 属性错误 : 'Timestamp' object has no attribute 'timestamp
在将 panda 对象转换为时间戳时，我遇到了这个奇怪的问题。 Train['date'] 值类似于 01/05/2014，我正在尝试将其转换为 linuxtimestamp。我的代码: Train
python - 属性错误 : 'Timestamp' object has no attribute 'timestamp'
我正在努力让我的代码运行。时间戳似乎有问题。您对我如何更改代码有什么建议吗？我看到之前有人问过这个问题，但没能成功。这是我在运行代码时遇到的错误:'Timestamp' object has no
timestamp - AWS 雅典娜 SYNTAX_ERROR : not a valid timestamp literal
我正在尝试运行以下查询: SELECT startDate FROM tests WHERE startDate BETWEEN TIMESTAMP '1555248497'
sql - 亚马逊雅典娜 : Convert bigint timestamp to readable timestamp
我正在使用 Athena 查询以 bigInt 格式存储的日期。我想将其转换为友好的时间戳。我试过了: from_unixtime(timestamp DIV 1000) AS readab
sql-server - SQLServer异常: The conversion from timestamp to TIMESTAMP is unsupported.
最近进行了一些数据库更改，并且 hibernate 映射出现了一些困惑。 hibernate 映射: ...other fields 成员模型对象: public class Mem
Pandas : How to get timestamp. 天和 timestamp.month 填充零
rng = pd.date_range('2016-02-07', periods=7, freq='D') print(rng[0].day) print(rng[0].month) 7 2 我想要
Pandas : How to get timestamp. 天和 timestamp.month 填充零
rng = pd.date_range('2016-02-07', periods=7, freq='D') print(rng[0].day) print(rng[0].month) 7 2 我想要
Android - Firebase ServerValue.TIMESTAMP 返回 "{.sv=timestamp}"
我必须在我的数据库中保存 ServerValue.TIMESTAMP 但它必须是一个字符串。当我键入 String.valueOf(ServerValue.TIMESTAMP); 或 ServerVa
PostgreSQL select now()::timestamp 不同于默认的 now()::timestamp
在我的程序中，每个表都有一列 last_modified: last_modified int8 DEFAULT (date_part('epoch'::text, now()::timestamp)
python - 将 pandas._libs.tslibs.timestamps.Timestamp 转换为日期时间
我想将此时间戳对象转换为日期时间此对象是在数据帧上使用 asfreq 后获得的这是最后一个索引 Timestamp('2018-12-01 00:00:00', freq='MS') 想要的输出 2
mysql - 如何查找上一条记录[n-per-group max(timestamp) < timestamp]？
我有一个包含时间序列传感器数据的大表。大型是指分布在被监控的各个 channel 中的从几千到 10M 的记录。对于某种传感器类型，我需要计算当前读数和上一个读数之间的时间间隔，即找到当前读数之前的最

首页

博学

6Ren·AI

商城

python - Pandas 时间序列 : avg of a timestamp column