将字典和聚合值数据的python列表分组

goRunToStack

我有输入清单

inlist = [{"id":123,"hour":5,"groups":"1"},{"id":345,"hour":3,"groups":"1;2"},{"id":65,"hour":-2,"groups":"3"}]

我需要按“组”值对字典进行分组。在那之后,我需要在新的分组列表中添加小时的最小值和最大值。输出应如下所示

outlist=[(1, [{"id":123, "hour":5, "min_group_hour":3, "max_group_hour":5}, {"id":345, "hour":3, "min_group_hour":3, "max_group_hour":5}]),
     (2, [{"id":345, "hour":3, "min_group_hour":3, "max_group_hour":3}])
     (3, [{"id":65, "hour":-2, "min_group_hour":-2, "max_group_hour":-2}])]

到目前为止,我设法对输入列表进行了分组

new_list = []
for domain in test:
    for group in domain['groups'].split(';'):
        d = dict()
        d['id'] = domain['id']
        d['group'] = group
        d['hour'] = domain['hour']
        new_list.append(d)

for k,v in itertools.groupby(new_list, key=itemgetter('group')):
    print (int(k),max(list(v),key=itemgetter('hour'))

和输出是

('1', [{'group': '1', 'id': 123, 'hour': 5}])
('2', [{'group': '2', 'id': 345, 'hour': 3}])
('3', [{'group': '3', 'id': 65, 'hour': -2}])

我不知道如何按组汇总值?还有是否需要按键值对字典进行分组的更多Python方式?

阿兰·费

首先创建一个将组号映射到字典的字典:

from collections import defaultdict

dicts_by_group = defaultdict(list)
for dic in inlist:
    groups = map(int, dic['groups'].split(';'))
    for group in groups:
        dicts_by_group[group].append(dic)

这给了我们一个看起来像

{1: [{'id': 123, 'hour': 5, 'groups': '1'},
     {'id': 345, 'hour': 3, 'groups': '1;2'}],
 2: [{'id': 345, 'hour': 3, 'groups': '1;2'}],
 3: [{'id': 65, 'hour': -2, 'groups': '3'}]}

然后遍历分组的字典min_group_hourmax_group_hour为每个组设置

outlist = []
for group in sorted(dicts_by_group.keys()):
    dicts = dicts_by_group[group]
    min_hour = min(dic['hour'] for dic in dicts)
    max_hour = max(dic['hour'] for dic in dicts)

    dicts = [{'id': dic['id'], 'hour': dic['hour'], 'min_group_hour': min_hour,
              'max_group_hour': max_hour} for dic in dicts]
    outlist.append((group, dicts))

结果:

[(1, [{'id': 123, 'hour': 5, 'min_group_hour': 3, 'max_group_hour': 5},
      {'id': 345, 'hour': 3, 'min_group_hour': 3, 'max_group_hour': 5}]),
 (2, [{'id': 345, 'hour': 3, 'min_group_hour': 3, 'max_group_hour': 3}]),
 (3, [{'id': 65, 'hour': -2, 'min_group_hour': -2, 'max_group_hour': -2}])]

本文收集自互联网,转载请注明来源。

如有侵权,请联系 [email protected] 删除。

编辑于
0

我来说两句

0 条评论
登录 后参与评论

相关文章