COVID-19数据集

包含COVID-19历史报告病例、死亡、康复等数据的表格数据集,用于疫情监测与分析。

newaaa41newaaa41
GitHub
2022-01-10 更新
浏览 5
表格COVID-19表格数据

基本信息

模态
表格
创建/更新时间
2022-01-10

资源简介

该数据集包含COVID-19疫情报告的历史数据,截止至1月24日。数据包括新报告的病例、死亡、康复人数,以及累计报告的病例、死亡、康复人数和当前活跃病例数。数据已从多个来源清洗并整合,适用于疫情趋势分析和建模。

原始链接

https://github.com/newaaa41/COVID19-Dataset

访问原始数据

官方服务

如需原始数据获取支持或标注服务,请联系我们。

帮我联系

下载信息

注册下载

Tips: 该数据集需要在对应的数据源网站注册通过后,才能进行数据下载,注册有对应要求,或者需要收费。

暂未开放

公开下载

Tips: 该数据集属于公开下载,应该可以免费公开下载。

免登录

有偿下载

Tips: 该数据集 Qianfanghub 可以协助提供有偿下载服务,注意,服务不针对数据相关产权,只是技术服务费。

提供高速下载与技术交付服务(收技术服务费,非数据销售)

暂未开放

千方医数集,医疗数据集部分,是为社区服务的公开医疗数据集搜索引擎,并不存储或者下载原始的任何数据。 如果您有其他医疗数据需求,可以和客服联系,或者下工单。我们有强大的三甲医疗机构帮助您提供个性化的医疗数据定制、采集、标注服务。

使用方式

数据集获取

git clone https://github.com/newaaa41/COVID19-Dataset.git

curl -L -o repo.zip https://github.com/newaaa41/COVID19-Dataset/archive/refs/heads/master.zip
unzip repo.zip

源站 README 摘录(使用方式)

How To Use

The data is stored in four different forms

  1. in covid19_dataset.csv file which contains the whole dataset in one file
  1. in BY_REGION/ directory where the data has been split into different csv files based on countries and regions
  1. in BY_DAY/ directory where the data has been split into different csv files based on the date
  1. in BY_MEASUREMENT directory where the data has been split into different csv files based on the measurement (newly reported cases, newly reported deaths, newly reported recoveries, total reported cases, total reported deaths, total reported recoveries, currently active cases)

To use the data you can clone the repository

git clone https://github.com/newaaa41/COVID19-Dataset

or you can use the url for any csv file

example:

#python

import pandas as pd
df = pd.read_csv(''https://raw.githubusercontent.com/newaaa41/COVID19-Dataset/master/covid19_dataset.csv'')
#R

data <- read.csv(''https://raw.githubusercontent.com/newaaa41/COVID19-Dataset/master/covid19_dataset.csv'')

数据加载示例(表格/文本类)

import pandas as pd, glob, os

files = (glob.glob(os.path.join(path, "**", "*.csv"), recursive=True)
       + glob.glob(os.path.join(path, "**", "*.tsv"), recursive=True)
       + glob.glob(os.path.join(path, "**", "*.xlsx"), recursive=True))
print("数据文件:", files)
df = pd.read_csv(files[0])
print(df.shape); print(df.columns.tolist()); print(df.head(3))

完整仓库:github.com/newaaa41/COVID19-Dataset

精度瓶颈?数据缺失?

当前公开数据无法满足您的算法精度?千方提供针对 新型冠状病毒肺炎 的高质量、多模态真实临床数据定制解决方案。

获取专属数据定制方案