Tell HN: The Police Data Accessibility Project

栏目: IT技术 · 发布时间: 5年前

内容简介:It is my belief that perhaps the single most effective way we can make concrete progress toward this goal is in the realm of data accessibility. Specifically the accessibility of granular, local, county level police citation data.This however, is a big und

Police Data Accessibility Project

Rough Mission Statement (will evolve):

It is my belief that perhaps the single most effective way we can make concrete progress toward this goal is in the realm of data accessibility. Specifically the accessibility of granular, local, county level police citation data.

This however, is a big undertaking, as most counties have clumsy/antiquated systems for searching for and extracting this type of data, and almost none have fully exposed datasets.

Ultimately, the future goal for this initiative is to:

Request (via FOIA) or Scrape, and then clean and aggregate county level police citation data for as many counties as possible, and to most importantly make this data open and free to the public. I believe doing so will enable citizen data scientists and data journalists to then make progress on analysis, whereby they will serve the public by looking for and finding trends and anomalies in police behavior, leading to more accountability for both individual police officers and their larger police organizations.

Initial Goals:

Scraping:

  1. Many counties outsource their court records data to third party vendors such as Tyler Technologies. Finding and building scrapers for portals that are the same for many counties seems like a great early goal. A list of counties court record systems and their vendors must be made. This will be done collaboratively in this Google Sheet . For more details see https://github.com/Police-Data-Accessibility-Project/Police-Data-Accessibility-Project/issues/6 .

  2. Finding and writing scrapers for other large counties. Prioritize counties with easier to scrape systems first.

For guidelines to contributing to scraping, please see https://github.com/Police-Data-Accessibility-Project/Police-Data-Accessibility-Project/blob/master/CONTRIBUTING.md

Freedom of Information Act Requests:

  1. Researching a data request template with all the data we want to ask for in FOIA requests

  2. Submitting FOIA requests and monitoring responses

I will be adding tasks to the projects section of this repo, so we can all keep track of them there.

If you are looking to start building a scraper, the csv file above has the URLS of all most US counties' public records portals.

The fields we would like to make sure to collect at a minimum from any scrape are:

_id _state _county CaseNum FirstName MiddleName LastName Suffix DOB Race Sex ArrestDate FilingDate OffenseDate DivisionName CaseStatus DefenseAttorney PublicDefender Judge ChargeCount ChargeStatute ChargeDescription ChargeDisposition ChargeDispositionDate ChargeOffenseDate ChargeCitationNum ChargePlea ChargePleaDate ArrestingOfficer ArrestingOfficerBadgeNumber


以上就是本文的全部内容,希望对大家的学习有所帮助,也希望大家多多支持 码农网

查看所有标签

猜你喜欢:

本站部分资源来源于网络,本站转载出于传递更多信息之目的,版权归原作者或者来源机构所有,如转载稿涉及版权问题,请联系我们

创新公司

创新公司

[美]艾德·卡特姆、埃米·华莱士 / 靳婷婷 / 中信出版社 / 2015-2 / 49.00元

●《玩具总动员》《海底总动员》《机器人瓦力》《飞屋环游记》等14部脍炙人口的动画长片, 近30次奥斯卡奖, 7部奥斯卡最佳动画长片,7次金球奖; ●几乎每一部电影一上映都位居票房榜首,所有电影都曾进入影史票房总榜前50,每一部电影都是商业与艺术的双赢。 ●即便新兴动画公司不断涌现,皮克斯始终保持动画界的王者之位,这一切背后的秘密就在于:不断推动创新的创意管理方式。 你可以从本书......一起来看看 《创新公司》 这本书的介绍吧!

CSS 压缩/解压工具
CSS 压缩/解压工具

在线压缩/解压 CSS 代码

Base64 编码/解码
Base64 编码/解码

Base64 编码/解码

HEX HSV 转换工具
HEX HSV 转换工具

HEX HSV 互换工具