- Data Analysis
- Data Cleaning
OpenRefine is a powerful tool for working with messy data.
OpenRefine is a standalone open-source desktop application for data cleanup and transformation to other formats, the activity known as data wrangling. It is similar to spreadsheet applications; however, it behaves more like a database. It operates on rows of data which have cells under columns, which is very similar to relational database tables. An OpenRefine project consists of one table. The user can filter the rows to display using facets that define filtering criteria. Unlike spreadsheets, most operations in OpenRefine are done on all visible rows: transformation of all cells in all rows under one column, creation of a new column based on existing column data, etc. All actions that were done on a dataset are stored in a project and can be replayed on another dataset.