Forgot your password?

typodupeerror
The Media Technology

How Journalists Data-Mined the Wikileaks Docs 59

Posted by samzenpus
from the reading-between-a-million-lines dept.
meckdevil writes "Associated Press developer-journalist extraordinaire Jonathan Stray gives a brilliant explanation of the use of data-mining strategies to winnow and wring journalistic sense out of massive numbers of documents, using the Iraq and Afghanistan war logs released by Wikileaks as a case in point. The concepts for focusing on certain groups of documents and ignoring others are hardly new; they underlie the algorithms used by the major Web search engines. Their use in a journalistic context is on a cutting edge, though, and it raises a fascinating quandary: By choosing the parameters under which documents will be considered similar enough to pay attention to, journalist-programmers actually choose the frame in which a story will be told. This type of data mining holds great potential for investigative revelation — and great potential for journalistic abuse."

This discussion has been archived. No new comments can be posted.

How Journalists Data-Mined the Wikileaks Docs

Comments Filter:

The world will end in 5 minutes. Please log out.

Working...