About the Bundle
Get the data through scraping - and then turn it into stories with a quick heist! Buy both Scraping for Journalists and The Data Journalist Heist together for a discounted price.
Scraping for Journalists (2nd edition)
How to grab information from hundreds of sources, put it in data you can interrogate - and still hit deadlines
Scraping - getting a computer to capture information from online sources - is one of the most powerful techniques for data-savvy journalists who want to get to the story first, or find exclusives that no one else has spotted. Faster than FOI and more detailed than advanced search techniques, scraping also allows you to grab data that organisations would rather you didn’t have - and put it into a form that allows you to get answers.
Scraping for Journalists introduces you to a range of scraping techniques - from very simple scraping techniques which are no more complicated than a spreadsheet formula, to more complex challenges such as scraping databases or hundreds of documents. At every stage you'll see results - but you'll also be building towards more ambitious and powerful tools.
You’ll be scraping within 5 minutes of reading the first chapter - but more importantly you'll be learning key principles and techniques for dealing with scraping problems.
Unlike general books about programming languages, everything in this book has a direct application for journalism, and each principle of programming is related to their application in scraping for newsgathering. And unlike standalone guides and blog posts that cover particular tools or techniques, this book aims to give you skills that you can apply in new situations and with new tools.
Data Journalism Heist
How to get in, get the data, and get the story out - and make sure nobody gets hurt
Data journalism is a key skill for journalists to differentiate themselves in a world where almost anyone can publish, and competition for journalism jobs is fierce.
Whether it's hard stories from government spending and MPs' expenses, or softer stories from sports data, fashion trends or music and social activity, our increasingly digital world is providing a rich range of potential new story sources - and new forms of storytelling too.
This short ebook introduces you quickly to key techniques in finding that data and turning it into stories - through a 'Data Journalism Heist'.
This isn't about the huge investigative projects that you hear about, but the everyday stories that you can do with speed and simplicity. It's about getting in, getting the data, and getting the story out safely. No one gets hurt.
In the process you'll learn about:
- Sources of data - where to find data stories and leads
- Typical data stories - how to find simple stories in data
- Basic spreadsheet techniques - finding the biggest and smallest values; calculating averages and totals.
- Top techniques for getting stories against a deadline - using filters and pivot tables to get to the story quickly
- Making a clean getaway - avoiding mistakes in data journalism
- Telling the story - basic techniques in visualising and humanising your data-led story
As the book is published you'll receive regular updates as you build on previous skills towards the final story. User feedback, examples and ideas will be incorporated as they come in.
The Leanpub 45-day 100% Happiness Guarantee
Within 45 days of purchase you can get a 100% refund on any Leanpub purchase, in two clicks.
See full terms
Free Updates. DRM Free.
If you buy a Leanpub book, you get free updates for as long as the author updates the book! Many authors use Leanpub to publish their books in-progress, while they are writing them. All readers get free updates, regardless of when they bought the book or how much they paid (including free).
Most Leanpub books are available in PDF (for computers), EPUB (for phones and tablets) and MOBI (for Kindle). The formats that a book includes are shown at the top right corner of this page.
Finally, Leanpub books don't have any DRM copy-protection nonsense, so you can easily read them on any supported device.