Extracts the title, publication time, body text, images, and video from a news page Works across sites in any language, and stays fast and stable under load Source: https://github.com/crawlerclub/ce Demo: http://goobot.org/