Newsscraper - A simple Python 3 module to get crypto or news articles and their content from various RSS feeds.

Last update: Jan 02, 2022

Overview

NewsScraper

A simple Python 3 module to get crypto or news articles and their content from various RSS feeds.

🔧 Installation

Clone the repo locally.
Use the package manager pip to install the requirements.

pip install -r requirements.txt

✨ Basic Usage

import NewsScraper

all_data = NewsScraper.fetch_all()
news_data = NewsScraper.fetch_news_data()
crypto_data = NewsScraper.fetch_crypto_data()

fetch_all()

Returns a set of NewsScraper.Result containing fetched results from all available RSS feeds

Can include categories: GLOBAL, US, EU, CRYPTO, BLOCKCHAIN, BTC, ETH, LTC.

fetch_news_data()

Returns a set of NewsScraper.Result containing fetched results from CNN, ABC News, Yahoo News, Fox News RSS feeds

Can include categories: GLOBAL, US, EU.

fetch_crypto_data()

Returns a set of NewsScraper.Result containing fetched results from CoinJournal, Crypto Currency News RSS feeds.

Can include categories: CRYPTO, BLOCKCHAIN, BTC, ETH, LTC.

🔨 Advanced Usage

NewsScraper.Result class

A class used to represent a returned article.

Attributes

context : str

A string describing the category of the article.

ex. "GLOBAL", "US", "BLOCKCHAIN", "BTC".
title : str

A string containing the name of the article.
summary : str

A string containing the summary of the article.

NOTE: sometimes it can have the value of "", because the RSS feed didn't provide a summary.
content : str

A string containing the content of the article.

Methods

Result.json()

Returns a dictionary with the attributes of the class formatted in JSON.

ex.

{
  "context": "global",
  "title": "title of the article",
  "summary": "summary of the article",
  "content": "content of the article"
}

News RSS Feeds

All of these functions return a set of NewsScraper.Result containing fetched results of the described RSS feeds.

fetch_abc()
fetch_cnn()
fetch_yahoo()
fetch_fox_news()

Can include categories: GLOBAL, US, EU.

Alternatively, you can use fetch_news_data() to receive results from all of them.

Crypto RSS Feeds

All of these functions return a set of NewsScraper.Result containing fetched results of the described RSS feeds.

fetch_coinjournal()
fetch_cryptocurrencynews()

Can include categories: CRYPTO, BLOCKCHAIN, BTC, ETH, LTC.

Alternatively, you can use fetch_news_data() to receive results from all of them.

🤝 Contributing

Pull requests are welcome. For major changes, please open an issue first to discuss what you would like to change.

📝 License

This project is licensed under the MIT license.

Newsscraper - A simple Python 3 module to get crypto or news articles and their content from various RSS feeds.

Related tags

Overview

NewsScraper

🔧 Installation

✨ Basic Usage

🔨 Advanced Usage

NewsScraper.Result class

context : str

title : str

summary : str

content : str

Result.json()

News RSS Feeds

Crypto RSS Feeds

🤝 Contributing

📝 License

Owner

Rokas

News, full-text, and article metadata extraction in Python 3. Advanced docs:

A web scraper which checks price of a product regularly and sends price alerts by email if price reduces.

jd_maotai rpa 基于selenium驱动的jd抢购rpa机器人

This is a module that I had created along with my friend. It's a basic web scraping module

Automatically scrapes all menu items from the Taco Bell website

Scraping Thailand COVID-19 data from the DDC's tableau dashboard

An helper library to scrape data from Instagram effortlessly, using the Influencer Hunters APIs.

A web service for scanning media hosted by a Matrix media repository

Searching info from Google using Python Scrapy

A simple app to scrap data from Twitter.

Twitter Claimer / Swapper / Turbo - Proxyless - Multithreading

Quick Project made to help scrape Lexile and Atos(AR) levels from ISBN

A leetcode scraper to compile all questions in leetcode free tier to text file. pdf also available.

PaperRobot: a paper crawler that can quickly download numerous papers, facilitating paper studying and management

Web Crawlers for Data Labelling of Malicious Domain Detection & IP Reputation Evaluation

Facebook Group Scraping Using Beautiful Soup & Selenium

Nekopoi scraper using python3

A Python package that scrapes Google News article data while remaining undetected by Google.

Google Developer Profile Badge Scraper

Scraping weather data using Python to receive umbrella reminders