Submit Your Requirement

Content Extraction Service

Once we were provided with the list of source websites and data points, our team started working on the project. As this use case was for news data, the frequency of scrapes had to be very high.

Scroll down to discover

Content Extraction Service

Client

A popular media website from Brazil looking for content extraction and mining service.

EMAIL : sales@promptcloud.com
Submit requirement

Challenge

The client wanted content to be extracted on a continuous basis from Brazilian news sites to power their news portal. The list of websites included popular blogs, news sites, forums and a few content bookmarking sites. The required data points were date of publishing, author name, title, main text content and tags.

Solution: Content Extraction

Once we were provided with the list of source websites and data points, our team started working on the project. As this use case was for news data, the frequency of crawls had to be very high. This meant fresh data sets had to be provided every day. Since each site in the list had a different structure and design, site specific crawl and extraction was the solution used for this case.
Once our team finished setting up the web crawlers, the data started flowing in. The data was then cleaned and formatted to be uploaded to the client’s Dropbox servers in XML format. The number of records being delivered per day was above 300,000.

Benefits to the client:

  • Our team handled every technical aspect of the crawling process
  • The initial setup only took 3 days to get completed and the supply of data was consistent
  • Since monitoring was set up for each site, the quality and consistency of data was top notch
  • Although some of the sites used dynamic coding practices, our tech stack could handle them well
  • The client could launch their news portal with the data in a short notice
  • The cost incurred was way less than what an in-house crawling set up would have costed them

Related Use Cases

Use Case
Concept,Online,Sopping.,Boxes,And,Shopping,Bag,With,Smartphone,Online

Continuous News Feeds In Near Real Time

Read More

Use Case

Scraping Product Prices And Reviews

Read More

Use Case

Get Amazon Reviews Data

Read More

RECOMMENDED

Classified scraper

Classifieds Scraper For Listings

 

Read More

social media data mining

Social Media Data Mining Service

 

Read More

Walmart price scraper

Walmart Price Scraper

 

Read More

Contact Us

Contact Us
© Promptcloud 2009-2020 / All rights reserved.
To top