z-logo
open-access-imgOpen Access
IMPLEMENTATION of DSL for WEB SCRAPING
Author(s) -
Shail K. Shah,
Shashank Shyam Shankar,
N Rachana,
S Preetha
Publication year - 2020
Publication title -
international journal of recent trends in engineering and research
Language(s) - English
Resource type - Journals
ISSN - 2455-1457
DOI - 10.23883/ijrter.2020.6028.lbifz
Subject(s) - digital subscriber line , world wide web , computer science , telecommunications
The main goal of this project is to implement a DSL for Web Scraping. A Domain Specific Language or DSL in short is a language that is created for solving a single purpose. It is a language that is used in only one domain. In our project, that domain is web scraping. Our main aim is to create a simple scripting language with easy to use syntax with many features that help the user scrape the web easily. Currently, web scraping is a tedious process. At the moment, the majority of web scraping is done by the means of modules in high level languages. This would require the user and in-depth knowledge of the high-level language as well, and thus precludes many laymen from easy web scraping. This project will provide a DSL with highly simplified syntax which does not assume any skill from the user. Thus, anyone would be able to use this language to scrape the web with no previous knowledge of the domain. This DSL has been implemented using Python and its scraping libraries. With this, many features and functionalities can be implemented in the DSL thus providing an effective tool for web scraping without compromising on simplicity Keywords— Domain Specific Language, Web Scraping, Python, Beautiful Soup 4

The content you want is available to Zendy users.

Already have an account? Click here to sign in.
Having issues? You can contact us here
Accelerating Research

Address

John Eccles House
Robert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom