内容简介:The documentation in this repository describe the FullStack webscrapping platform for use in Machine learning.
The documentation in this repository describe the FullStack webscrapping platform for use in Machine learning.
Architecture
We first break the architecture into four distictive components namely Front-End, API, Scrapers and Database. The user sends information from the front-end to the API, the fron-end connects the API through a form. Inputs like the youtube URL are sent through front-end. Later the scrapers through the API pulls the necessary data and is saved to the database. Afterwhich the data is served to the front-end.
The Tech Stack are as below
- Front-End - javascript
- API - express
- scraper - puppeteer
- db - mysql (typeorm)
Also we need nodejs, npm and mysql.
The Architecture consists of several components:
Front End
For the Front-end we will have a header, an input box and a button. Below which we will have render boxes which renders relevant info from json. This will send data to the API.
API
We will have to create a single route with two methods GET and POST. We use nodejs and simple backed framework express.
Scraper
This function takes in URL and reaches out to YouTube, fetch the relevant data and then store it into the database.
Database
We use mySQL here. Here we add id, name, avatar and channelURL
To run the program
First go into server
$ npm install init
Install all the necessary packages
$ npm install express $ npm install body-parser
Run the index.js script
$ node index.js
Thanks to Aron from Uber
以上就是本文的全部内容,希望对大家的学习有所帮助,也希望大家多多支持 码农网
猜你喜欢:本站部分资源来源于网络,本站转载出于传递更多信息之目的,版权归原作者或者来源机构所有,如转载稿涉及版权问题,请联系我们。
百度SEM竞价推广
马明泽 / 电子工业出版社 / 2017-5 / 59
竞价推广已成为企业昀主要的网络营销方式,《百度SEM竞价推广:策略、方法、技巧与实战》以百度竞价推广为基础,全面阐述了整个竞价推广过程中的重要环节,涉及大量账户操作实战技巧,以及解决各类难点的方法,其中包括搜索引擎营销基础、百度搜索推广介绍、账户结构搭建技巧、关键词与创意的使用技巧、质量度优化与提升、账户工具的使用、百度推广客户端的使用、企业搜索推广方案制作、百度网盟推广、着陆页分析、效果优化与数......一起来看看 《百度SEM竞价推广》 这本书的介绍吧!