You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
//get crawler and set urls output filename under "urls" directory
$crawler = new GithubCrawler("out.txt");
//crawl each repo on github which got over 3000 stars and do the interest value analyse immediately with store path data.txt which under "data" directory
//crawl and save all repo html source under resource directory by default urls file which contains all urls records. in this case, the file path is urls/out.txt (GithubCrawler constructor parameter)
//$crawler->crawlEachRepoHtml();
//crawl and save all repo html source under resource directory by given file which contains all urls records. in this case, the file path is myurls.txt