Commit Graph

233 Commits (e64178b072db11533a0cbf56cace64246cc6553c)

Author SHA1 Message Date
yihua.huang 807aefe9df change EntityUtil to IOUtil because some encoding error 11 years ago
yihua.huang 00b0a751b4 #33 ignore 'content-encoding' when redirect 11 years ago
yihua.huang 8f774afc84 add direct download 11 years ago
yihua.huang c18b603399 optimize long compare 11 years ago
yihua.huang ed3f3583cc downloader refactor 11 years ago
yihua.huang a37f40e6e6 add cookie supoort 11 years ago
yihua.huang 3c6fced48e update connection client 11 years ago
yihua.huang 09153ff715 #22 http proxy support #32 update httpclient to 4.3.1 11 years ago
yihua.huang edfc319c45 update httpclient to 4.3.1 11 years ago
yihua.huang 160a149b05 todo bugfix 11 years ago
yihua.huang 583a0eba8c #29 refactor some method name 11 years ago
yihua.huang 6fa82a418b #29 seed urls with more information 11 years ago
yihua.huang 1446ada732 some refactor 11 years ago
yihua.huang 84976c81ec remove useless code 11 years ago
yihua.huang b4fcf41168 add exit when comlete option 11 years ago
yihua.huang 352887870c remove shutdown call 11 years ago
yihua.huang a3f9ad198f refactor multi thread code in Spider 11 years ago
yihua.huang 7fb44d2eec #30 reuse PoolingClientConnectionManager for HttpClientDownloader 11 years ago
yihua.huang 5a226387e0 #27 nullpointer fix 11 years ago
yihua.huang 16e12e3bc9 #27 customize http header for downloader 11 years ago
yihua.huang 1a2c84ea78 #27 add timeout config to site 11 years ago
yihua.huang 372cc0ad06 update jar 12 years ago
yihua.huang 4acbc19cee [maven-release-plugin] prepare for next development iteration 12 years ago
yihua.huang cc3b787991 [maven-release-plugin] prepare release webmagic-0.3.2 12 years ago
yihua.huang b131878123 add example 12 years ago
yihua.huang 95ab4edec3 some bugfix 12 years ago
yihua.huang fba330872b fix a thread pool exception 12 years ago
yihua.huang 3c79d031bd fix thread pool 12 years ago
yihua.huang a2fba8caa2 update to 0.3.1 12 years ago
yihua.huang fb693a4ac4 [maven-release-plugin] prepare for next development iteration 12 years ago
yihua.huang bfaaa042b9 [maven-release-plugin] prepare release webmagic-parent-0.3.1 12 years ago
yihua.huang c17a31a21d fix null pointe exception #26 12 years ago
yihua.huang d2e0f0cd33 #25 use URL api in UrlUtils.canonicalizeUrl() 12 years ago
yihua.huang ef4cf49fee add stop method to spider #24 12 years ago
yihua.huang 58150a090d update jar 12 years ago
yihua.huang 57556ab879 merege 12 years ago
yihua.huang 692de76f86 fix issue #21 charset detect error 12 years ago
yihua.huang e7bf425df4 [maven-release-plugin] prepare for next development iteration 12 years ago
yihua.huang 77ff252316 [maven-release-plugin] prepare release webmagic-0.3.0 12 years ago
yihua.huang 1fc8e104ab add cycle retry 12 years ago
yihua.huang d141541ef3 add retry 12 years ago
yihua.huang a1ef2523cc update xsoup version 12 years ago
yihua.huang aefd0569a5 update version 12 years ago
yihua.huang 194518fd82 add switch 12 years ago
yihua.huang 326b97c65a update 12 years ago
yihua.huang 2c3574537a refactor in selectors 12 years ago
yihua.huang 85b7cf1563 complete test 12 years ago
yihua.huang d7cd9e5747 update pom 12 years ago
yihua.huang 55d4a76ab7 newselectors 12 years ago
yihua.huang d7abbd0e4b fix compile error 12 years ago
yihua.huang 5e9e8b2541 add TextContentSelector 12 years ago
yihua.huang 0cc0ccee35 add charset specific for easy call of HttpClientDownloader 12 years ago
yihua.huang 91dcccf7b5 add a sample 12 years ago
yihua.huang ad66d33f38 [maven-release-plugin] prepare for next development iteration 12 years ago
yihua.huang 9dc6b11954 [maven-release-plugin] prepare release webmagic-parent-0.2.1 12 years ago
yihua.huang 4f62dfc8a4 release 12 years ago
yihua.huang 74c940c758 [maven-release-plugin] prepare for next development iteration 12 years ago
yihua.huang a4bb4e3429 [maven-release-plugin] prepare release webmagic-parent-0.2.1 12 years ago
yihua.huang 194f16aa75 update 12 years ago
yihua.huang 0f0f1a9bcd release notes 12 years ago
yihua.huang c1471718df extractors 12 years ago
yihua.huang 20705b34ac add more option to extractors 12 years ago
yihua.huang c70ed57025 remove PriorityScheduler to core 12 years ago
yihua.huang 7003426898 update pom 12 years ago
yihua.huang 606417fdc7 update pom 12 years ago
yihua.huang d460e136ef update version 12 years ago
yihua.huang c79d6ecf09 complete all comments 12 years ago
yihua.huang 90bbe9b951 webmagic-core 12 years ago
yihua.huang 17f8ead28f update comments for selector 12 years ago
yihua.huang 77e6ca2945 update comments 12 years ago
yihua.huang 5073258237 closable 12 years ago
yihua.huang d01c0eb8ce update comments of spider 12 years ago
yihua.huang 5f1f4cbc46 update comments 12 years ago
yihua.huang 1148450ff9 update filecache to more useful 12 years ago
yihua.huang 3ba7a76f44 add combo extract to replace Extract2 Extract3... 12 years ago
yihua.huang 5cb45af3a4 +doc 12 years ago
yihua.huang ef673b985e add a method for httpclientdownloader 12 years ago
yihua.huang 067f3ea0cb add some null pointer check for httpclientdownloader 12 years ago
yihua.huang 9e82256ce3 update docs 12 years ago
yihua.huang 0a902b441c update docs 12 years ago
yihua.huang 0f2c5b5723 update redisscheduler 12 years ago
yihua.huang 787b952932 release notes and docs 12 years ago
yihua.huang 8b15f3c63d add test 12 years ago
yihua.huang ade5714d50 add https support 12 years ago
yihua.huang 21eca688e9 complete docs 12 years ago
yihua.huang 17d2d98cec remove invalid @date 12 years ago
yihua.huang 268bd8d0c4 remove saxon to extension 12 years ago
yihua.huang cff943f698 fix path format error 12 years ago
yihua.huang 5ef231a768 update version 12 years ago
yihua.huang 570533cce5 update readme 12 years ago
yihua.huang 36494bcfa5 add xpath2.0 api 12 years ago
yihua.huang 5c96407a3d fix a null domain error 12 years ago
yihua.huang c7005a0227 json fix 12 years ago
yihua.huang e5f4b3916f change file dir 12 years ago
yihua.huang 7d277e84d4 update lucene pipeline 12 years ago
yihua.huang b40cca1122 move model package to plugin 12 years ago
yihua.huang 4eb3d60083 fix nullpointer exception 12 years ago
yihua.huang b0af45f4bb complete redis support 12 years ago
yihua.huang f3a29d9315 fix pagedmodel bug 12 years ago
yihua.huang 629f8ac2d1 add extractors chain 12 years ago