alpcentaur
|
89dcca2031
|
added further handling for javascript links not being urls, made config for giz work
|
2023-11-28 15:27:39 +00:00 |
|
alpcentaur
|
a0075e429d
|
added further database in config.yaml, added new exception for downloading js generated html pages
|
2023-11-27 15:10:11 +00:00 |
|
alpcentaur
|
b2cf4b67ce
|
added first config parameters for search on not uniform entries
|
2023-11-15 17:27:54 +00:00 |
|
alpcentaur
|
ff23c22e3c
|
added working bund.de-bekanntmachungen config with new example of xpath contains
|
2023-11-13 16:44:11 +00:00 |
|
alpcentaur
|
06fa81e549
|
added function find config parameter and changed core spider
|
2023-11-10 01:12:49 +00:00 |
|
alpcentaur
|
a846ce04cc
|
specifying the links, new exception clause if soupparser does not work
|
2023-11-07 14:55:05 +00:00 |
|
alpcentaur
|
c078ee4b1b
|
first function works, actuall xml parser has still problems with certain xml types
|
2023-11-06 19:17:45 +00:00 |
|
alpcentaur
|
8b20bc178f
|
added multi pages configuration and code
|
2023-11-06 18:17:32 +00:00 |
|
alpcentaur
|
7aa903883b
|
update to config.yaml
|
2023-11-03 12:23:04 +00:00 |
|
alpcentaur
|
5ac07d151a
|
added first config.yaml template and started creating folder structure
|
2023-10-31 17:41:44 +00:00 |
|