Sun, 14 Mar 2021 21:28:02 +0000 |
Henry S. Thompson |
prepare for real parallel distribution
|
Sun, 14 Mar 2021 21:25:01 +0000 |
Henry S. Thompson |
environment improvements
|
Wed, 03 Mar 2021 19:33:56 +0000 |
Henry S. Thompson |
trying to move to slurm
|
Sat, 09 May 2020 16:16:28 +0100 |
Henry S. Thompson |
improved F handling/logging
|
Fri, 08 May 2020 19:52:36 +0100 |
Henry S. Thompson |
keep separate antecedants separate, buggy?
|
Thu, 07 May 2020 18:47:24 +0100 |
Henry S. Thompson |
track redirects, need to us full crawldiagnostics.warc.gz for "location:" and "Uri:"
|
Thu, 07 May 2020 11:33:24 +0100 |
Henry S. Thompson |
refactor, change summary print (problem?)
|
Wed, 06 May 2020 18:28:52 +0100 |
Henry S. Thompson |
bare framework working
|
Wed, 06 May 2020 14:25:44 +0100 |
Henry S. Thompson |
starting on tool to assemble as complete as we have info wrt a seed URI
|
Wed, 06 May 2020 14:24:42 +0100 |
Henry S. Thompson |
use local .m2/repository for Hadoop 3.4.0
|
Wed, 06 May 2020 14:23:33 +0100 |
Henry S. Thompson |
works for big files with Hadoop 3.4.0
|
Wed, 06 May 2020 14:22:48 +0100 |
Henry S. Thompson |
x
|
Tue, 28 Apr 2020 19:02:34 +0100 |
Henry S. Thompson |
log trucations
|
Tue, 28 Apr 2020 19:02:14 +0100 |
Henry S. Thompson |
impose some limits
|
Tue, 28 Apr 2020 19:01:41 +0100 |
Henry S. Thompson |
x
|