01:50:29qwertyasdfuiopghjkl94 joins
01:53:05qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds]
01:58:03qwertyasdfuiopghjkl94 is now known as qwertyasdfuiopghjkl
02:11:05Jake3 (Jake) joins
02:11:10Jake quits [Ping timeout: 250 seconds]
02:11:10Jake3 is now known as Jake
02:22:52ThreeHM quits [Ping timeout: 250 seconds]
02:24:56ThreeHM (ThreeHeadedMonkey) joins
02:56:38qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds]
03:50:52qwertyasdfuiopghjkl joins
04:27:32qwertyasdfuiopghjkl quits [Client Quit]
04:28:44qwertyasdfuiopghjkl joins
06:08:46atphoenix_ (atphoenix) joins
06:11:25atphoenix quits [Ping timeout: 252 seconds]
07:50:05<wolfin>The ZoomInternet directory list is ready, https://p202.p0.n0.cdn.getcloudapp.com/items/xQu6zrJn/4757fd06-b05e-49a8-ac83-b018109ef1ef.csv?v=daeba0749b86be923f34faa6c9b7c6f7
08:48:19<Ryz>Directory? Like a list of stuff to throw into ArchiveBot?
09:58:08qwertyasdfuiopghjkl65 joins
09:58:45qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds]
10:01:22@hook54321 quits []
10:02:21hook54321 (hook54321) joins
10:02:21@ChanServ sets mode: +o hook54321
10:09:54IDK (IDK) joins
11:18:11qwertyasdfuiopghjkl65 is now known as qwertyasdfuiopghjkl
13:09:24qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds]
13:16:59qwertyasdfuiopghjkl joins
13:38:49qwertyasdfuiopghjkl quits [Client Quit]
13:40:11qwertyasdfuiopghjkl joins
15:33:11masterX244 quits [Quit: No Ping reply in 180 seconds.]
15:34:45masterX244 (masterX244) joins
16:28:50qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds]
16:28:54qwertyasdfuiopghjkl74 joins
16:44:29<wolfin>Correct, the list of user dirs and finable internal directories
16:55:11Iki quits [Ping timeout: 244 seconds]
17:41:10qwertyasdfuiopghjkl74 quits [Ping timeout: 244 seconds]
21:48:53<Ryz>Ah, a relatively small list of stuff in comparsion, huh wolfin~
21:55:30<Ryz>JAA, worth running another queuebot with what wolfin provided up? The specific instructions for each of the job would be:
21:55:32<Ryz>--explain "https://wiki.archiveteam.org/index.php/ISP_Hosting - ISP userpage hosting accounts under http://users.zoominternet.net/" --concurrency 1
22:04:37<@JAA>Sure, but definitely needs some further processing first. For example, it contains http://users.zoominternet.net/~basma/freedom%202021/ but that's linked from http://users.zoominternet.net/~basma/pages/news.htm so we'd be grabbing duplicates.
22:35:56<Ryz>JAA, would something like this be applicable on larger lists? Or theoretically infinite lists?
22:41:45<@JAA>I mean, ideally, we'd have a way to recurse based on a specified prefix with some seed URLs rather than what AB does now.
22:42:55<@JAA>The alternative solution is to first run the highest-level URLs, then check whether the lower-level ones were included there. But that wouldn't catch the example above because the link actually doesn't go to the directory but to .../index.htm.