| 01:50:29 | | qwertyasdfuiopghjkl94 joins |
| 01:53:05 | | qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds] |
| 01:58:03 | | qwertyasdfuiopghjkl94 is now known as qwertyasdfuiopghjkl |
| 02:11:05 | | Jake3 (Jake) joins |
| 02:11:10 | | Jake quits [Ping timeout: 250 seconds] |
| 02:11:10 | | Jake3 is now known as Jake |
| 02:22:52 | | ThreeHM quits [Ping timeout: 250 seconds] |
| 02:24:56 | | ThreeHM (ThreeHeadedMonkey) joins |
| 02:56:38 | | qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds] |
| 03:50:52 | | qwertyasdfuiopghjkl joins |
| 04:27:32 | | qwertyasdfuiopghjkl quits [Client Quit] |
| 04:28:44 | | qwertyasdfuiopghjkl joins |
| 06:08:46 | | atphoenix_ (atphoenix) joins |
| 06:11:25 | | atphoenix quits [Ping timeout: 252 seconds] |
| 07:50:05 | <wolfin> | The ZoomInternet directory list is ready, https://p202.p0.n0.cdn.getcloudapp.com/items/xQu6zrJn/4757fd06-b05e-49a8-ac83-b018109ef1ef.csv?v=daeba0749b86be923f34faa6c9b7c6f7 |
| 08:48:19 | <Ryz> | Directory? Like a list of stuff to throw into ArchiveBot? |
| 09:58:08 | | qwertyasdfuiopghjkl65 joins |
| 09:58:45 | | qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds] |
| 10:01:22 | | @hook54321 quits [] |
| 10:02:21 | | hook54321 (hook54321) joins |
| 10:02:21 | | @ChanServ sets mode: +o hook54321 |
| 10:09:54 | | IDK (IDK) joins |
| 11:18:11 | | qwertyasdfuiopghjkl65 is now known as qwertyasdfuiopghjkl |
| 13:09:24 | | qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds] |
| 13:16:59 | | qwertyasdfuiopghjkl joins |
| 13:38:49 | | qwertyasdfuiopghjkl quits [Client Quit] |
| 13:40:11 | | qwertyasdfuiopghjkl joins |
| 15:33:11 | | masterX244 quits [Quit: No Ping reply in 180 seconds.] |
| 15:34:45 | | masterX244 (masterX244) joins |
| 16:28:50 | | qwertyasdfuiopghjkl quits [Ping timeout: 244 seconds] |
| 16:28:54 | | qwertyasdfuiopghjkl74 joins |
| 16:44:29 | <wolfin> | Correct, the list of user dirs and finable internal directories |
| 16:55:11 | | Iki quits [Ping timeout: 244 seconds] |
| 17:41:10 | | qwertyasdfuiopghjkl74 quits [Ping timeout: 244 seconds] |
| 21:48:53 | <Ryz> | Ah, a relatively small list of stuff in comparsion, huh wolfin~ |
| 21:55:30 | <Ryz> | JAA, worth running another queuebot with what wolfin provided up? The specific instructions for each of the job would be: |
| 21:55:32 | <Ryz> | --explain "https://wiki.archiveteam.org/index.php/ISP_Hosting - ISP userpage hosting accounts under http://users.zoominternet.net/" --concurrency 1 |
| 22:04:37 | <@JAA> | Sure, but definitely needs some further processing first. For example, it contains http://users.zoominternet.net/~basma/freedom%202021/ but that's linked from http://users.zoominternet.net/~basma/pages/news.htm so we'd be grabbing duplicates. |
| 22:35:56 | <Ryz> | JAA, would something like this be applicable on larger lists? Or theoretically infinite lists? |
| 22:41:45 | <@JAA> | I mean, ideally, we'd have a way to recurse based on a specified prefix with some seed URLs rather than what AB does now. |
| 22:42:55 | <@JAA> | The alternative solution is to first run the highest-level URLs, then check whether the lower-level ones were included there. But that wouldn't catch the example above because the link actually doesn't go to the directory but to .../index.htm. |