00:15:09Arcorann (Arcorann) joins
00:32:45dm4v_ joins
00:33:39dm4v quits [Ping timeout: 265 seconds]
00:33:39dm4v_ is now known as dm4v
00:33:40dm4v quits [Changing host]
00:33:40dm4v (dm4v) joins
00:50:16kallsyms quits [Quit: probably segfaulted]
00:50:31kallsyms joins
00:58:47Discant quits [Ping timeout: 265 seconds]
01:01:50dm4v_ joins
01:03:08dm4v quits [Ping timeout: 265 seconds]
01:03:08dm4v_ is now known as dm4v
01:03:09dm4v quits [Changing host]
01:03:09dm4v (dm4v) joins
01:12:39igloo22225 quits [Quit: The Lounge - https://thelounge.chat]
01:14:03igloo22225 (igloo22225) joins
03:30:27sec^nd quits [Remote host closed the connection]
03:32:16sec^nd (second) joins
04:43:42Wohlstand quits [Client Quit]
05:27:52BlueMaxima quits [Read error: Connection reset by peer]
05:56:02DogsRNice quits [Read error: Connection reset by peer]
08:15:08sec^nd quits [Remote host closed the connection]
08:16:04nothere quits [Quit: Leaving]
08:17:30sec^nd (second) joins
08:26:57nothere joins
09:02:27Chenjesu quits [Remote host closed the connection]
09:03:27Discant joins
10:00:07nepeat quits [Ping timeout: 245 seconds]
10:29:18spirit joins
10:36:22nepeat (nepeat) joins
12:01:51thetechrobo_ joins
12:02:00qwertyasdfuiopghjkl quits [Remote host closed the connection]
12:02:00dm4v quits [Client Quit]
12:02:00Arcorann quits [Remote host closed the connection]
12:02:00TheTechRobo quits [Remote host closed the connection]
12:02:08dm4v joins
12:02:10dm4v quits [Changing host]
12:02:10dm4v (dm4v) joins
12:06:42ave quits [Client Quit]
12:06:46ave9 (ave) joins
12:07:37Arcorann (Arcorann) joins
12:23:49qwertyasdfuiopghjkl joins
12:44:45Sutilbo joins
12:47:00<Sutilbo>Hello, could you save the files listed in http://0x0.st/oMCu.7z ? I was trying to download hispafiles.ru with wget, but the problem is that you have to download the threads (and the files they contain) individually, otherwise it first create files like hispafiles.ru/c and when you want to download something like hispafiles.ru/c/res/31742.html then it cannot be saved because the directory cannot be created due to there is already a fi
12:47:00<Sutilbo>le with the same name. I imagine that ArchiveBot must work in a similar way.
12:47:01<Sutilbo>I would have no problem doing it myself, but my connection is quite slow and I don't know how big the site is, so it could take me days or weeks to download everything. Also, recently the owner of the site said that "it will be alive for the time being" (which does not give me much confidence) so I would like to save everything I can in case it is closed at any moment as it happened with hispachan.org a few days ago.
12:47:01<Sutilbo>By the way, if you download the files from that list, you should activate something like the --no-parent parameter to avoid the problem that I mentioned above (that is, ArchiveBot should only download the listed files and not go around the site if it finds a 404 error or similar).
12:57:37Discant quits [Ping timeout: 245 seconds]
13:16:19<thetechrobo_>Sutilbo: ArchiveBot does not work like that. It saves files into a format called WARC which is required for the Wayback Machine. It stores all HTTP traffic sent and received, and there can be multiple records with the same URL in the same file and it would still work fine.
13:17:15thetechrobo_ is now known as TheTechRobo
13:20:40<Sutilbo>But tell me one thing thetechrobo_, apart from generating the WARC, does ArchiveBot download the files locally just like wget does (at least as a cache or something like that)?
14:09:23Arcorann quits [Ping timeout: 265 seconds]
14:14:47<TheTechRobo>Sutilbo: Yes, but they're discarded after they're written to the WARC.
14:17:19<TheTechRobo>I'm not sure why it hasn't been run through yet.
14:18:18<TheTechRobo>I asked again in #archivebot-bs
14:18:22<TheTechRobo>*#archivebot
14:18:27<TheTechRobo>my brain is not functioning
14:18:46<TheTechRobo>I recommend moving to #archiveteam-bs for further discussion.
14:29:14<Sutilbo>TheTechRobo: Sure, I will ask there.
14:36:15spirit quits [Client Quit]
14:38:29spirit joins
14:41:33spirit quits [Client Quit]
14:46:22spirit joins
14:55:29spirit quits [Client Quit]
15:02:59ssb joins
15:19:07Brella quits [Ping timeout: 265 seconds]
15:57:07ats quits [Quit: new irssi]
15:57:29ats (ats) joins
17:56:02zoe joins
19:00:18knecht4202 quits [Read error: Connection reset by peer]
19:00:20knecht42024 (knecht420) joins
19:04:53zoe quits [Client Quit]
19:47:47DiscantX joins
21:12:50internet_od (internet_od) joins
21:23:38dm4v_ joins
21:23:38dm4v quits [Ping timeout: 265 seconds]
21:23:38jamesp quits [Ping timeout: 265 seconds]
21:23:38capjamesg quits [Ping timeout: 265 seconds]
21:23:38dm4v_ is now known as dm4v
21:23:38thuban quits [Ping timeout: 265 seconds]
21:23:38Barto quits [Ping timeout: 265 seconds]
21:23:38asie quits [Ping timeout: 265 seconds]
21:23:38marked1 quits [Ping timeout: 265 seconds]
21:23:38phuzion quits [Remote host closed the connection]
21:23:38summerisle quits [Remote host closed the connection]
21:23:40fionera quits [Remote host closed the connection]
21:23:40dm4v quits [Changing host]
21:23:40dm4v (dm4v) joins
21:23:42capjamesg joins
21:23:43jamesp joins
21:23:43jamesp quits [Changing host]
21:23:43jamesp (jamesp) joins
21:23:49thuban joins
21:23:53Barto (Barto) joins
21:23:55asie joins
21:23:57marked1 (marked1) joins
21:24:22phuzion (phuzion) joins
21:24:23summerisle (summerisle) joins
21:24:34fionera (Fionera) joins
21:28:23qwertyasdfuiopghjkl quits [Client Quit]
21:41:28internet_od quits [Client Quit]
21:41:57qwertyasdfuiopghjkl joins
21:57:08HP_Archivist (HP_Archivist) joins
22:04:42Mateon1 quits [Ping timeout: 245 seconds]
22:06:42Mateon1 joins
22:08:30DiscantX quits [Ping timeout: 265 seconds]
22:08:43DiscantX joins
22:14:33HP_Archivist quits [Client Quit]
22:44:16DiscantX quits [Ping timeout: 265 seconds]
22:50:59michaelblob quits [Read error: Connection reset by peer]
22:52:17michaelblob (michaelblob) joins
23:16:17cj joins
23:19:54KabukiFan joins
23:20:08<KabukiFan>Hey, I was wondering if I could get someone's help backing up https://www2.ntj.jac.go.jp/dglib/contents/learn/ebook01/mainmenu.html to the internet archive?
23:20:23<KabukiFan>The reason I ask is that it is quite an old website and it is invaluable for researchers of Japanese theatre.
23:20:43<KabukiFan>I know they may be doing upgrades at some point and this, frankly quite old, site might be a casualty.
23:37:09<TheTechRobo>KabukiFan: Try asking in #archivebot