[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/pol/ - Politically Incorrect

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
Flag
File
  • Please read the Rules and FAQ before posting.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


Janitor acceptance emails will be sent out over the coming weeks. Make sure to check your spam folder!


[Advertise on 4chan]


File: monk_scribe.jpg (87 KB, 550x417)
87 KB JPG
Anon you need to back up **EVERYTHING THAT YOU DON'T WANT TO VANISH FOREVER**.

- It started with search engine filtering for political reasons. **Fine, we'll use alternative search engines.**
- Scammers started abusing SEO and flooded search results with ever more crap. **Fine, at least we can still find the real content with some effort.**
- Walled gardens like Discord, Telegram and other platforms start to attract new content not accessible on the web. **Fine I guess we will just install all of these apps.**
- Free content gets paywalled as greedy subhumans want shekels and sabotage free alternatives via lawfare & illegal means. **Fine, we'll just pay for it.**
- Data rot starts to snowball, old websites simply aren't renewed and entire free hosters shut down. **Fine, at least we have archive.org.**
- Archive.org crawler is incomplete/doesn't archive exclusions/removes content upon request of (((certain organisations))). **Fine uh... well shit.**

> you are here

Web 1.0 is already dead and Web 2.0 is dying a slow death. There will be no Web 3.0. AI garbage content is way worse than the SEO spam of old and soon the HTTP internet as such will be so rife with jeet generated filth that it will become entirely useless. Meanwhile heebs like (((Dario Amodei))) and shabbos goys like Dean Ball start preaching data authority centralization. They not only want to be the new Google they want to become the entire oracle of delphi and single focal point of truth.

**What you can do**

Back up everything worth saving. Books, websites, images, videos, audio, anything you can.

If you're technically skilled, use these for websites: httrack, wget
If you're technically retarded, use these for websites: singlefile, obsidian web clipper
For videos/audio of any kind use yt-dlp, it also works on instagram, tiktok and other sites.
>>
**How to store your data**

Have at least a +1 copy of any of these:
If you're a richfag, simply buy large HDDs and put them into cold storage (less use=less wear).
If you're not that well off, buy used HDDs, they will still be fine but more prone to failure.
If you're a poorfag, use USB sticks.
If you're literally homeless, mega.co.nz comes with 20gb of free storage.

HDDs last about 5-10 years while USB sticks last about 1-5 years. Check your backups regularly and replace when needed.

**What to do with your data**

- Share them with other anons. Make a torrent or put your collection on a file hoster like litterbox, mega, make a torrent, etc.
- Organize it for your own use. Use systems like Hydrus (autistic tier), Obsidian (normie tier) or nested folders (retarded oldfag tier).
- Upload content to archive.org but be aware that they absolutely censor and take down sensitive content.
- Upload books to anna's archive. They don't censor but may add "probable spam" notifications on an entry if someone reports (downloading still works).

**My data collection is pure chaos nobody would want it**

With the advent of local vision models it's actually extremely easy to transcribe your meme folder, your book folder or even your video/audio folders via local models like whisper. From there you can easily use another logic model to create tags out of your content and voila, you end up with a fully indexed searchable cloud of your own and you didn't do shit to organize it. This will only become easier as these systems get streamlined together by hobbyists. Even if you think you personally have no use for it now, just wait and see, or share your data with someone that will have use for it.

**How is this political**

Free access to data and knowledge is extremely political as is its decay, deprival and censorship.
>>
bumping potentially decent thread
>>
>>539766853
Seems as though we don't truly understand the implications of what you're talking about. This should be alarming, but almost everyone is either ignoring it or is unaware of the danger that awaits us.
>>
File: 1785164645896453.jpg (79 KB, 1103x1280)
79 KB JPG
>>539768316
Thanks bud

>>539769156
The free and open internet was basically only a thing of the elitist and white early days before scamming jeets and outraged normies got onto it. Ever since that happened we're seeing increasingly dire consequences as everything either needs to be milked for money, censored or destroyed altogether. Anything free and uncensored is counter to the profits of the greedy or the eyes of the outrage mob. The only thing we can currently do to counter that is via backing up data on our own before they manage to shred it forever.

In the near future you will be able to train your own AI on that data that you saved. For now you can use it as your own knowledge base with existing tools. If we don't do that, we will be at the whims of the centralized AI heebs for good.
>>
File: AscendedApu.png (845 KB, 827x699)
845 KB PNG
Anon.....
it's over
stop trying to save the materialistic world
let's escape the samsara, together
>>
>>539769156
I think the main problem is that it's just a huge task. There's so much stuff to go grab and it's a fair bit of effort to get it put somewhere. It's intimidating, makes you not want to start on it.

I guess the main helpful thing I could say is don't worry about organizing it, just grab the data for now. Organization can be done later.

Would be nice to have legitimate libraries for this sort of thing. I always thought it'd be one thing I'd do with a homestead if I had one; get a Starlink, set up some solar panels, build my own earth-bermed data center and start storing stuff in raid arrays in there. Once you've got a way to deal with humidity that kind of cold underground storage can keep stuff intact for a very long time. And if you're out inna woods, you'll be resilient against things like arson or social upheaval.

Although really ideally, what you would want long-term is to convert all your digital storage into physical storage, i.e. print out all your books. Paper doesn't last forever but it lasts longer than any digital storage, doesn't require electricity and you can't lose the ability to read it. Like what if you somehow lost the ability to read PDF files? Video's even worse and it'd be really difficult to print that stuff to film which you could store, if that's even possible.

The more complex any technology is, the more points of failure it has, just a fact of life.
>>
>>539766853
*bump*
>>
>>539766853
Bump based Swiss anon

TAKE NOTE YOU ABJECT FAGGOTS

DOWNLOAD FUCKING EVERYTHING
SAVE EVERYTHING MUSIC VIDEOS WIKIPEDIA AI FUCKING EVERYTHING
>>
>>539766853
Nice thread. I periodically back up Wikipedia and take a gander at archive.org content and /t/ for rare torrents.
I have two NAS back ups of everything.
I aim to contribute and data dump some how when they start to remove the last great places for content retrieval. Others will as well.
>>
>>539766853
>Back up everything
I'm a physical medium whore already. My shit has been "backed up" forever.
>>
>>539770400
I started with full discographies of my favorite artists like Deftones and or labels like waxtrax, then books, just stuff I liked at first. Lovecraft and Phillip K Dick, stuff like that I haven’t yet checked out yet like Asimov. Then what I felt were legendary films like Gremlins 2 the new batch and Brain Candy. Then I started hording apps, good ones like audacity and Winamp. It kept going from there, addi by and more. I could lose the internet forever and have enough content to keep me busy until I die at this point. The emulator and rom collection was my favourite part of building the collection so far.
>>
File: new_burden.png (130 KB, 640x480)
130 KB PNG
>>539770207
I still have some lifeblood to give, Apu. Not yet.

>>539770400
The task is a big as you make it. You can start small with ebooks. PDF or epub format hardly takes up any space and has condensed knowledge. Focus on censored and rare books, stuff that's already on anna's archive is likely there to stay. Aside from that, image archiving is decently important as well. Quite a few of the memes/infographics here are good but also at risk of vanishing. What will really start to pile up is archiving video material such as censored documentaries or entire creator channels on youtube. For that you need the big boy cash if you want to do it for a prolonged time. I have an 80TB mirrored NAS with an attached JBOD for backup for that.

> I guess the main helpful thing I could say is don't worry about organizing it, just grab the data for now. Organization can be done later.
100% and like I said with local AI models it's already possible to auto organize your collections. You can even vibe code it together with a bit of python and a small local model like I laid out in my initial posts.

> Although really ideally, what you would want long-term is to convert all your digital storage into physical storage, i.e. print out all your books.
That would be cool but it will require a lot of space and like you said humidity becomes a problem. Alternatives would be tape drives or well stored Bluray discs (decently cheap now but not much storage).

> Like what if you somehow lost the ability to read PDF files? Video's even worse and it'd be really difficult to print that stuff to film which you could store, if that's even possible.
In theory, at worst, you could keep a cheap smartphone in cold storage in a faraday cage and you would always have a "PDF reader" no matter what. As long as you can still shovel data on there you'd be fine. Question is what are you preparing for? World war 3? EMP bombs? Or just fighting the centralization heebs destroying the internet?
>>
Government shuts down porn. Internet dies. Porn hoarder data becomes the new bitcoin.
>>
>>539770975
>The emulator and rom collection was my favourite part of building the collection so far.
Don't forget the offbeat weird shit too, Pocket Fighter for the Wonderswan is pretty fun.
>>
File: tlu_castle.jpg (281 KB, 1200x1600)
281 KB JPG
>>539770664
> wikipedia dumps
Definitely a must if you want to prepare for shtf scenarios. Not too hard of a hit either at around 25gb.

> I aim to contribute and data dump some how when they start to remove the last great places for content retrieval. Others will as well.
Long term we will need a new platform of sorts. Torrenting works for now but getting torrents and advertising them is an issue. Maybe some sort of distributed image board clone with integrated file sharing would be a good idea. Preferably where you can sync entire collections and not just individual files.

>>539770890
Good job but don't forget to share it if possible.

>>539770975
Definitely check out soulseek if you're into music. Lots of fantastic collections on there and you can message the owners and chat with them. Great place to make friends too. Some people even share rare non-music stuff on there (books, videos, games). I've downloaded a few collections of rare videos of political speeches that I hadn't seen anywhere else before that way.
>>
>>539766853
Did you know that literally like a block away from the library of Alexandria there was a jewish quarter? Really makes you think.
>>
File: siege_of_jerusalem.jpg (55 KB, 600x600)
55 KB JPG
>>539772864
I didn't but I'm also not the least bit surprised if history repeats itself. In dubio culpa iudaeus.
>>
>>539770400
I've been fiddling around with this idea of trying to archive as much as possible for a few years now. Httrack was how I first started doing it but the problem is that it's not able to crawl most modern websites because of the way they're structured. Probably due to server side processing.

Httrack does work very well for Wikipedia though if anyone is interested. I have over 100 gigs of Wikipedia backed up, lol. Sometimes I'll restart the mirroring and let it run over night.
>>
Also just a pointer for lazyfags who may be interested in this (everyone should be).

The singlefile extension has an autosave feature so you can set it to save every page either after loading or when you close a tab. As long as your computer isn't a piece of shit, it runs seamlessly in the background. I have thousands of zipped html files thanks to that very useful feature.
>>
The burning of Alexandria's library is a nothing burger that no one cared about. All of it's books had already been moved to the replacement location and all that remained was worthless copies or slop tomes. What we're facing has no historical precedent.
>>
File: 4theGorillianthTime.jpg (86 KB, 978x703)
86 KB JPG
>>539766853
bump
>>
File: you_fools.jpg (156 KB, 727x1000)
156 KB JPG
>>539774361
For wikipedia you don't need httrack. They provide official dumps. Here's all articles but without pictures(~25gb):
https://dumps.wikimedia.org/enwiki/latest/enwiki-latest-pages-articles.xml.bz2
Here's all articles + pictures (~115gb):
https://lb.download.kiwix.org/zim/wikipedia/ (look for: wikipedia_en_all_maxi_*)

I'd also recommend backing up metapedia but unfortunately they don't provide official dumps. I did notice this here though https://en.metapedia.org/wiki/Special:Export

> httracks is not able to crawl most modern websites
You can use ArchiveBox or Browsertrix if you want to back up websites which use modern faggot JS frameworks. They are quite resource heavy projects though as they emulate virtual browsers.
>>
If you don't know how to make a book, you're not going to make it.
Learn the art of stitching signatures together and how to bind them.
https://www.youtube.com/watch?v=8gc9wnUCfIk
>>
>>539770608
This, tech companies are wiping everything.
>>
File: smaug_treasure.jpg (49 KB, 394x507)
49 KB JPG
>>539776730
We can only hope that they get hacked and the scanned files leak onto the internet. Or maybe there will be some heroic insider who will sacrifice himself. I don't even care that much about the weights but the source material is invaluable. Amodei should be sentenced to scaphism for destroying rare books.
>>
>>539766853
At the end of the day, why should we care? Regular people have access to way more information than they ever had and yet act like worse goycattle than illiterate peasants did a thousand years ago.
>>
>>539777490
And don't get me wrong, I archive what I can, but it's for my own use only. I don't give a shit about what happens once I kick the bucket.
>>
>>539777490
The reason why sharing is important is because it'll make the data resilient. Now as for why you would share it: If not to help others then why not out of spite to those who try to keep it from you? That's half the reason why I do it.

> Regular people
Regular people will be fine enjoying their AI generated crap. It's the above average intellect ones who want to research, create, build and explore.
>>
>>539766853
the problem is
how do you know what you want to keep
if you dont know that it exists that is worth saving ?
like actually usefull stuff ?
and then archived and curated your library
whom are you going to show it to ? normies ?
sorry i am too blackpilled
>>539770608
>MUSIC VIDEOS WIKIPEDIA AI
music videos ? that are 100% cultural subversion ?
wikipedia that is largely biased and has low information value / is completely trashed with tl:dr info so using it becomes even worse than researching by yourself in the actual literature ?
ai that is so bad you cant trust it asking simple questions without having to doublee check and correct everything ?
this is not going to be good without someone who has the knowledge putting in effort and curating the data
what use are 9000 books on woodworking if they all contain information that is not properly matched ?
what do you need textbooks for when you have no documents that show you how to practically do the thing you are reading about ?
this has been kepot apart for a reason by the gatekeepers who maintain the educational system
why do you think are there no more tv shows like telekolleg ( if you even know what that is ) who do practical examples and demonstrate things ?
and even these were too basic so you would lack insight
for the english speaking media there are even less educational video series who transmit practical knowledge, to keep people dependent for knowledge on the institutions that would brainwash them
that is why you need to curate even harder
so just saving anything wont work / produce raw slop archives
you MUST curate what you archive, and INDEX so that the information becomes accessible
nobody profits from good info in a book of 1000+ pages on page 323-342 that is labeled wrong

the archivist must at the same time become a teacher that prepares lessons
>>
good thread
>>
>>539766853
>>
File: franz_von_stuck.jpg (216 KB, 1080x1350)
216 KB JPG
>>539779525
> how do you know what you want to keep
I think it's easy to judge. You found some old pdf of a rare book that isn't on anna's archive? Upload it there and hold onto a copy yourself. Same goes for old memes or infographics if they're rare or the data in the infographics is useful. Same for videos, maybe you had enough insight to backup some youtube channels pre all the ban waves over the years and they contain valuable data. Even images themselves can be useful. For example historic ones like third reich pics in good quality are increasingly not only locked behind paywalls but even behind institutional restrictions for academics only.

> whom are you going to show it to ? normies ?
Depends on your use case. All my images, pdfs, videos, audios, etc in my source files are fed into a database which is easy for me to search. I'm taking part in a discussion about a certain topic? I'll have the tags, OCR and transcripts at the ready to refresh my brain. Likewise I can share that data to anyone who also wants to build their own knowledge base or simply is interested in a few files. The important part is having and sharing the data. Whether you personally use it or just hoard it is not that important to be honest. Eventually local AI will become affordable enough for everyone to just feed it your own data collections and you can curate the AI's "knowledge" away from propaganda and towards truth. Or even train your own from scratch.

> you MUST curate what you archive, and INDEX so that the information becomes accessible
Yes agreed. If you want to actively use the data right now you must use one of the few strategies I laid out in this thread. If someone is too lazy to do that they can use the local model AI chain I recommended or if they're still too lazy they can wait until the complete toolset gets made by someone (maybe me, I already have a version).

> the archivist must at the same time become a teacher that prepares lessons
That would be optimal.
>>
how do you know what is worth saving when you dont know it exists ?

e.g. does someone who never heard of the book industrial society and its future look for it, find it, and archive it to prevent it from being "digitally book burned" ??

why would someone save this ? its just a rant basically it provides no useful information beyond "told you so" and this is basically the same for any book except maybe some
"how to" /diy/ books from the 80s on how to do pottery at home
how to build your own clay turning table with pictures and clear instructions, what tools and materials etc. and then again the problem
how do you know that you even have the book if you have 90000+ books stored and archived ?
can you remember all of that ?

one would need a local llm that has a ocr part which slowly sifts through all the books and memorizes what is in the archive and enables a searchable index, if you one day decide you need info about how to do pottery and make your own clay pots you ask the computer for pottery and it cvan point you to the file
or something similar more crude even
but what good is a database you search for "pottery" and it gives you 500 books that you "archived" that you then ahve to look through one by one until you find the one gem that shows you how to build a clay turning table or a pottery burning oven

does someone have answers to this problem ?
a hint please ?
>>
File: Kaneo salute.jpg (13 KB, 250x140)
13 KB JPG
>>539766853
Already ahead of you boss, I might as well buy two more external HDDs this year
>>
>>539766853
This. Also acquire hardcopies of everything you can. The new Dark Ages are coming.
>>
>>539766853
I saw a similar thread the other day but there anon claimed the world would get EMPd and all electronic devices would be toast (OP there said to stack physical books, to which I replied that it's better to have an e-reader with a 1000 books instead).

This made me wonder if a faraday cage would protect an electronic device from an EMP attack
>>
File: Why not both.gif (917 KB, 338x208)
917 KB GIF
>>539781102
>>
>>539780918
This is why we don't talk about German Monks recording ancient documents. We're not pondering life here.
>>
>>539781102
Here's an easier way to out it. 1/3rd of the Internet has been lost forever.
>>
Use neocities! I made a site a couple years ago and it's fun. Just like the old angelfire and geocities and even early myspace. You can even have your favorite gayi help you with the html and make a really nice looking page. Retvrn to web 1.0! (I was a supporter but I haven't updated my payment so I can't host video atm but the free plan is still awesome.)
>>
File: albert_bierstadt.jpg (23 KB, 390x512)
23 KB JPG
>>539780918
> why would someone save this ?
With technology advancing, especially hard drive space still expanding (albeit slower) we're at the point where unless we're talking some 4 hour long 4k video the cost for data storage is quite trivial. We are in a luxury age where, especially when it comes to books, you can save thousands for pennies, easily. So what should you save? Anything you deem useful, anything you can afford. You're interested in gardening? Wood working? What's keeping you from downloading 5000 books on that topic and feeding it to a local LLM or RAG? If you pick up that hobby in the future and ever have a question just query it locally and it will even spit out the source file for its answer. No propaganda, no lies, just the data you choose.

> can you remember all of that ?
You don't need to remember it. You can create a knowledge base (automated or by hand) or you feed it to a model and query it.

> one would need a local llm that has a ocr part which slowly sifts through all the books and memorizes what is in the archive and enables a searchable index
I've already been doing that for over a year now it works great and will only get easier.

> but what good is a database you search for "pottery" and it gives you 500 books that you "archived"
You're talking about the difference between a plain search and an llm. The later can be queried for interpretative parameters such as "give me the pottery recipe that best describes how to sculpt a human figure".

>>539780941
Saluting you, Ameribro.

>>539781102
EMP's are mostly a danger to plugged in and powered on devices. If you have a HDD unpowered in some storage it will practically already be immune to anything but the most giga power EMP blasts and yes those will still not go through a faraday cage. Hint: Microwaves act like faraday cages. Maybe pick up an old damaged one from your local trash heap if you're frugal and paranoid.
>>
>>539780848
i already have a stack that uses a tool for the few digital pdfs that i have archived and another tool to "read"/scan the mountain of "regular" pdfs, do the tex tthin, feed it to my model and create the index, but there is no way that this will be even remotely useful, as i have mentioned (hopefully before, i think) the information itself is practically useless if you dont have someone (or some thing (doubt)) that is knowledgeable in the topic and can determine if the literature is good / useful or redundant / lowq etc and then distill it to produce a useful thread from it
you can do so many books on origami index them, catalog them etc, but when someone with no knowledge needs the info about origami, they need the real deal experience not, just fold paper like so and so, this goes for basically any topic, repairing things, making things, etc, its good to have deep literature such as professional engineering texts, standards, etc, but these act as reference for professionals, while the information in them is super important for anyone, a beginner will hardly be able to get that info from there.

sorry i am jus ttired

thanks for the thread tho, its important
i just need a break
>>
File: 1784728194439461.jpg (55 KB, 458x466)
55 KB JPG
>>539780918
someone who asks ten questions at the same time isn't looking for answers, they are trying to blackpill.
>>
>>539781102
Thats one of the points of a faraday cage.
If you were truly autistic and dedicated to this task, you would have a small extremely well build faraday cage as a storage box, and you would put hard drives and shit in there. Then all you have to do is get electricity going again for yourself and you have your private data intact
>>
>>539781663
(((they))) want to erase knowledge so it's easier for them to lie. i don't care about origami but history is pretty important. and i don't care if 90% of it is propaganda, by crossreferencing you can usually determine what's true and what's not.
>>
>>539766853
Way ahead of you.
I had a weird feeling a month ago I should start rebuilding my library again. Mostly primary source histories, older Based books (((they))) don't want you reading, political philosophy, religious works, etc..

What I can't get in physical is all separately stored in secure places. I keep older machines that can still process formats from their time.

Make totally sure you're saving art and images, too, as well as music and film in unaltered states. Everything is .webp format now to prevent image sharing. Streaming services are always going to comply with any edits made to older versions of film or music. Audiobooks also aren't safe. They will yank it from your library any time they feel like it. If you play vidya, start ripping things from Steam or switch to other platforms to get digital copies. Physical ones are impossible to get for many of them.

If you think it's totally nebulous the fact AI companies are literally cannibalizing physical books and then saving digital scans and they won't use AI or other software to edit them to remove critical passages you are hopelessly naive and ngmi.
>>
>>539781102
>This made me wonder if a faraday cage would protect an electronic device from an EMP attack
yes you can 100% store a tablet computer inside a EMF shielded container, build it from muMetal if possible ( if you can get large enough sheets ) or copper sheets, you basically have to solder your device in for after the event, if i would do that i would chose one of the older models where you could easily dismantle it and swap the power pack/barttery because storing the abttery for a long time without charging / discharging will deteriorate it. and you maybe will think about alternate power sources that are easier to replace than those specialised batteries especially after shtf, but to protect the ICs of the device from EM you can definitely do that, for a start look into how to shield a electric guitar ( commonly done by instrument builders ) using mu metal / copper / foil and solder, same principle
>>
File: Talmudic_Jew_Cover.jpg (197 KB, 400x610)
197 KB JPG
>>539781663
keeping religious texts in the public domain is also important. pic related.
>>
>>539781159
Because physical books are big and heavy, I have like 30 and they take up an entire travel bag and are heavy as all shit. Also I am planning to move to another city and live in my car for a while so I can't take the books, but I can take my e-reader
>>
>>
>>539766853
good thread, thanks OP
>>
>>539781711
no i am having my own problem
i forget
and i am then unable to remember because my knowledge is gone
i build a vague memory of seemingly trivial datapoints that somehow resurface like driftwood in a river
so i tried if i can automate the computer to take a look at all the ebooks / texts i have saved over almost two decades i can at least create a searchable index that somehow limits the area where i need to look further for info
but then i created just another layer that i need to dive deeper
i think i am suffering from information overflow or something like that
how do you know that something is "not normal" without any reference ?
only after years, decades you eventually become able to realise that while doing all that all the time without knowing
i am not intentionally trying to blackpill anyone, i guess if i am coming off like that it is because i am blackpilled
sorry to blackpill you
goodbye
>>
>>539766908
>web 2.0 is dying a slow death
>there will be no Web 3.0
I’m not saying I disagree with you but wouldn’t the technocratic oligarchy that’s currently running the country (palantir, SpaceX, etc) be intrinsically tied to the rollout of a Web 3.0?
>>
>>539766853
>- Archive.org crawler is incomplete/doesn't archive exclusions/removes content upon request of (((certain organisations))). **Fine uh... well shit.**
And Archive.org is ran by some obscure character of unknown allegiance, he may very well be a glowie.
> Back up everything worth saving. Books, websites, images, videos, audio, anything you can.
And stick to physical media, secure books and guard them, as well.
Jews are infamous for corrupting certain texts in order to advance their narrative.
>>
>>539766853
good. jew media and jewish subversive knowledge needs to be eradicated again
>>
>>539766908
No, there needs to be a decantrilized way for establish data authenticity.
Maybe multiple anons can take a hash of a single source using some standarised way and collectively validate it across the wire.
Ultimately all of this is a consequence of the death of physical media.
>>
File: thomas_cole.jpg (950 KB, 2561x1730)
950 KB JPG
>>539781663
I think you're mixing up two or three different goals here. One is data preservation which is the main goal. Another is personal usefulness of that data which is what you went into and then a third one which you seem to bring up is distilling your own model via weights and curation from said sources. You should keep those things apart and it would feel less overwhelming to you. Because all of these are solvable. But only the data preservation is the most urgent task at the moment.

>>539781948
Yes good points. Some truth will also self distill itself simply out of large data sets.

>>539781956
Yes definitely back up music and audiobooks. All those services like Audible, Kindle, Spotify won't give you anything and at any time can cut you off. You need the real files on your own storage.

> Everything is .webp format now to prevent image sharing.
You mean sharing it on imageboards? You can convert it in theory but that's a hassle.

> to edit them to remove critical passages you are hopelessly naive
This has already happened with Israel paying OpenAI for positive propaganda for them. It will only get worse.

>>539782112
I gave up on physical books. I used to collect rare ones but it became a hassle as space ran out quickly. Started scanning them, uploading to anna's archive and donating the books to interested friends or reselling them.

>>539782086
JR Books is fantastic. Highly recommended. You might also want to check https://sacred-texts.com/index.htm if you're into religious/spiritual texts.
>>
>>539781956
Webp can be converted, fyi.
And while it is shitty for sharing, it will work when opened up with apps
>>
File: richard_moult.jpg (30 KB, 452x442)
30 KB JPG
>>539782345
Welcome, Burgerbro.

>>539782414
You're suffering from analysis paralysis. Take a break and then calmly plan out step by step what you want to do and you will succeed. If you're having issues, (ironically), ask an LLM to guide you.

>>539782540
Web 3.0 will be apps and monolithic AI. They have no need for an open web when they can discard it entirely and keep the customer in an ever surveilled box that they then also get paid for.

>>539782730
> No, there needs to be a decantrilized way for establish data authenticity. Maybe multiple anons can take a hash of a single source using some standarised way and collectively validate it across the wire. Ultimately all of this is a consequence of the death of physical media.
Fascinating because I had a similar idea a few years ago. The problem is you could still game such a system unless you whitelist who you let curate. Otherwise bots could ruin it easily. But a centralized system made up of individual clients, maybe over torrent or IPFS would be very helpful. Then instead of tripcodes you sign with secure keys and you could whitelist which anons you deem trustworthy. Who knows. I think it's feasible but would need a lot of thought.
>>
>>539783419
>But a centralized system made up of individual clients, maybe over torrent or IPFS would be very helpful. Then instead of tripcodes you sign with secure keys and you could whitelist which anons you deem trustworthy. Who knows. I think it's feasible but would need a lot of thought.
A centralized system is not a good idea, since it requires a lot of trust, which doesn't come by so easily on the internet.
And while such a system can be gamed by bots, you can theoretically even do data validation in real life without the internet even.
If you wanted to verify the authenticity of a dataset, find a bunch of anons you trust and compare your hashes.
It's just a way to ensure that the data you are holding is legit and has not been corrupted.
Corruption is worse than censorship.
>>
>>539783419
>Web 3.0 will be apps and monolithic AI. They have no need for an open web when they can discard it entirely and keep the customer in an ever surveilled box that they then also get paid for
There needs to be a cultural change in altitude towards the internet as a whole, it must no longer be regarded as some glorious uncensored libertarian utopia like it used to be at some point.
Today, western cultural trends are driven by algos under the control of a handful of big tech kike oligarchs, you can't fight them or build alternatives since they have a shitton of money behind them, delegitimizing them is therefor the best move.
>>
>>539766853
There was a very obscure wiki I came across late last year, some real niche-ass info on there. FF a couple months and there's a message from the admin saying he shut it down for now due to LLM crawling the shit out of his site or something. The message points to web-archived copies of the site that do not even exist

Pissed me off ngl, if it ever goes back online I'll save it wholesale for sure. Or maybe I'll get in touch w/the admin, I really want to skim thru the rest of that wiki's contents
>>
>>539783101
There's no problem with saving the format itself. No idea why this guy says webp is to prevent sharing. Pretty sure it's just a newer slightly inefficient compression technique.
>>
File: richard_moult.jpg (210 KB, 1200x997)
210 KB JPG
>>539784283
Not centralized like a traditional server, but multiple clients making up a whole. You would still need trackers though, or something like DHT. Or how would people exchange data? Completely samizdat style "it's your job to exchange dumps or usb sticks" style? I think it would need to be a fusion of both an image board AND a file sharing platform with integrated file curation/validation.

> And while such a system can be gamed by bots, you can theoretically even do data validation in real life without the internet even.
The image board part would help there. Say you and me trust each other we could whitelist each other's keys, maybe set a note or even link the key whitelisting to a thread/post we made and voila we have trust established securely and easily (and we can always revoke it).

> It's just a way to ensure that the data you are holding is legit and has not been corrupted.
Legit as in signed data that you know is coming from a whitelisted user? Or as in using the file hash to ensure a file was not tampered with? An user could sign their "validations" with a key and if the key is whitelisted (trusted) by you they get added to your database on the corresponding file. Say "User YYYY: This video is AI generated, do not believe it". Something like that?

>>539784476
> There needs to be a cultural change in altitude towards the internet as a whole
I don't see normies changing unless the grass is greener somewhere else unfortunately.

> delegitimizing them is therefor the best move.
How would you delegitimize them though? You would need to have propaganda/media power of your own to even reach the masses. Or a really cleverly crafted campaign that hits people on an emotional level (since most people are emotional thinkers).

>>539784659
You can find similar websites and wikis on yandex. They index some curious websites on there that you will never find otherwise. Too bad about that wiki.

>>539784738
If 4ch would just support it it wouldn't be an issue.
>>
>>539766853
Stacher is a good GUI for yt-dlp if you don't mind it being a tracker like everything else easy to use.



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.