• Usenet archive, TAFKAC replacement?

    From Thomas Prufer@prufer.public@mnet-online.de.invalid to alt.folklore.urban on Wed Sep 9 09:39:50 2026
    From Newsgroup: alt.folklore.urban

    Took this for a quick test drive: I think it works better than the Google offering...


    Thomas Prufer


    On Sun, 06 Sep 2026 06:12:18 GMT, in alt.folklore.computers Craig Stadler <cstadler18@hotmail.com> wrote:

    I spent the last year gathering and normalizing as many resources for
    Usenet messages and built a search engine over public Usenet
    discussions going back to 1981. Posting here because this seemed like
    the one group that would actually care.

    https://www.usenet-rewind.com/

    Since Google Groups stopped indexing new Usenet content a
    while back, and search over its historical archive has always been
    rough- I wanted something that treats this as a real research
    archive -- full-text search, date-range/filtering,etc and most
    importantly as much Usenet data as possible going back as far as I could >find.

    The index currently holds (to date) roughly 980 million messages and is
    still growing. Sources are archival backups, data donations, and ongoing >crawls of several thousand still-active public NNTP servers. Binary and >yEnc-encoded content is stripped where possible to keep the index
    focused on text.

    The index was built using Apache Solr 10, MariaDB 12 (rocksdb), Ubuntu >22.04.5 LTS & custom Python scripts, all on nvme disk.

    This is public archival content, same legal basis as DejaNews and
    Google Groups before it. There's a removal process for anyone who
    wants their own posts taken down, and author contact info is masked
    by default.
    --- Synchronet 3.22a-Linux NewsLink 1.2
  • From Don Freeman@Don@cosmoslair.com to alt.folklore.urban on Sun Sep 13 17:26:24 2026
    From Newsgroup: alt.folklore.urban

    I tried it and went down several rabbit holes while doing so. Used up my
    25 free searches (only same day limitation). But I'll be back. And
    definitely better than what Google had set up.

    On 9/9/2026 1:39 AM, Thomas Prufer wrote:
    Took this for a quick test drive: I think it works better than the Google offering...


    Thomas Prufer


    On Sun, 06 Sep 2026 06:12:18 GMT, in alt.folklore.computers Craig Stadler <cstadler18@hotmail.com> wrote:

    I spent the last year gathering and normalizing as many resources for
    Usenet messages and built a search engine over public Usenet
    discussions going back to 1981. Posting here because this seemed like
    the one group that would actually care.

    https://www.usenet-rewind.com/

    Since Google Groups stopped indexing new Usenet content a
    while back, and search over its historical archive has always been
    rough- I wanted something that treats this as a real research
    archive -- full-text search, date-range/filtering,etc and most
    importantly as much Usenet data as possible going back as far as I could
    find.

    The index currently holds (to date) roughly 980 million messages and is
    still growing. Sources are archival backups, data donations, and ongoing
    crawls of several thousand still-active public NNTP servers. Binary and
    yEnc-encoded content is stripped where possible to keep the index
    focused on text.

    The index was built using Apache Solr 10, MariaDB 12 (rocksdb), Ubuntu
    22.04.5 LTS & custom Python scripts, all on nvme disk.

    This is public archival content, same legal basis as DejaNews and
    Google Groups before it. There's a removal process for anyone who
    wants their own posts taken down, and author contact info is masked
    by default.
    --
    -a-a__
    (oO) rCarCarCa*www.cosmoslair.com*
    /||\ rCaCthulhu Saves!!! (In case he needs a midnight snack)
    --- Synchronet 3.22a-Linux NewsLink 1.2