Profile pic

Mike, michael@piefed.chrisco.me

Instance: piefed.chrisco.me
Joined: 7 months ago
Posts: 12
Comments: 93

    RSS feed

    Posts and Comments by Mike, michael@piefed.chrisco.me





    Whats even better is that the squirrel can stop it from attacking for at least 15 turns.

    Im just imagining something like Cthulhu coming out about to destroy the word and a squirrel hurls itself at it….and now it has to wait another year for another go.


    First time Ive heard about Milk-V Duo S. Interesting. Looks very cheap!



    Piped.chrisco.me

    I removed a couple. Mostly porn. I dont really need that for the family. But that was 2 servers out of hundreds.


    So!! I went from 60 instances to over 900 in the span of less than 24 hours! Also most of the content is pretty good. So thats nice.


    Hell yeah! Thats an awesome setup.



    Also if you want to add my little instance and see more dog videos: https://piped.chrisco.me/c/teddy_the_dog/videos feel free to add

    https://piped.chrisco.me

    to the list


    So im not seeing the option for “Automatically follow platforms of a public index” after federation.

    image

    I believe(?) I found it under Settings->Configuration->Basic:

    Automatically follow instances of a public index
    ⚠️ This functionality requires a lot of attention and extra moderation.
    See the documentation for more information about the expected URL

    Im going to replace the main https://instances.joinpeertube.org with your index and see what happens.


    Ill give it a shot and let you know how it goes!


    Or make your own. We are pretty flexable.


    Looking forward to the new book! DCC 😁

     reply
    4



    Im not sure. I agree with you.

    All I can tell you is that my server was hammered by a couple of IP addresses. When I did a lookup they were ALL openai and they just would not stop. Until I added protection that is. Then they got bogged down and eventually stopped.

    Ive heard talk that their engineers said they dont need to abide by robots.txt since they are not an indexer like google. Which is BS.


    The bigger companies are looking at robots.txt to see if they can scan your stuff for AI scraping purposes. I get a couple from google, bing and others. Im not sure about facebook, but I do see the bigger ones usually abide by robots.txt and stop there. It doesn’t stop them from hammering your robots.txt though.

    If you get fail2ban and/or block the ip range from one actor, it usually goes away.

    The worst offenders is openai which does NOT hit robots.txt and just scrapes/DDOS my small site. Until I put in a couple of infinite loop/nefarious solutions on the server. Then you have fail2ban see what ip addresses try and go deep and block them.


    RSS feed

    Posts by Mike, michael@piefed.chrisco.me

    Comments by Mike, michael@piefed.chrisco.me





    Whats even better is that the squirrel can stop it from attacking for at least 15 turns.

    Im just imagining something like Cthulhu coming out about to destroy the word and a squirrel hurls itself at it….and now it has to wait another year for another go.


    First time Ive heard about Milk-V Duo S. Interesting. Looks very cheap!



    Piped.chrisco.me

    I removed a couple. Mostly porn. I dont really need that for the family. But that was 2 servers out of hundreds.


    So!! I went from 60 instances to over 900 in the span of less than 24 hours! Also most of the content is pretty good. So thats nice.


    Hell yeah! Thats an awesome setup.



    Also if you want to add my little instance and see more dog videos: https://piped.chrisco.me/c/teddy_the_dog/videos feel free to add

    https://piped.chrisco.me

    to the list


    So im not seeing the option for “Automatically follow platforms of a public index” after federation.

    image

    I believe(?) I found it under Settings->Configuration->Basic:

    Automatically follow instances of a public index
    ⚠️ This functionality requires a lot of attention and extra moderation.
    See the documentation for more information about the expected URL

    Im going to replace the main https://instances.joinpeertube.org with your index and see what happens.


    Ill give it a shot and let you know how it goes!


    Or make your own. We are pretty flexable.


    Looking forward to the new book! DCC 😁

     reply
    4



    Im not sure. I agree with you.

    All I can tell you is that my server was hammered by a couple of IP addresses. When I did a lookup they were ALL openai and they just would not stop. Until I added protection that is. Then they got bogged down and eventually stopped.

    Ive heard talk that their engineers said they dont need to abide by robots.txt since they are not an indexer like google. Which is BS.


    The bigger companies are looking at robots.txt to see if they can scan your stuff for AI scraping purposes. I get a couple from google, bing and others. Im not sure about facebook, but I do see the bigger ones usually abide by robots.txt and stop there. It doesn’t stop them from hammering your robots.txt though.

    If you get fail2ban and/or block the ip range from one actor, it usually goes away.

    The worst offenders is openai which does NOT hit robots.txt and just scrapes/DDOS my small site. Until I put in a couple of infinite loop/nefarious solutions on the server. Then you have fail2ban see what ip addresses try and go deep and block them.