• 688 Posts
  • 13.3K Comments
Joined 3 years ago
Aquileo | cake
Cake day: October 4th, 2023

Aquileo | help-circle



  • taltoLemmyTodayTesting server load with login required on old.lemmy.today
    Aquileo | link
    Aquileo | fedilink
    English
    Aquileo | arrow-up
    5
    ·
    Aquileo | edit-2
    13 hours ago

    Yeah, I don’t want the whole world of social media to wind up in a “get an account everywhere to view anything” situation.

    I think that it’s a reasonable stopgap fix from an admin standpoint, because there aren’t a whole lot of great levers for admins to pull to try to deal with bots beating the hell out of a server. And one can’t just have one’s server become unusable.

    But it should really be something that is addressed on the dev side. Like, we need a real, long-term solution for the Web.

    Also, while improving server performance and failure modes under load would help, the “badly-written, aggressive scraper bots that ignore robots.txt and are given loads of network resources clobbering servers” is something that affects many, many different Web servers out there. This isn’t a Lemmy problem, nor even just social media problem. It’s a Web problem.

    I think what really needs to happen is some form of mechanism that makes it hard for purely-automated systems to hammer a server, something that segregates bots and humans. Maybe a CAPTCHA or other proof-of-resource thing or something like that.

    It’s going to also create server admin problems. Think of things like Web search engines indexing sites — server admins normally want their servers to be indexed by “good” bots — so it’s going to create additional headache in that even after successfully creating a system that is able to segregate bots and humans, admins have to be able to identify “good” bots — maybe “good” IP ranges or have “good” bots use public keys or something — and then whitelist those. Maybe have some mechanism to distribute “good bot” whitelists and an Apache mod that downloads them occasionally, lets an admin subscribe to a whitelist or something.

    I also kind of like the ability to occasionally wget -r sections of sites, and a solution of the above sort will break that without logging in in a browser and handing off browser cookies to wget, which is not ideal.

    And there’s some resource cost to humans — just lower — for proof-of-resource solutions, and human time cost for CAPTCHAs, all of which are undesirable.

    considers

    One unorthodox solution might be making it easier to snapshot a site. Like, okay. One of the things that irritates the hell out of me is that the bots are doing this to a number of websites that go out of their way to make it really easy to distribute the data that they’re after.

    Like, GitHub — which has lots of open-source code, which is useful to train coding AI models — ran into load problems with bots scraping it. That was induced by the same data being massively downloaded over via inefficient Web views. Bots would follow every single link, many of which displayed the same data formatted slightly differently. But…you can already get all of the code far more efficiently, which would be better for GitHub and better for bot operators, by just using git clone to pull the code (and all of its history!). Bot operators could just say “oh, this is GitHub” and just git clone everything and the load would be completely ignorable. I’m sure that people have mass-git-cloned GitHub lots of times before. The reason bot operators presumably aren’t doing it is because it takes some amount of time and dev effort on their side compared to just “build generic Web spider and clobber everything”.

    Same thing for the Threadiverse. The Threadiverse will let you set up a Threadiverse instance and subscribe to everything, efficiently feed all the posts and comments you want to your instance, the moment they come in. In nice, machine-readable form, rather than in something intended for humans that you have to scrape and post-process. But…it takes more dev effort to set up something specific to the Threadiverse than to just treat it like another website.

    If there were some sort of widely-adopted API for “request site snapshot” or “request snapshot of changes since time X”, widely-enough that it were worth using, maybe bot operators would use that instead. If they don’t have to write a “detect GitHub site and git clone” and “detect Lemmy and subscribe to posts and comments” system, but just have a single universal “dump changes” API, they might use that; less dev effort on their end.

    considers

    I guess the problem is that some website operators might treat the “snapshot API” itself as an opportunity to discriminate between bots and humans, and just return garbage or nothing as a snapshot. Some websites don’t want to be scraped at all. If many websites returned incorrect data in response to such an API request, that’d kill the incentive of bot operators to use such an API.


  • taltoLemmyTodayTesting server load with login required on old.lemmy.today
    Aquileo | link
    Aquileo | fedilink
    English
    Aquileo | arrow-up
    3
    ·
    Aquileo | edit-2
    14 hours ago

    It’s probably not directly related to these changes if you’re only doing it Saturday and it’s only on old.lemmy.com, but if you’re looking into load problems, might be indirectly relevant. In the past…I think two days, using the standard Lemmy Web UI, I’ve seen a high rate of the server returning a page just reading:

    Error!

    There was an error on the server. Try refreshing your browser. If that doesn’t work, come back at a later time. If the problem persists, you can seek help in the Lemmy support community or Lemmy Matrix room.

    Sometimes, but not always, the above error text is also followed by:

    The server returned this error: Error. This may be useful for admins and developers to diagnose and fix the error

    This has not, as far as I can tell, shown up when trying to view lists of posts (like, when I view all subscribed posts) but it does show up when trying to view the page for a post. It appears to affect all posts for me, not just specific ones.

    I initially thought that its just that the post had been deleted, but I could view said posts fine on the original, remote instance.

    One such example that I just hit, tried to load a number of times unsuccessfully:

    https://lemmy.today/post/57605044

    Which has, as a remote post:

    https://nord.pub/c/games@lemmy.world/p/326761/is-routine-worth-playing

    I reloaded five or six times, and eventually it came up locally on lemmy.today.

    When it happens, it seems to affect all of the post pages that I try to load.

    When I see this, the error page comes up pretty quickly, maybe in a second, so it’s not some (long) timeout. Maybe hitting some limit on concurrent database queries or something like that?

    tries loading page again

    And now that post is back to showing an error. Also, I saw it without doing a cache-invalidating reload on Firefox (shift-reload rather than a normal reload), so whatever the server is trying to generate, and failing, it’s not driven by the server trying to respond to something that my browser will have cached.

    When the problem is happening, the same problem also shows up when I try to view my user profile page (so doesn’t as best I can tell, affect lists of posts, either “all subscribed” or viewing posts in a community, but does affect individual post pages and does affect user profile pages).


  • There used to be a convention of having authors who would, at least occasionally, write deliberately-enraging opinion articles in magazines. In my experience, they were typically at the end of magazines. Maybe not outright wrong articles, but they’d make statements that many readers of the magazine would object to.

    The idea, as I understand it, was to — and this was before social media, in the Internet-based sense of the word, developed a focus on this — drive engagement. People would write letters to the editor complaining and arguing against it, and then the magazine could print some of these in the next issue in the letters section.

    John C. Dvorak was the most-prominent example of this that I recall — he put articles in a few computer magazines — and was notable in that I recall him publicly-admitting on one occasion that he did this, that it wasn’t just that he happened to have controversial opinions.

    Huh. Looking at his WP article, it looks like he just died last week.

    Anyway, my point is that I could believe that publications might still have authors who do this sort of thing.



  • I personally would much rather have a PC, given a choice between having one or the other. When I have had both in the past, I have always wound up with a console that’s gathering dust.

    However, that depends on what you’re doing with it and what you want. Consoles have their own strong points:

    • Less setup/troubleshooting/maintenance.

    • Because the platform is locked down, harder for other players to cheat in multiplayer competitive online games. The PC is open.

    • Everyone has more-or-less the same hardware, which is typically desirable for 3D multiplayer competitive video games. No “get a more expensive and powerful video card to get a slight edge”.

    • There is basically one standard controller, so games in the game library are more likely to support the features of your controller (assuming, of course, that you’re using a controller on the PC).

    • Some platform-exclusive games (and obviously that cuts both ways, and the PC library is much larger, but if you very specially want some of the PS5 exclusives, you need a PS5 to do it).




  • taltoAsk Lemmy@lemmy.worldIf you could have a small farm, what would you grow/keep?
    Aquileo | link
    Aquileo | fedilink
    English
    Aquileo | arrow-up
    1
    ·
    Aquileo | edit-2
    1 day ago

    I don’t know about a whole farm, but I suppose that the plant that I’d most personally be likely to grow would be some sort of hot peppers.

    https://en.wikipedia.org/wiki/List_of_Capsicum_cultivars

    Not a lot of work, and nice fresh in small quantities — so having some growing there is nice. Plus, it’s something where a grocery store is only gonna have a limited selection of options relative to what’s out there.

    Whether that’s a good choice is gonna depend on climate where one lives.

    I also kind of like pineapple guavas. Those are hard to find in stores (unavailable in many places) as they don’t ship well.

    https://en.wikipedia.org/wiki/Feijoa

    If I didn’t personally have to take care of them, probably a variety of fruit trees, grapes, berries, maybe squash. Dad planted stuff around the back yard when I was a kid planned so that there was always something edible out there year-round. Grafted stuff on there too, so you had plumcots on one of the apricot trees, a peach tree that was half yellow and half white, a bunch of types of apples and plums (IIRC seven types on one tree) and so forth. It was neat from a consuming standpoint, but I wouldn’t personally want to deal with caring for each. If you enjoy the hobby, though…shrugs More fun than just having a ton of one thing. Not optimal from a cash crop standpoint, though.

    EDIT: Oh, strawberries. Not that they’re hard to find, but store-bought strawberries are really watery, as they’re sold by weight or something like that. If you’ve ever had strawberries that aren’t as heavily watered, like, wild strawberries, it’s a neat experience. Much stronger flavor.


  • We can desalinate all the water we want in California — we already have a bunch of desalination facilities, and can build those out as far as need be — but desalinated water isn’t as cheap as already-fresh water. That’s mostly an annoyance for, say, residential use. But it’s harder to competitively do agriculture on desalinated water. Israel has done some limited agriculture on desalinated water, but farmers are not generally going to be okay with that as an option. Most agriculture uses a lot of water.

    searches

    Someone ran the numbers to see roughly what it would cost if we did only desalinated water for everyone in the US.

    https://hannahritchie.substack.com/p/how-much-energy-does-desalinisation

    Producing enough drinking water for someone — assuming 3 litres per day — costs just $2.30 for the entire year. That’s less than the cost of a single bottle of water in many countries.

    We actually use way more water than we do for drinking — we’re pretty prolific with water to do things like lawns.

    If the average person in the United States uses 310 litres of water per day for domestic uses, it would cost them $0.42 per day. Or $154 per year.

    We’d probably want to conserve more nationally in such a scenario (Californians use less than the national average). But even with no conservation, it’s very much doable.

    It’s a cost that we’d rather not pay if we don’t have to, as a society, but we certainly are able to do so for residential needs.

    But if there isn’t enough freshwater to do agriculture, then who gets the freshwater determines which farmers go out of business (or at least have to shift to something else) and which get to keep operating.






  • taltolinuxmemes@lemmy.worldSpeed Dating
    Aquileo | link
    Aquileo | fedilink
    English
    Aquileo | arrow-up
    14
    ·
    Aquileo | edit-2
    3 days ago

    She went for him in late 1993:

    https://en.wikipedia.org/wiki/Linus_Torvalds

    Linus Torvalds is married to Tove Torvalds (née Monni), a six-time Finnish national karate champion, whom he met in late 1993. He was running introductory computer laboratory exercises for students and instructed the course attendees to send him an e-mail as a test, to which Tove responded with an e-mail asking for a date.

    Fedora won’t have been around.

    https://en.wikipedia.org/wiki/Fedora_Linux

    Fedora Linux[10] is a Linux distribution developed by the Fedora Project. It was originally developed in 2003 as a continuation of the Red Hat Linux project.

    It looks like there were Linux distributions at the time, but they were very early:

    https://en.wikipedia.org/wiki/Linux_distribution

    Early distributions included:

    • Torvalds’ “Boot-Root” images, from v0.95a onwards “Root” was maintained by Jim Winstead Jr., the aforementioned disk image pair with the kernel and the absolute minimal tools to get started (0.10: 4 November 1991)[12][13][14][15]

    • MCC Interim Linux (3 March 1992)[16]

    • TAMU Linux, based on MCC with simpler installation and X11 (14 July 1992)[17]

    • Softlanding Linux System (SLS) which included the X Window System and was the most comprehensive distribution for a short time (15 August 1992)[18]

    • H.J. Lu’s “bootable rootdisks” (23 September 1992),[19][20] and “Linux Base System” (5 October 1992)[21][22]

    • Yggdrasil Linux/GNU/X, a commercial distribution (alpha: 8 December 1992)

    • The two oldest, still active distribution projects started in 1993. The SLS distribution was not well maintained, so in July 1993 a new SLS-based distribution, Slackware, was released by Patrick Volkerding.[23] Also dissatisfied with SLS, Ian Murdock set to create a free distribution by founding Debian in August 1993, with first public BETA released in January 1994 and first stable version in June 1996.[24][25]

    I think that there are two possible takeaways from this:

    • Not clear what distro was used — if any — but if it was, it won’t have been a presently existing one. It’s possible that he was just rolling his own setup, not using a distro at all.

    • Linus’s date with Tove derived from an email exchange. What ninja chicks look for in a guy is possibly an email client, not a Linux distribution at all. It sounds like Linus was probably using Pine at the time (though someone could go dig up mail headers from old emails to check).




  • On further thought, one edge use case that might make sense — suppose you’re a multiplayer online gamer who is only interested in getting an edge in multiplayer online games and don’t really care at all about what a game looks like. OLED displays can offer considerably higher reasonable framerates than past monitors, maybe 500 Hz or more. You might be able to crank the texture settings in a number of games all the way down, just look at a blurry mess, and use an OLED display and have a video card that can, as long as it isn’t crashing into memory limitations, render high framerates.

    There are also a handful of games that don’t use textures on polygons, sort of the Avara or Star Fox style. Most of those are older and target systems before hardware acceleration, but I can think of a few newer ones. I personally like Carrier Command 2*, and that has almost-entirely untextured polygons, and where it does use textures, they’re low resolution. I don’t know what the VRAM requirements are for the geometry, but the listed system requirements on Steam are for 1GB VRAM cards. I don’t think that any of those are actually games where framerate matters much, but if you wanted to render those at very high resolution or on multiple monitors or something (if the game can support that?), it might make sense.

    I mean, this is really just a mental exercise — I don’t think that the OEM in question is really trying to develop a system that’s actually well-suited to any entry-level-gaming users. But if you were looking for very specialized gaming niches, you could maybe find something that the system could actually serve well at.