Upcoming 0.18 upgrade, 404 errors and infrastructure costs

sunaurus@lemm.ee · edit-2 1 year ago

Upcoming 0.18 upgrade, 404 errors and infrastructure costs

Jenga@lemm.ee · 1 year ago

I’m new to this instance, but loving the communication so far! Definitely happy with pinned posts to share updates like this.

InEnduringGrowStrong@lemm.ee · 1 year ago

Choosing the “right” home instance isn’t too obvious as a noob.
This transparency, not only on funding, but also at the technical level is very nice.

marshoepial@lemm.ee · 1 year ago

I don’t mind the pinned posts! I like keeping up to date on the lemmee goings on

kneelknee 🐖@lemm.ee · 1 year ago

I also don’t mind the pinned posts. I like to know if anything new is going on, and it’s pretty easy to just scroll past them when there isn’t. If these pinned posts do get unpinned after a certain amount of time, I’d prefer a little longer (maybe 48 hours?)

coliseum@lemm.ee · 1 year ago

Thank you for having a working instance. I tried to sign up on so many others, and they just plain did not function. It got so bad that I considered running my own…

mmhmm@lemm.ee · 1 year ago

💯this is basically how i found lemm.ee

Brad@lemm.ee · 1 year ago

I like the pinned posts personally. I like to keep informed about what’s happening & what your plans are. Also thank you to the folks that have already donated, I hope I’ll have the money in the coming months to also be a regular contributor to the network.

holgersson@lemm.ee · 1 year ago

Is it possible to also support you by other means than donations aswell? Im a DevOps engineer and could possibly help with Cloud infra, observability or automation.

Thank you for your work so far! I actually like the pinned updates, they show that you care about the server and transparency.

sunaurus@lemm.ee · 1 year ago

Thanks a lot for the offer! The operational work is totally manageable for me so far, but if that changes, then I’ll definitely reach out to the community to try and find some help.

Notorious@lemm.ee · 1 year ago

Hey @sunaurus I really appreciate all of your hard work. I already sponsored you on Github, but money only helps so much - don’t burn yourself out. If you need anyone to lend a hand please don’t hesitate to ask. Infra is one of my specialties, but happy to mod or whatever else you need.

Prefix@lemm.ee · 1 year ago

+1! I am not an infra guide, but a SWE by trade. Happy to lend a hand however I can.

killjoy@lemm.ee · 1 year ago

New here but your communications and transparency have felt like a warm hug.

Thank you for everything!

TWeaK@lemm.ee · 1 year ago

As of today, we have 4 supporters who have signed up for monthly (!!) contributions on my GitHub sponsors as well as one supporter who has donated money through my Ko-Fi page.

What’s your preferred method for donations?

sunaurus@lemm.ee · 1 year ago

Thanks for asking! The difference is not huge, but Ko-Fi does take a tiny bit more fees than GitHub sponsors (especially for any recurring payments Ko-Fi takes 5%), so GitHub would be the preferred option for now.

OneDimensionPrinter@lemm.ee · 1 year ago

Just signed up for a monthly thing on GitHub. Did a ko-fi donation before I saw this comment though. But here’s hoping we can get you more than enough to fund the instance!!

sunaurus@lemm.ee · 1 year ago

Thank you mate, that’s much appreciated!

Ledditor@leddit.minnal.icu · 1 year ago

Great Job @[email protected] !!

borstis@lemm.ee · 1 year ago

Additional posts will no longer automatically appear in your feeds while you’re scrolling

Oh yeah!

sleepyducky@lemm.ee · 1 year ago

Thank you for everything you have been doing! I am loving it here so far! I enjoy the pinned posts. I might be missing a regular update post when I am working but with the pinned posts I don’t feel like I am missing anything. I love the transparency of it all

FerretOnLemmy@lemm.ee · 1 year ago

I’m guessing you’re a Reddit refugee like me, and yeah Lemmy seems pretty chill so far. I hope only the good people of Reddit migrate here tho, Reddit was mostly porn and we don’t need that crap here

JakeBacon@lemm.ee · 1 year ago

Thankfully, it’s a little more convenient to block out porn on Lemmy as the majority of it comes from the one instance.

BeegYoshi@lemm.ee · 1 year ago

Is there a way for users to block instances? I’ve only been able to figure out how to block individual communities

thegiddystitcher@lemm.ee · 1 year ago

Don’t believe so yet. Maybe add it to the long list of requests on Github if it’s not there already?

barsoap@lemm.ee · edit-2 1 year ago

To block a community just click block (right next to subscribe). As far as blocking whole instances goes: As a user, not yet, but lemmynsfw.com has a strict “tag everything nsfw as nsfw” policy and I think lemmy.ee filters everything tagged such by default. That is, you can’t enable “show nsfw” in the user settings. They don’t want to get defederated either so they’re keeping their ship tidy.

two_wheel2@lemm.ee · 1 year ago

Looks great! Thanks for doing this! I don’t see anywhere here the approximate monthly costs… only what the money is being spent on. Do you have a figure for how much goes into running this instance?

sunaurus@lemm.ee · 1 year ago

The current projected bill for our whole infra in the month of June is $147. This covers the load balancer, 3 servers + database server, object storage for image uploads and our e-mail service. This may increase a little bit if we go higher than expected on bandwidth, object storage or outgoing e-mails.

two_wheel2@lemm.ee · edit-2 1 year ago

Alright, I’m tossing my tiny hat into the sponsor ring. Thanks so much for putting this community together! I’m excited to see it grow. Just out of curiosity, what does the incremental cost look like? Does it scale well with users? Or does it explode a little bit?

sunaurus@lemm.ee · edit-2 1 year ago

That’s greatly appreciated!

In terms of costs of scaling, I would say we’re positioned a bit better than many other Lemmy instances at the moment, thanks to the fact that we employ horizontal scaling as much as possible for the Lemmy software itself.

By the way, AFAIK, lemm.ee is the only non-experimental Lemmy instance that has chosen to go with horizontal scaling so far. If anybody knows of any other instance that is doing it, I would be super interested to know about it! All the admins I’ve spoken to so far myself have confirmed that they are only doing vertical scaling.

More technical details below for anybody who is interested:

There are two approaches you can generally take for scaling - horizontal, where you add more load balanced nodes of more or less the same power, or vertical, where you increase the power of an individual node (of course a mixture of both is also possible).

One of the benefits of horizontal scaling is that in most cases, it’s significantly more flexible compared to vertical scaling. For example, at my current cloud provider, the only upgrade path for vertical scaling a server would be 8 CPU -> 16 CPU - 32 CPU -> 40 CPU. So if you’re on a 16 CPU server, and you need just a little bit more headroom, then your only option is to upgrade to the 32CPU server, which is straight up double the power (and cost!). Meanwhile, with horizontal scaling, you can just keep adding smaller servers (say 2 CPU each) one at a time, thus growing costs more gradually and appropriately for your actual needs.

So for lemm.ee, this horizontal scaling means that when our backend servers start getting overloaded, I can just add one or two more servers without exponentially increasing costs.

electromage@lemm.ee · 1 year ago

Are your servers in one geographic region? Could you scale across regions for better performance?

sunaurus@lemm.ee · 1 year ago

I am already leveraging Cloudflare’s globally distributed cache, which helps improve performance even if you’re far away from the backend server. But this only helps partially, not with all types of requests.

lemm.ee is hosted in central Europe, and based on monitoring, it does seem that most users are having a pretty decent experience on lemm.ee regardless of their geographic location so far. One key exception to this are short windows of database load spikes, which last for roughly 10 seconds every 5 minutes. For these spikes, everybody is suffering equally, regardless of where they are in the world 😅.

But in general I agree with the sibling comment by @Notorious - rather than scaling one instance to be some massive globally distributed powerhouse, it makes sense to spread out the load amongst a lot of different instances.

electromage@lemm.ee · 1 year ago

Thank you for your work and communication! I agree it doesn’t make sense to invest in global infrastructure unless everyone does it, and the return wouldn’t be worth it. We’ll just have to get used to some performance issues as the fediverse takes off!

OneDimensionPrinter@lemm.ee · 1 year ago

Are the DB spikes ACTUALLY every 5 minutes or is that just kind of a guess? I ask because if it’s consistent, it’s gotta be some sidecar process somewhere in the stack that can be fiddled with.

That said, it really sounds like you know what you’re doing already so I’ll just go play with my new communities.

sunaurus@lemm.ee · 1 year ago

The spikes are caused by a specific reoccurring process which happens every 5 minutes. I have already significantly optimized it with a patch on lemm.ee, I’m working on getting it merged upstream as well!

Notorious@lemm.ee · 1 year ago

Personal opinion is that is outside the scope for a single instance. The whole idea behind Lemmy is to have multiple instances to accommodate different geos and different languages.

electromage@lemm.ee · 1 year ago

I think this could be problematic if instances aren’t providing a consistent user experience in different regions. If my Flashlight community is on an instance in California, and my Linux community is in Finland, I’m going to have a very asymmetric experience.

sunaurus@lemm.ee · 1 year ago

Home instances act as mirrors for posts and comments, so the experience should still be quite symmetric for you overall if you’re browsing both communities from the same instance

OneDimensionPrinter@lemm.ee · 1 year ago

As someone who has “been there and done that” at a much larger scale than many devs may ever get a chance to (not a brag, it can suck royally) this really seems like the smart choice.

This is effectively a basic web server scenario and horizonal scaling tends to with really well to a point. And frankly it’ll be a long while before that becomes the bottleneck.

Smart choices you’re making. All the best and I’m happy to help out monetarily where I can!

xavier666@lemm.ee · 1 year ago

For storage, I can understand how horizontal scaling works (add more storage nodes to, say glusterfs). But how does it work for CPU? Since adding a 2CPU VM can be physically on another server, it would need lemmy to work in a highly distributed manner, i.e., CPU instructions need to cross the network.

Is this distributed feature a part of lemmy or is there another abstraction layer?

sunaurus@lemm.ee · edit-2 1 year ago

This is where our load balancer comes in. All requests go through the load balancer, and this load balancer will try to evenly distribute the requests to all of our backend servers.

Is this distributed feature a part of lemmy … ?

In fact it’s the opposite - Lemmy has so far had some assumptions built in to the code which make it quite hard to run on multiple servers. I have made some modifications in order to improve this (and contributed those modifications back to the main repo as well). It’s one of the things I want to keep improving as we grow.

xavier666@lemm.ee · edit-2 1 year ago

Here is my oversimplified understanding of the backend of lemm.ee This

Am I correct? Or is there another loadbalancer in front of the DB?

Sorry for asking so many questions, but I’m new to system design and trying to learn about practical deployments.

sunaurus@lemm.ee · edit-2 1 year ago

That’s pretty close, but there are some nuances.

One of the servers is currently exclusively dedicated to handling images (processing, indexing, resizing, uploading to object storage)
One of the servers is only handling Lemmy HTTP requests
One of the servers is handling Lemmy HTTP requests + at the same time also handling Lemmy background tasks (different cleanups, updating the front page rankings, etc)

Additionally, we are not using Docker at all for lemm.ee. Not that I have anything against Docker - I use it regularly in other projects - it just wouldn’t provide any advantages for lemm.ee at the moment.

bric@lemm.ee · 1 year ago

Same. I’m not putting in a ton, but monthly donations go a long way to help with monthly server costs. We just need 150 people to put in $1 a month and we’ll be covered indefinitely

two_wheel2@lemm.ee · 1 year ago

Exactly. I’ve tossed in $5/mo and I literally just realized that with Reddit never in my WILDEST DREAMS would I have imagined kicking in some money for something like gold or trophies or even Apollo (RIP), but $5 a month to contribute to supporting a distributed community of people beyond myself feels like nothing to me. I think that speaks to the potential federation + good will can offer the world

OneDimensionPrinter@lemm.ee · 1 year ago

Shit. That’s a bunch of hardware/services. I hope the donations keep coming in. I’ll gladly drop a few bucks a month for quality updates and a relatively stable instance.

Thank you for running this so I don’t have to deal with it myself.

dan@upvote.au · 1 year ago

If they’re all powerful servers, $147 is pretty good for that many of them! Out of curiosity, are you using Hetzner? VPSes or physical servers?

sunaurus@lemm.ee · 1 year ago

Not Hetzner. We’re on VPSes for now!

dan@upvote.au · 1 year ago

Oh cool. That makes sense. Which provider?

AndromedusGalacticus@lemm.ee · 1 year ago

Yeah, I feel being transparent of cost will bring a lot more goodwill and help people want to participate in the goals of the server.

dkd4brw@lemm.ee · 1 year ago

Out of curiosity, can you explain more about what the infrastructure setup is? ie. what cloud provider, what services (managed databases or raw VMs) etc

sunaurus@lemm.ee · 1 year ago

We’re running on a managed database for now, which is great because it gives us things like high availability and point in time recovery out of the box, but the trade-off is that it’s also is a bit harder to tune for our needs. I think it’s a good choice for the time being, but let’s see what the future brings.

Sorry for not being more specific though, I’m a bit hesitant about going into too many details. I am afraid that putting all the details out there will allow malicious people to easily figure out what the potential bottlenecks are in our infrastructure, thus giving them a specific target to try and overwhelm.

There are for sure folks out there who want to see Lemmy fail, but maybe I’m being a bit too paranoid about this 😃 still, for now, I would rather be safe than sorry. Maybe in the future I will feel more comfortable with sharing a lot more.

Gatsby@lemm.ee · 1 year ago

Just because you’re paranoid doesn’t mean people aren’t really trying to kill you😉

dkd4brw@lemm.ee · 1 year ago

I understand, no worries! Thanks for sharing :)

Tankton@lemm.ee · 1 year ago

I agree about the pinned posts being a little too much, but overall its fine… :)

eyy@lemm.ee · 1 year ago

thank you so much!

Upcoming 0.18 upgrade, 404 errors and infrastructure costs

Upcoming 0.18 upgrade, 404 errors and infrastructure costs

Hello, fellow lemmings!

Upcoming 0.18 upgrade

Why do we even want 0.18?

Random 404 errors

Server costs

Pinning updates on the front page