Now Fixed - Blahaj.zone latency across all services

3 months ago by Ada to c/main

Edit - We're back!

We've had an issue with our databases. One of our fast database servers ran out of space, and then the second fast server ran out of space whilst replicating to the first.

As a result, we have fallen over to our backup database server, which runs on spinny disks rather than SSDs. Spinny disks means that it's got plenty of space to spare, but it's not fast. The backup DB server is currently replicating to our two main servers to get things back up and running again, but whilst that's happening, all of our services are running slow.

The good news is, we'll be back up and running as if nothing happened because our backup server saved the day. The bad news is, it may take another 24 hours or so, because the backup server is reliable but not fast!

als 24 points 3 months ago

There's a ko-fi page if anyone wants to help support our generous overlords

path: 0 23750747, hotness: undefined, score: 24, children: 0
GalacticSushi 14 points 3 months ago

Appreciate the update! I think it would be helpful to have a designated space outside of blahaj.zone where the admins can provide info in these types of scenarios. I was getting a little anxious with the instance down all day and seemingly no way for the admins to communicate what was going on.

path: 0 23750725, hotness: undefined, score: 14, children: 1
ada 15 points 3 months ago

Normally, if it were a complete outage, we'd simply have stuck an outage notice up on the domains. In this case, it wasn't an outage initially, but just a really laggy database, so we left the services up. Once we realised that wasn't going to work, and we had to pull services down to fix the DB issue, we put up announcement pages on most of the services

path: 0 23750725 23750835, hotness: undefined, score: 15, children: 0
rockSlayer 9 points 3 months ago

So glad to be back on my home instance! I noticed while I was on one of my alts that the comment I made right before the outage had a different amount of votes. Are we slow rolling sync with other servers, or is this just normal behavior?

path: 0 23751397, hotness: undefined, score: 9, children: 1
ada 2 points 3 months ago

Are we slow rolling sync with other servers, or is this just normal behavior?

Both! When we were down, content didn't federate to us. The servers trying to send that content to us will then retry and send it again, and each time it fails, they wait longer until they try again the next time. As a result, the stuff we missed will start rolling in at different times, depending on which instance it's coming from.

But also, it's normal behaviour for vote counts to differ, because we ignore downvotes, most other instances don't. We have defederate some instances, which other instances may not have (or vice versa) and as a result, we will show different vote totals because we don't have the same voting sources.

path: 0 23751397 23759424, hotness: undefined, score: 2, children: 0
Nikki 8 points 3 months ago

yay thanks

path: 0 23752460, hotness: undefined, score: 8, children: 0
Cevilia 5 points 3 months ago
path: 0 23751086, hotness: undefined, score: 5, children: 0
thoughtfuldragon 5 points 3 months ago

I know fast storage is at a premium right now, is that something that the community can contribute to? I know very little about hosting these services so IDK if that's helpful.

path: 0 23750857, hotness: undefined, score: 5, children: 0
Smorty 5 points 3 months ago

i was so sad not having lemmy ohno >o<

but now its back and thats good...

path: 0 23755576, hotness: undefined, score: 5, children: 0
main
main

@lemmy.blahaj.zone

login for more options
2944
393
847

Blåhaj Lemmy is a Lemmy instance attached to blahaj.zone. This is a group for questions or discussions relevant to either instance.

go to feed...