HN user

antongribok

1,355 karma

I'm just a Linux Sysadmin with 25 years of experience, the last 17 years working on distributed storage.

Posts26
Comments260
View on HN
www.youtube.com 5mo ago

NTSB Animation of Flight 5342 2025 Potomac River mid-air collision [video]

antongribok
3pts0
www.dexerto.com 9mo ago

Armed police swarm student after AI mistakes bag of Doritos for a weapon

antongribok
693pts436
www.cnn.com 2y ago

Why scientists say we need to send clocks to the moon

antongribok
35pts29
www.bbc.com 2y ago

Electric buses withdrawn in south London after fire

antongribok
46pts40
storagesearch.com 3y ago

SSD Market History (1970s to 2018)

antongribok
2pts0
www.bbc.com 4y ago

Russian film team boldly shoot towards space station

antongribok
1pts0
www.redhat.com 4y ago

Looking back on 30 years of Linux history with Red Hat's Richard Jones

antongribok
4pts0
www.artstation.com 5y ago

Not a Portrait

antongribok
1pts0
duckduckgo.com 6y ago

1 Gallon = 3.999987 Quarts

antongribok
4pts3
www.washingtonpost.com 6y ago

DEA seized father’s life savings at airport without alleging any crime occurred

antongribok
60pts23
www.bbc.com 6y ago

A billionaire retailer whose shops had no stock

antongribok
105pts50
www.thedrive.com 6y ago

Tragic Tale How NASA's X-34 Space Planes Ended Up Rotting in Someone's Backyard

antongribok
2pts0
www.washingtonpost.com 6y ago

Call of Duty Modern Warfare aims to be most realistic war game on the market

antongribok
3pts3
www.bbc.com 7y ago

Call of Duty Hoax Caller Tyler Barriss Jailed

antongribok
2pts0
www.washingtonpost.com 7y ago

Yext to hire 500 for work in Arlington, VA

antongribok
1pts0
www.washingtonpost.com 7y ago

SCOTUS – Constitutional Protection Against Excessive Fines Applies to States

antongribok
73pts2
www.kwkpromes.pl 7y ago

Safe House

antongribok
2pts0
www.theregister.co.uk 8y ago

Who had Intel in the 'discrimination lawsuit' pool? Congratulations

antongribok
1pts1
www.washingtonpost.com 8y ago

Nationwide, police shot and killed nearly 1,000 people in 2017

antongribok
3pts1
daniel.haxx.se 9y ago

Denied Entry

antongribok
284pts70
www.washingtonpost.com 9y ago

A Union Station ad screen played PornHub videos Monday night

antongribok
3pts0
www.bbc.com 9y ago

Miami's fight against rising seas

antongribok
2pts0
www.motortrend.com 9y ago

Elon Musk Is the 2017 Motor Trend Person of the Year

antongribok
2pts1
www.bbc.com 10y ago

The bicycle is making a comeback in US cities

antongribok
89pts104
en.wikipedia.org 10y ago

Airship Italia

antongribok
1pts0
www.washingtonpost.com 10y ago

What it’s like to be a hot girl online, when you’re a nerdy guy in real life

antongribok
1pts0

I think more people should know about the existence of ZRAM on modern Linux distributions. It's really changed the way I look at swap configs.

ZRAM is a compressed block device that is stored in RAM. It's great!

Previously, if I ever had high memory pressure situations, I really dreaded the slowdowns. Now, with swap sitting on top of /dev/zram0 it's a completely different experience.

I have ZRAM enabled on all of my personal machines, both laptops with limited memory, and desktops with 64 or 128GB of RAM. It's rarely used, but it is nice to have that extra room sometimes.

The performance of a zram device is so much faster than even the latest NVMe drives.

Not sure what you mean about Ceph wanting to be in a single rack.

I run Ceph at work. We have some clusters spanning 20 racks in a network fabric that has over 100 racks.

In a typical Leaf-Spine network architecture, you can easily have sub 100 microsecond network latency which would translate to sub millisecond Ceph latencies.

We have one site that is Leaf-Spine-SuperSpine, and the difference in network latency is barely measurable between machines in the same network pod and between different network pods.

Monash University is also a Ceph Foundation member.

They've been active in the Ceph community for a long time.

I don't know any specifics, but I'm pretty sure their Ceph installation is pretty big and used to support critical data.

I've lived most of my adult life in houses with forced air furnaces (albeit powered via natural gas, not propane), and what you are saying is inaccurate regarding indoor air pollution unless your furnace is in need of immediate replacement.

A modern furnace works via a heat exchanger, where the combustion produced pollutants never mix with the indoor air being pushed through. All pollutants are expelled outside via a property functioning chimney. This is one reason why you should have the furnace (and chimney function) inspected annually. Aging heat exchangers will show hotspots before there is a possibility of air being mixed, giving plenty of time to plan for a replacement. Of course there is a possibility of failure, which is why you should have a carbon monoxide detector.

I know I'm going to sound crazy here, but there is one more alternative. How about: Reduce, Reuse, Repair, Recycle?

I recently got a sewing machine for an unrelated project and around the same time I ordered it I had one of these cloth reusable bags rip, because I put too many heavy things in it. When I got the sewing machine, for practice I decided to see if I could fix the bag. It turned out to be surprisingly quick and easy. I didn't use any extra material besides the thread, and I believe the bag is much stronger now.

While most of what you speak of re Ceph is correct, I want to strongly disagree with your view of not filling up Ceph above 66%. It really depends on implementation details. If you have 10 nodes, yeah then maybe that's a good rule of thumb. But if you're running 100 or 1000 nodes, there's no reason to waste so much raw capacity.

With upmap and balancer it is very easy to run a Ceph cluster where every single node/disk is within 1-1.5% of the average raw utilization of the cluster. Yes, you need room for failures, but on a large cluster it doesn't require much.

80% is definitely achievable, 85% should be as well on larger clusters.

Also re scale, depending on how small we're talking of course, but I'd rather have a small Ceph cluster with 5-10 tiny nodes than a single Linux server with LVM if I care about uptime. It makes scheduled maintenances much easier, also a disk failure on a regular server means RAID group (or ZFS/btrfs?) rebuild. With Ceph, even at fairly modest scale you can have very fast recovery times.

Source, I've been running production workloads on Ceph at fortune-50 companies for more than a decade, and yes I'm biased towards Ceph.

Thank you for this detailed reply!

I only want to add a small suggestion. I get that large distributed production systems will occasionally go down, but it would be great if you could look into reducing the latency of your status page.

By my count there was at least a 35 minute delay between when things broke and before the status page (https://fastmailstatus.com) was updated.

Also, I think it would have been nice to have a bit more explanation on this event than simply "database issues" [1]. Being able to know that this was related to an upgrade would have made me feel a bit better during the time the status page was updated and until the issue was resolved.

Thank you for your hard work and an excellent email service!

-A long time customer.

[1] https://fastmailstatus.com/cme1fq7ej002dh0iu6z8pey4f

Having the same problems.

It started out as not being able to search, but the situation is quickly deteriorating and now I'm unable to open pretty much any email message.

Some content seems to briefly show up and then it quickly disappears and after that, it's as if cache has been invalidated and you can't get back into it.

Distro: Fedora (default Gnome)

Laptop: LG Gram 16"

There are several variations of the laptop with spec differences, the main thing you want is Intel CPU with integrated graphics.

Great screen (16:10 ratio), great battery life (80Wh), dual NVMe slots (if you care about bit rot).

Last, but not least, very, very light.

This is complete nonsense. No one running business critical installs of Ceph runs single-monitor.

You can also tell Ceph to use a single disk as your failure domain. No one does that either. Homelabbers maybe, but then why are you comparing such setups with Google?

We run Ceph with a failure domain of an entire rack. We can literally take down (scheduled or unscheduled) an entire rack of 40 servers, and continue to serve critical, latency sensitive applications, with no noticeable performance loss.

We have a Ceph footprint 5x larger than CERN run by a team of 4-5 people.

I had to get a new setup spun up very quickly, didn't know what to get, so I went out and got a bunch of cameras ranging from $49 to $500 from MicroCenter. For my requirements the cheapest cameras ended up serving my needs the best.

I ended up returning all of the expensive stuff and only keeping a bunch of really cheap Amcrest PoE cameras.

Here is one such example: https://www.microcenter.com/product/634071/amcrest-5mp-ultra...

I've been very impressed with Amcrest cameras.

They support being configured without an outside internet connection.

They support dual streams.

They have all kinds of tuning settings, and come with sane defaults.

They support H.265 encoding, so you get good quality at small file (and bandwidth) sizes.

They also have MicroSD slots, and some cloud stuff that I don't use.

They work great with Fridate and Google Coral with is what I use them with.

Highly, highly satisfied and recommend.

Ceph actually works very well with swap on zram. I have some older Ceph clusters that have 48GB zram devices used for swap.

It sometimes helps with defragmenting memory, memory pressure in general, and depending on your workload and other things running on the box, you could have better performance by being able to have more things in page cache.

CephFS works great at home. I've got an all-Linux setup, and it works great on Laptops as they go to sleep and resume. Works great over Tailscale on a modile hotspot too.

My home cluster is running on some 7 Raspberry Pis currently, so it's not very performant, but the uptime over the past 2-3 years has been unbeatable.

With cephadm it's very easy to stand up, and upkeep has been basically zero.

You're mixing software defined storage options with public cloud offerings, which is kind of strange.

If you have your own hardware, Ceph is solid.

I wouldn't put more than 200 million objects in a single bucket, but other than that it can be very reliable.

I have ~ half an exabyte on-prem, and sleep very well at night.

Not sure if I agree with you on RAID 0, but RAID 1 certainly is, and that's indeed what I'm running.

The 3x20TB btrfs setup I mentioned is configured with RAID1C3 for metadata, and RAID1 for data, and works just fine with even or odd number of drives.

It's funny how people assumed RAID5 when they saw 3 drives.

I switched to this after years of running on ZFS, and for my workloads btrfs is faster on Linux (not to mention the licensing/packaging mess).

I actually recently switched from my laptop to my desktop as my main work machine, and due to some weird partition choices previously (long story), I temporarily ended up with my /home/<work_user> directory on a btrfs filesystem that's sitting on top of 3 Seagate Exos 20TB drives (instead of my main NVMe).

Hearing the drives has been really nice actually, and got me noticing all kinds of interesting and sometimes unexpected behavior going on with my system, and actually helped find a bug with my terminal multiplexer.

With 64GB of RAM my entire home directory fits, so only writes go to the drives, and it's been surprisingly performant for my workloads.

I've been using Xonsh as my main shell for about the same time. I love it as well.

However, I feel like your criticism re fzf is not really fair, because I run into this with other tools quite often. So often in fact that it took me only a second to convert your command in my head to this:

  git checkout @($(git branch |fzf).strip())

In case you ever have to work with a directory that has lots of files in it, sometimes it's useful to remember the -U flag in ls. For example to count the number of files, I sometimes use this:

  ls -U1 |wc -l
Can be very fast with lots of files on a modern Linux box.

This checks out with my anecdata...

My wife, a licensed teacher in our state, was switching jobs from one school district to another, and watching her go through that process was interesting...

The interview process for any kind of school administrative position consisted of multiple rounds of interviews.

The interview process for a teaching position consisted of them confirming that my wife had a pulse.