Saving another 100TB of RAM
103 points by f311a 3 hours ago | 19 comments

ricardobeat 41 minutes ago
These optimizations are impressive, but it gets me thinking: at what point does a company become a collection of impenetrable siloes, where nothing really does what you expect? Maybe know with AI this is less of an issue as exploring a codebase is also much faster.
reply
BobbyTables2 3 minutes ago
I feel like any company whose products have RESTful interfaces are already there…

One wants to turn on an indicator on a remote device. A simple Boolean value. But we need networking, TLS, authentication plugins, certificate validation, distributed logging, containers, orchestration, HTTP client/server, interprocess communication, daemon dependency management, …

Sure, one can say each of these layers and abstractions has an important and justifiable purpose. But one can also step back and start wondering - what the hell are we really doing???

At some level, it seems like each layer of abstraction has to manage others, only simply because they exist.

Imagine the simplicity of 1800s telegraph signaling - no software!

Too often we build systems with Fortune-50 style hierarchies when a 5-person team could do the whole job.

reply
simonjgreen 7 minutes ago
My intuition around larger companies is they are already impenetrable silos, and AI makes it worse
reply
proc0 44 minutes ago
The only Rust section is the one on storage improvements about the struct that stores the hash, but do they really need that many hashes that 2 bytes makes that big of a difference? Article doesn't expand, but I guess it's a hash for every task on every computer, so maybe yes.
reply
agosta 41 minutes ago
That's exactly his point/the area of cost saving - that they didn't actually need as many hashes as they had started with. The trick was in finding out how many hashes they could cull without degrading load balance.
reply
agosta 45 minutes ago
Bang up article! As someone who doesn't get to do enough (almost any) calculus in my daily programming assignments, I thoroughly enjoyed reading about Kevin's dive into that derivation (linked in the supplemental article). All the people being negative here can swallow raisins
reply
terabyteoff 20 minutes ago
Thanks! Maybe dial it back or people are going to think I paid you
reply
johnnyApplePRNG 2 hours ago
[flagged]
reply
officialchicken 60 minutes ago
Come on CF, you can't even be bothered vibe a README for the crate? Somebody tear down this post until this garbage is documented properly.

I found this in the code, it looks like a rewrite-in-rust:

//! # pingora-ketama

//! A Rust port of the nginx consistent hashing algorithm.

//!

//! This crate provides a consistent hashing algorithm which is identical in

//! behavior to [nginx consistent hashing](https://www.nginx.com/resources/wiki/modules/consistent_hash...).

//!

//! Using a consistent hash strategy like this is useful when one wants to

//! minimize the amount of requests that need to be rehashed to different nodes

//! when a node is added or removed.

reply
globnomulous 5 minutes ago
If you have a real, actual, substantive critique of either the post or package itself, I'd be interested in reading that. What you posted doesn't provide that. I'm not sure who you're talking to or what you expect your comment to accomplish.
reply
cyberpunk 44 minutes ago
Anyone have an idea how it behaves differently from google's jump hash algorithm? The cool thing about google's one is it's so short I can include it in a HN comment:

    int32_t JumpConsistentHash(uint64_t key, int32_t num_buckets) {
      int64_t b = 1, j = 0;
      while (j < num_buckets) {
        b = j;
        key = key * 2862933555777941757ULL + 1;
        j = (b + 1) * (double(1LL << 31) / double((key >> 33) + 1));
      }
      return b;
    }
https://arxiv.org/pdf/1406.2294
reply
cyberpunk 31 minutes ago
Well I looked it up; nginx, apparently, uses ketama -- it's a ring-style hash probably works better for web backends than the jch above, as when given [0,1,2,3] and replacing the server in slot 1 you're going to have a lot of hash moves. With ketama, you'd only have the '1' hashes moving. You can't really beat google's for brevity, though.
reply
agosta 49 minutes ago
We can tell you didn't read the post because it is definitively NOT garbage. Very interesting write up by the Cloudflare team - the man literally did calculus to improve something. When's the last time any of us did Calculus to improve anything? Bang up job Kevin and everyone!!
reply
go_elmo 2 hours ago
Was this vibecoded? I hope so, would expect models to be good enough by now, its not that creative.
reply
n738 2 hours ago
go_elmo is unimpressed everyone. Pack it up. Time to go home.
reply
Maxion 2 hours ago
I think it's more important now than it was last year to differentiate pure yolo vibecoding from "AI assisted engineering" or "AI engineering", I.e. deliberate and careful use of AI to speed up coding but without creating too much slop. Maybe aineering?
reply
wild_pointer 2 hours ago
Nah, AI code is black or white, and whether the code is good depends on your religion.
reply
kawogi 2 hours ago
AIded development?
reply
derwiki 2 hours ago
Just like the original harness, aider.chat
reply
readthenotes1 2 hours ago
I expect a another post in a year where they get a performance improvement by reducing the number of calls to create integers from bytes.
reply