logoalt Hacker News

psanford • today at 4:20 PM • 7 replies • view on HN

Object store is quickly becoming the new core data substrate. Lets build kafka, but on s3. Lets build github, but on s3. It feels like we going to see more and more "object-store first" systems in the next few years.

I am excited about this future. Give me stateless servers and a storage bucket over having to manage systems with disks any day.

I do wonder if we will see an expansion of the s3 api to support more of these use cases. S3 added a janky file append operation to their new express-one-zone bucket type, and limited to 10k total file append operations. I wonder what else we will get in the next few years.


Replies

shye • today at 7:05 PM

It makes perfect sense: it's a extremely reliable, infinitely-scalable, strongly consistent (in some implementation), cheap, and extremely simple to use key-value store. As such, it transparently solves a lot of the problems distributed systems have to be engineered around.

As long as you can run on a CSP, and can engineer around the high-ish latency (most business cases can), it's extremely expensive to try engineer around it.

necubi • today at 4:31 PM

Definitely agree that every data system that doesn't need <100ms latency is moving to object storage.

> I do wonder if we will see an expansion of the s3 api to support more of these use cases

This is actually an area where I think we have a big leg up on folks building on top of S3. My team (which built K2) sits next to the R2 team, and we have the opportunity to co-evolve the products in mutually beneficial ways.

➕ show 1 reply
khazit • today at 6:03 PM

I still think S3 is underutilized.

The range of things you can do with blob storage and a (very simple) auth model are surprisingly broad.

We recently replaced our Docker container registry with S3 using a tiny tool [1] we built in-house. I think that even with current capabilities, we can still model a lot services as a very thin layer over object storage.

[1]: https://github.com/Simple-Observability/grue

➕ show 3 replies
someonebaggy • today at 8:16 PM

How will you prevent this system from becoming a server with extra steps?

madjam002 • today at 7:27 PM

Came across this the other day which lets you run a etcd compatible API/Kubernetes on top of S3

https://github.com/t4db/t4

Haven't tried it yet but looks nice for simpler K8s deployments

6thbit • today at 4:59 PM

doesn't that make egress fees egregious? or still cheaper than disks?

➕ show 3 replies
Onavo • today at 4:28 PM

People like S3 because they have hard engineering guarantees around bit rot and work well as a high level abstraction of a network filesystem with all of the low level failure recovery built-in. You don't have to worry about doing your own RAID configs. The bigger question is whether non-AWS services can offer the same level of guarantees. I have heard horror stories for example when it comes to downtime on Hetzner's S3 object store.

I am currently using Cloudflare R2 right now and if you see their forums, there's always the occasional post about objects going missing.

➕ show 2 replies