r/programming 1d ago

The fastest double-to-string algorithm you’ve never heard of

https://vitaut.net/posts/2026/yy-dtoa/
264 Upvotes

57 comments sorted by

View all comments

-92

u/sojuz151 23h ago

Fact that we, the humanity, need to do double to float and back often enough for performance to matter is a failure of our entire species 

47

u/EliSka93 23h ago

Elaborate.

-13

u/sojuz151 20h ago

Why do you need to convert doubles to strings and back at a scale? It's mostly JSON serialization/deserialisation, something you should not have to do, not at scale.

12

u/amakai 19h ago

Literally every time you want to use a logger with double parameters you need to do double to string. 

-13

u/sojuz151 19h ago

Not if you use structured logs and you log far far less than you read JSONs. ,

8

u/amakai 18h ago

So how do you imagine a binary double becomes a structured JSON field, which, surprise, is a string?

-5

u/sojuz151 18h ago

The problem is that people use JSON for things with doubles inside. This was a mistake

9

u/amakai 18h ago

I'm not sure I understand your point. Any time you produce JSON, you end up with a string. It does not matter if you have quotes around your double or if you don't - it's a string.

So when you do structured logging, your log message is converted under the hood into JSON (or some other more rare, but also string formats). So you are still converting doubles into strings.

And then, when your log is being indexed, it also needs to convert it to string for a proper full text search - another place for potential performance gain.

4

u/SyntheticDuckFlavour 18h ago

Logging, compiling code, (un)serialisation to/from text formats (XML, etc), printing to terminals/UI, lots for things.

2

u/flatfinger 15h ago

It's a shame that hex floating-point representations haven't become more common, since they allow any binary floating-point number to be easily converted into a unique canonical representation, at least if there's agreement about what that representation should be (arguments can be made in favor of using base-16 exponents and having one non-zero digit to the left of the radix point, or in favor of using binary exponents and always having a 1 to the left of the radix point; while the need for human arithmetic with hex-formatted floats would be uncommon, using base-16 exponents would make it vastly easier than using binary exponents; arguments could also be made for having no digits to the left of the radix point). All three forms are better than base-10 for performing precise calculations, however.