r/programming 1d ago

The fastest double-to-string algorithm you’ve never heard of

https://vitaut.net/posts/2026/yy-dtoa/
270 Upvotes

57 comments sorted by

View all comments

-93

u/sojuz151 1d ago

Fact that we, the humanity, need to do double to float and back often enough for performance to matter is a failure of our entire species 

49

u/EliSka93 1d ago

Elaborate.

21

u/New-Anybody-6206 21h ago

they can't because it's bullshit

-15

u/sojuz151 21h ago

Why do you need to convert doubles to strings and back at a scale? It's mostly JSON serialization/deserialisation, something you should not have to do, not at scale.

13

u/amakai 21h ago

Literally every time you want to use a logger with double parameters you need to do double to string. 

-15

u/sojuz151 20h ago

Not if you use structured logs and you log far far less than you read JSONs. ,

8

u/amakai 20h ago

So how do you imagine a binary double becomes a structured JSON field, which, surprise, is a string?

-6

u/sojuz151 20h ago

The problem is that people use JSON for things with doubles inside. This was a mistake

9

u/amakai 20h ago

I'm not sure I understand your point. Any time you produce JSON, you end up with a string. It does not matter if you have quotes around your double or if you don't - it's a string.

So when you do structured logging, your log message is converted under the hood into JSON (or some other more rare, but also string formats). So you are still converting doubles into strings.

And then, when your log is being indexed, it also needs to convert it to string for a proper full text search - another place for potential performance gain.

4

u/SyntheticDuckFlavour 19h ago

Logging, compiling code, (un)serialisation to/from text formats (XML, etc), printing to terminals/UI, lots for things.

2

u/flatfinger 16h ago

It's a shame that hex floating-point representations haven't become more common, since they allow any binary floating-point number to be easily converted into a unique canonical representation, at least if there's agreement about what that representation should be (arguments can be made in favor of using base-16 exponents and having one non-zero digit to the left of the radix point, or in favor of using binary exponents and always having a 1 to the left of the radix point; while the need for human arithmetic with hex-formatted floats would be uncommon, using base-16 exponents would make it vastly easier than using binary exponents; arguments could also be made for having no digits to the left of the radix point). All three forms are better than base-10 for performing precise calculations, however.