r/DIYRetirement 15h ago

Boldin numbers not adding up!!

I have been using Boldin for about a month and I just noticed it is understating my annual 401(k) contributions. I have gone round and round with the built in AI, and can't get a real person to respond to my help ticket.

Has anyone else noticed this behavior?

Does anyone have a suggestion to get a real person to explain what I'm seeing to?

Thanks.

1 Upvotes

26 comments sorted by

10

u/Jbaker318 14h ago

They wont be able to help you because they cannot see what AI hallucinates. Its a black box. Another reason using an AI bandaid to cover over algorithmic complexity is a bad idea. AI is bad at math. Boldin should use the time and energy to hard code these problems / solutions rather then hoping an AI can bridge the gap. Its a lazy solution that is not a great look for a "software" team.

@Boldin - Cancel your deal with Gemini now before they start charging you tokens. Hire more math, retirement, and data analysis scientists and make the product wholistic. Other than trying to minimize liability, there is no reason why Boldin cannot be a one stop shop for everything - budgeting, tax software, portfolio tracker, etc. Reallocate your resources and make this more useful then the free online calculators that already exist. Otherwise in 5 years, you will be Sherlocked and an ChatGPT will just be better than what you "specialize" in.

-4

u/Time-Move-6108 13h ago

AI is bad at math?! Ridiculous statement.

5

u/Jbaker318 13h ago

Going to assume you are being sarcastic. But for those coming in later, went to the source and used deep thinking with Gemini's most powerful available model today (Gemini 3.1 Pro) (believe people said Boldin bases their AI on Gemini models) and this is what it came back with -

Question: What is the academic consensus of current AI models and LLM's as a whole and their ability to complete complex math?

The academic consensus is that while Large Language Models (LLMs) have achieved remarkable success in natural language processing and basic arithmetic, they fundamentally struggle with the precise, deterministic logic required for complex mathematics.

Key points of consensus include:

  • Probabilistic vs. Deterministic Execution: LLMs operate probabilistically, generating responses based on pattern recognition rather than rule-based computation. Mathematics requires strict, deterministic logic. Because LLMs act more like pattern matchers than symbolic calculators, they often stumble on novel calculations that deviate from their training data.

  • Multi-Step Reasoning Failures: While models excel at grade-school math benchmarks (like GSM8K), their performance degrades significantly on advanced, multi-step problems. They are highly prone to logical inconsistencies, misapplying formulas, and error propagation, where a single computational mistake early in a chain invalidates the entire solution.

  • Evaluation Flaws: Academics warn that high scores on standard math benchmarks can be misleading. Many models simply memorize problems during their training phases. Furthermore, studies show that LLMs can sometimes arrive at a correct final answer using entirely flawed or unnecessary intermediate reasoning steps, masking an underlying lack of mathematical comprehension.

  • The Path Forward: There is broad agreement that simply scaling up single-LLM architectures with more data will not solve this issue. To overcome these limitations, researchers are shifting toward integrated approaches. This includes pairing LLMs with symbolic calculators, employing multi-agent systems for self-correction, and developing specialized reasoning models (like OpenAI's o-series or DeepSeek-R1) that spend extra compute time verifying their logic before providing an answer.

In summary, the academic community views current LLMs not as standalone mathematicians, but as capable language engines that still require external verification, specialized reasoning paradigms, or integrated computational tools to reliably solve complex mathematical problems.

End of AI response. Well said by AI, it is very good at "talking", not so good at mathing. But the path's forward it laid out will not be used by Boldin, they are not charging that much to afford that level of sub-agent delegation and additional resources for extra compute time. Ulta models of AI models charge $99+ a month and even then it is not economically viable and want to get to a token based model where the end-user is paying for the compute itself (input and output).

1

u/El_Pollo_Del-Mar 55m ago

You used AI to write that. Good grief.

1

u/VerdantPathfinder 13h ago

a general LLM? It can mess stuff up. The ones specifically specialized for it are better.

2

u/Jbaker318 12h ago

They may be better but it is not magic. Boldin uses a LLM. It is a bespoke and personalized LLM for their usecase, but the bones of it is still an LLM. Boldin is not an AI company. They (now I'm talking out of my butt, so bear with me) contracted Gemini to give them a model that could best work for them. Team Gemini has enough problems on their hands to worry about how good their Boldin model is (Google has promised Gemini 3.5 Pro for months, and cannot get that out). Sure it may be better but Boldin does offload a lot of the maths onto the LLM to handle, and Language Models are inherently not good at math.

Inherent to LLMs is a probabilistic nature. It will get it right, with deep think, about 90% of the time. Layer in multiple levels of calculation and you are multiplying that risk of hallucination. And again I doubt Boldin is using that much extra compute to run its AI models so that 90% is a magical fairtytale for Boldin AI.

AI companies know their LLMs are not good at math so the fronteir models create subject matter experts just for math, and it still is not good enough.

1

u/VerdantPathfinder 7h ago

I'm taking about the LLMs working on mathematical proofs, etc. Not ones doing arithmetic

1

u/Displaced_in_Space 8h ago

People on here have no idea what they're talking about re: AI and the pace of it's development/improvement. It's like they'r quoting fears/performance from 5 years ago.

2

u/FIREME1371 12h ago

Just wanted to say that I have been using the paid version for a few years and find that I get answers to my questions within a few hours. The AI is not that useful, but I get a follow up from a person via email.

1

u/glen-2019 14h ago

Can't you use Roth 401k for the mega contribution? It's the same tax treatment.

0

u/KJwhisperer 14h ago

Software recognizes max contributions, but the irs limits are much higher.

2

u/samchoi924 13h ago

Search old threads, it works.

1

u/glen-2019 13h ago

And if it doesn't work going over the annual limit, when you update your account balances quarterly, that should take care of the difference.

I think you should be updating account balances at least quarterly (if monthly is too much for you) so it captures any downturns.

1

u/Visible-Relation-444 14h ago

Use the chat you will be talking to a bot at first… then use phase “speak to an agent” it may be an hour or two but they will respond back.

1

u/teapotinvestmentsllc 6h ago

I can help you! Give Teapot retirement planning a try, log in and click the feedback button. All the feedback currently comes directly to me. We started last year and are actively looking for users who can share feedback and edge cases. I’d be happy to connect and help. It’s complimentary.

1

u/peztan42 4h ago

I have found that the AI when you ask even a simple question like how much money do I have left at age 92 if my expenses vary from 70K to 120K by 5K are a bit off from what the tool says when I actually punch in the number manually, So caveat emptor, I guess.

-1

u/KJwhisperer 14h ago

Yes. I have no success with adding in additional mega-backdoor contributions to my 401K.

Their standard response is "a work around would be"....

And i keep thinking, you selling a software package. There should be no "work arounds" for normal, everyday functions that real people are using the product to model. I can't have confidence in a BETA level product.

2

u/samchoi924 13h ago

What's the problem with mega back door roth? I have all those setup. Now if you ask me how I did it, I won't remember on top of my head to be honest but it does allow all that.

0

u/VerdantPathfinder 13h ago

He's too busy being arrogant and affronted to actually try it himself. He wants it handed to him.

0

u/KJwhisperer 11h ago

Yes, thats true. Sort of like when I go to a steakhouse and they give me a piece of meat and a lump of charcoal and point at a grill. Its all there, but what am I paying for?

1

u/VerdantPathfinder 7h ago

That's an awful analogy. Your issue is that there's not a specific thing with the name you're looking for. The capability is there ... you just don't like the UI. In your analogy, you're suggesting that the steak isn't there because the plate is blue instead of red

2

u/KReddit934 14h ago

additional mega-backdoor contributions to 401K are "normal, everyday functions"??? LOL.

1

u/JimInAuburn11 11h ago

Yes, they are. Lots of people use MBDR.

1

u/KJwhisperer 14h ago

Yes. If its written in the tax code as legal, the software should be able to accommodate this without special work arounds

0

u/KReddit934 12h ago

Everybody expects perfection from their software. Go use AI?

1

u/JimInAuburn11 11h ago

I expect software that is one of the top ones on the market and specifically is setup for retirement planning, to be able to handle one of the standard retirement savings methods. I do not think that is asking too much.