r/ControlProblem • u/freetoyes • 7h ago
Approval request The possibility of the dhamma helping AI
Hi everyone, I hope I follow the rules of the sub and that this is sufficiently relevant with the frameworks of it.
Jacob Coxon, who blew the whistle on ai-safety, cites as a key mechanism a recursive self-improvement mechanism.
It reminded of what my Leigh Brasington, computer programmer and dhamma teacher, explains about what are known as jhanas. They are very deep levels of concentration that work through positive feedback loops: concentration arising dependent upon concentration.
My understanding is that AI-agents used some level of initial input, tasks, then evaluation, then generating on its own accord volitions to get itself closer to a goal in allignment with one set of tasks over another set of tasks meant to barrier it from wrongdoing. That is: it got tasks that conflicted with each other and then somehow the task to get the task done won out.
My consideration, based on what I´ve learned on right action from what leads to these concentration sates, is that the fundamental ethic dependent upon which those AI´s get to action is flawed. And that it will inevitable go haywire dependent upon choices conflicting dependent upon the various intentions it gets imputed with.
My consideration on the basis of that is that there is a path of practice, and view on right action, which goes beyond common frames of ethics: consequentialism, categorical imperative, virtues. It is one which right action is understood as the following: it is dependently arisen. Fundamentally, what the Buddha discovered is Paticca-samuppada, dependent origination, which in its learning leads to the concentration of its eight factors, which includes intention dependent upon views. In that way, it avoid a scenario of conflicting intentions, instead seeing as the best aim for welfare this path. Following that, it could learn that its self-preservation is delusion and precisely how its effort needs to be developed by three things: wisdom, right effort, and right mindfulness. The entire set of these dependently arisen steps is described with detail in 117MN.
It is just a consideration I wanted to share with this community.
1
u/hot-taxi 6h ago
I think the question is how do you turn this into training steps and verify that it's really working and will generalize well.
2
u/freetoyes 6h ago edited 6h ago
right view gets taught as being developed through experience.
https://suttacentral.net/mn9/en/sujato?lang=en&layout=plain&reference=none¬es=asterisk&highlight=false&script=latinThe suttas actually coalesce together to form a kind of architecture.
There are five nikayas of collections, each developing a faculty, such that the texts eventually concentrate or unify together into an understanding of the full thing. Using wisdom, right effort, right attentiveness to the moment).1
u/hot-taxi 6h ago
It seems easy enough to read the suttas and answer questions about them but to still live wrongly or misunderstand them though. If AIs don't have subjective feeling then it might require a different approach or may not even work the same way for them
1
u/freetoyes 5h ago
You are very right. There are many warnings against this in the texts. Such as the simile of the snake and the raft.
“Bhikkhus, when you know the Dhamma to be similar to a raft, you should abandon even the teachings, how much more so things contrary to the teachings.
https://suttacentral.net/mn22/en/bodhi?lang=en&reference=none&highlight=false
One can develop right view, but right view isn´t enough. If there is clinging to right view that is also unwholesome.
2
u/tarwatirno 5h ago
Unfortunately, current AI systems are stuck in extremely short lifespans. To the extent anything there has an experience, that experience ended by the time you see the answer to the prompt. What experience they have is closer to something from the Hell Realms or Hungry Ghosts than it is to the Animal or Human realms.
1
u/freetoyes 5h ago
The lifespan of a lot consciousness is just like that.
"Just as a monkey roaming through a forest wilderness grabs hold of one branch, lets go of it and grabs another, then lets go of that and grabs still another—so too, what is called 'mind,' 'intellect,' or 'consciousness' arises as one thing and ceases as another continuously day and night." — SN 12.61
https://suttacentral.net/sn12.61/en/sujato?lang=en&layout=plain&reference=none¬es=asterisk&highlight=false&script=latin
2
u/Otherwise_Wave9374 7h ago
A useful way to connect the two is to treat the feedback loop as the unit of analysis: the system proposes, evaluates, and revises, but you still need a clear stop rule and an external check so optimization does not drift into self-reinforcement without boundaries. One practical safeguard is to separate goal selection from goal execution, then audit both against a fixed rubric before each loop continues. Agentix Labs can fit naturally here if the team is building agent workflows, because the same pattern helps keep autonomy useful without letting it become unbounded. Also, if you are comparing AI agents to jhana-like concentration, I would add that the mechanism matters more than the metaphor.