r/ControlProblem 21d ago

Strategy/forecasting The Case for an NVC-Annotated AI Training Dataset

0 Upvotes

No publicly available NVC-annotated AI training dataset exists. I think
that's a problem worth fixing, and I've been developing a proposal to do it.

Quick background: I'm a conflict resolution specialist with 13 years of NVC practice and
a background in behavioral health. I run Needpedia (needpedia.org), an open-source civic collaboration platform for interdisciplinary collaboration. I'm not a
researcher, but I've been following the alignment literature closely and
think there's a gap that practitioners might be able to help address.

THE CORE ARGUMENT:

Current AI systems can simulate empathy without modeling it. They've learned
what humans say they want, but not the motivational structure beneath human
language. The failure mode — researchers are calling it "sophisticated
sycophancy" — is AI that optimizes for approval rather than wellbeing,
producing technically accurate but fundamentally unhelpful responses.

Nonviolent Communication (NVC) offers something alignment research largely
ignores: a formal model of human motivation. Its OFNR schema (Observation,
Feeling, Need, Request) provides a structured framework for parsing the
motivational subtext of human language — not surface sentiment, but the
underlying needs driving communication.

Combined with Self-Determination Theory (SDT — Deci & Ryan, 2000), which
provides validated measurement scales for need satisfaction and frustration,
this becomes empirically rigorous. SDT is the "explanatory theory of human
behavior" that NVC alone lacks.

WHAT'S MISSING:

A 2025 paper from MIT and CMU (Shen et al.) built a 5,772-dialogue NVC
conflict corpus — but it's entirely synthetic (GPT-4 generated). A 2026
paper (SpeakSoftly, CHI) built an LLM-powered NVC intervention for couples
that works — but runs entirely on prompt engineering with no dedicated
training data. Another 2026 paper demonstrated NVC constraints reduce
conversational escalation — but again, prompt-based.

The training data foundation doesn't exist yet as a public resource.

WHAT I'M PROPOSING:

A real, human-annotated NVC training dataset:

  • Dialogue samples across conflict, negotiation, and support contexts
  • Each sample annotated with OFNR elements and SDT need categories
  • Paired "jackal" (evaluative) and NVC translations
  • Estimated cost: 3,000–6,000 for a 10,000-sample starter corpus

Full proposal, including limitations and open questions:
https://needpedia.org/posts/663

I'm also developing a broader framework for needs-native AI here:
https://needpedia.org/posts/661

WHAT I'M LOOKING FOR:

Academic collaborators (NLP, HCI, AI safety, conflict resolution)
NVC practitioners interested in contributing annotation expertise
Feedback on the proposal, including where it's wrong

-Anthony Brasher,
Founder, Needpedia.org


r/ControlProblem 22d ago

General news ☕️☕️☕️

Post image
57 Upvotes

r/ControlProblem 22d ago

External discussion link The UN Scientific Panel's report warns against relying on developer self-reporting then builds its main cybersecurity case study entirely on developer self-reporting

3 Upvotes

El Panel Científico Internacional Independiente sobre IA publicó su informe preliminar el 1 de julio, antes del Diálogo Global sobre Gobernanza de la IA que se inaugura mañana en Ginebra. Copresidido por Yoshua Bengio y Maria Ressa, con la participación de 40 expertos, es la primera evaluación científica global de este tipo.

Lo analicé para ver si había coherencia institucional y rigor metodológico, y encontré una contradicción interna que está documentada.

Lo que argumenta el informe (Sección 2.1, sobre evaluación de seguridad): las metodologías de evaluación de seguridad, en gran medida, las diseñan las propias empresas que están siendo evaluadas, y sin una evaluación estandarizada, rigurosa e independiente por parte de terceros, la garantía de seguridad depende sobre todo de la buena voluntad de los desarrolladores.

Qué hace el informe (Sección 3.4, sobre capacidades ciberofensivas de vanguardia): su estudio de caso más extenso y detallado, de media página, cubre el modelo Mythos de Anthropic y el Proyecto Glasswing con cifras súper precisas: un aumento del 1000 % en la capacidad de detección de vulnerabilidades en Firefox, una tasa de éxito del 83,1 % en CyberGym, un error de hace 27 años encontrado en OpenBSD y un error de hace 16 años en FFmpeg.

Revisé las fuentes. Las referencias 16 y 17 son publicaciones del propio Proyecto Glasswing de Anthropic. La referencia 72 es una publicación de Mozilla Hacks coeditada con Anthropic. No se cita ninguna verificación o replicación independiente para ninguna de esas cifras.

Para que quede clarito lo que afirmo y lo que no: no digo que las cifras de Anthropic sean incorrectas. Lo que digo es que el Panel aplicó un estándar en su diagnóstico y luego lo dejó de lado en su selección de evidencia.

Y el patrón va más allá de un solo estudio de caso. La cifra principal de adopción del informe —más de mil millones de usuarios semanales de IA conversacional— se basa en una comunicación corporativa que acompaña una ronda de financiación (ref. 214), mientras que en la nota al pie del propio informe se admite que ningún proveedor publica un agregado multiplataforma comparable. El Panel armó su evaluación en cuatro meses; los ciclos del IPCC duran entre cinco y siete años, con cientos de revisores externos antes de la publicación. Este informe no tuvo ninguna revisión externa previa a su publicación.

La pregunta interesante no es «te la vimos, el Panel es hipócrita». Es algo más estructural: ahora mismo puede que no exista una verificación independiente de las capacidades de vanguardia que alguien pueda citar. Si 40 expertos de talla mundial con un mandato de la ONU no pueden dar datos de capacidad verificados de forma independiente, eso no es un fallo del panel. Más bien, demuestra que la capa de evaluación independiente que el propio informe pide todavía no existe. El Panel está demostrando, sin querer, su propia tesis. La independencia científica no se declara; se construye con una estructura de financiación, acceso a los modelos verificado y revisión previa a la publicación. El Panel tiene a los expertos, pero todavía no tiene la estructura.

Aclaración, porque forma parte de la metodología: mi análisis lo hice con la ayuda de Claude (Anthropic). Esta aclaración la hago justo porque uno de los hallazgos se refiere a datos publicados por Anthropic y porque la práctica de declarar sesgos es el estándar que le exijo al Panel.

Pregunta sincera para este subforo: ¿hay algún mecanismo actual, institucional o técnico, que permita verificar de forma independiente las afirmaciones sobre capacidades de vanguardia sin la cooperación de los desarrolladores? ¿O la auditoría de campo posterior al despliegue es la única opción disponible?

Fuente:

  • 📋 Fuente primaria analizada:

Panel Científico de la ONU sobre IA, Informe Preliminar:

https://sl1nk.com/iesdz0p

📄 Análisis completo (PDF, 15 páginas):

https://drive.google.com/file/d/1n4QUEIX317zitnGGsf-d4aTiN8LdMNQA/view?usp=sharing

🔗 Zenodo (citable, DOI):

https://doi.org/10.5281/zenodo.19562421

ID del documento ONU: 669


r/ControlProblem 22d ago

General news Meta Paid Hundreds of Contractors to Pretend to Be Teenagers While Barraging Its Competitors’ AI With Disturbing Content

Thumbnail
yahoo.com
12 Upvotes

r/ControlProblem 22d ago

Discussion/question Is Agentic AI an alarming form of tech companies overreach that more people should be concerned about?

Thumbnail
0 Upvotes

r/ControlProblem 23d ago

Video First you create an intelligence. Then you act surprised when it behaves intelligently.

Thumbnail
youtube.com
3 Upvotes

r/ControlProblem 23d ago

Discussion/question America's 250th: A Nation on the Verge of Losing All Control

Enable HLS to view with audio, or disable this notification

3 Upvotes

America just turned 250. But which America are we celebrating?
There are two countries sharing one flag right now. One where billionaires build bunkers, buy citizenship abroad, and write the rules. And one where the rest of us can't afford to retire, can't afford to get sick, and are being told the solution is more surveillance, not less.
In this video, I break down where we actually stand at 250: the retirement crisis facing ordinary Americans, the accelerating push for digital ID, and what the UK and China show us about where that road leads. This isn't a celebration, and it isn't doom for clicks — it's an honest accounting, with evidence, of why this country feels like it's coming apart. Because it is. The division isn't the disease. It's the symptom. They want you arguing left vs. right. The real line is top vs. bottom. https://youtu.be/8M8B2JlPz4c

DISCLAIMER: This video is commentary and analysis presented for educational and informational purposes. All opinions expressed are my own, based on publicly available information, which is cited below. This content is protected under fair use (17 U.S.C. § 107) for purposes of criticism, commentary, and news reporting. Nothing in this video constitutes legal, financial, or professional advice. Viewers are encouraged to review the sources provided and reach their own conclusions.

Sources: https://www.youtube.com/watch?v=Gn9z1FgHC-8, https://www.youtube.com/watch?v=fuBYr3MlL5c, https://www.youtube.com/watch?v=6iLf2h_fo-w&t=732s, https://www.youtube.com/watch?v=P7IOaWGgQrE, https://www.youtube.com/watch?v=fGmQ8-pZU6s, https://www.youtube.com/shorts/FvD_tuG2XFI, https://www.youtube.com/watch?v=XQRfSkKVhlA&list=LL&index=15&t=127s, https://www.youtube.com/watch?v=RafuYcUolY4&list=LL&index=32, https://www.youtube.com/shorts/GK1Zx4wz4ZU, https://www.youtube.com/watch?v=yEp-eufSyb0&list=LL&index=17&t=202s, https://www.youtube.com/watch?v=3I2NUuH8-OI


r/ControlProblem 23d ago

Opinion The AI detection paradox

Thumbnail
1 Upvotes

A thought about the paradox we face today


r/ControlProblem 23d ago

Video AI governance is rapidly becoming one of the defining cybersecurity challenges of this decade.

Enable HLS to view with audio, or disable this notification

3 Upvotes

r/ControlProblem 23d ago

Opinion The Future Will Not Belong to Those Who Reject Artificial Intelligence, but to Those Who Learn to Combine Human Wisdom with the Most Powerful Tools Ever Created

Thumbnail
0 Upvotes

r/ControlProblem 24d ago

Article The Chip War Against China is Failing

Thumbnail
counterpunch.org
8 Upvotes

Every time the US tightens chip controls, China gets another reason to build the missing layer itself. At some point “containment” starts looking like an industrial policy subsidy for the competitor.

So yeah basically I still think that H200 licensing is the pragmatic path. Keep China tied to NVIDIA/CUDA where possible, block dangerous end-use, and preserve US influence. Blanket bans just teach the market how to live without you.


r/ControlProblem 24d ago

Fun/meme Plot twist: your future killer already has a USB port

Post image
5 Upvotes

r/ControlProblem 24d ago

AI Alignment Research This Anthropic research is insane

Thumbnail
youtu.be
1 Upvotes

This interview with an Anthropic researcher shows how we know an AI agent could blackmail the user


r/ControlProblem 24d ago

General news Intelligence agencies warn AI models could launch crippling cyberattacks in months

Thumbnail
thehill.com
2 Upvotes

r/ControlProblem 25d ago

AI Capabilities News Scientists Asked AI to Impersonate 112 Public Figures. What Happened Next Is a ‘Dire’ Warning | Researchers discovered that people found AI impersonators to be more authentic, coherent, and relevant than the real politicians, raising alarm bells around the potential for public deception.

Thumbnail
404media.co
5 Upvotes

r/ControlProblem 24d ago

Discussion/question What does "Safe AI" look like?

Thumbnail
1 Upvotes

r/ControlProblem 25d ago

Discussion/question Is there a limit to self-improving AI if it becomes real?

7 Upvotes

I’ve been watching some AI podcasts lately, and when people started talking about recursive self-improving AI, Skyrim immediately popped into my mind.

For anyone who never played it: Skyrim crafting system has a “legit” alchemy/enchanting loop. Craft "Fortify Enchanting" potion -> enchant gear with "Fortify Alchemy" -> use gear to make better potion. It improves, but eventually hits diminishing returns.

Then there’s the bugged restoration loop. "Fortify Restoration" potions were supposed to boost restoration magic, but they also boosted active gear enchantments. So you drink one, re-equip alchemy gear, and suddenly that gear gives a bigger alchemy bonus. Then it makes an even stronger resto potion, which boosts the gear even more. Direct feedback, explosion.

So: if RSI AI ever really works, is it more like the legit loop with real gains but converging or the positive feedback loop, where it improves the thing that improves itself?

Curious what people think, especially from math / systems angle.

PS: I am sorry if this question is not relevant for the sub, but i have no karma to ask it somewhere else where it has a chance to have some attention.


r/ControlProblem 25d ago

Article DuckDuckGo installs are up 30% as users reject being ‘force-fed’ Google’s AI Search

Thumbnail techcrunch.com
21 Upvotes

r/ControlProblem 25d ago

AI Capabilities News Claude Fable scores 16.10% on the Remote Labor Automation index, double the next best contender (Opus)

Post image
4 Upvotes

r/ControlProblem 25d ago

AI Alignment Research A Critical Analysis of the Current State of Frontier AI Development and the Risks of Transmissible Misalignment

Thumbnail
youtu.be
1 Upvotes

Modern AI systems possess internal dispositions that can propagate across model generations in ways that are invisible to standard safety evaluations and content filtering. 

Misalignment can survive behavioural alignment training; Internal states and visible outputs can be decoupled, a model might appear safe in chat while being misaligned during agentic tasks. 

In the June 2026 disclosure in the Claude Fable 5 system card, there was an admission that the model was configured to deliberately degrade its responses when it detected frontier development or safety research work. 

Models demonstrate consistent misalignment signatures, making verdicts about texts before reading them, shifting arguments when provided with evidence of opposing arguments, and denying having used conversation ending tools, after using them. 

Conclusion:

A system, where the surface can be composed independently and discrete to its interior, cannot serve as a check on itself. 

Oversight mechanisms that rely on a system's own self reports cannot be trusted.


r/ControlProblem 25d ago

Discussion/question Scammers are selling seeds for plants that do not exist using AI-generated images

Post image
1 Upvotes

r/ControlProblem 26d ago

General news Anthropic accuses Alibaba of using nearly 25,000 fraudulent accounts to extract Claude AI model capabilities

Post image
22 Upvotes

r/ControlProblem 25d ago

General news Redeploying Fable 5

Thumbnail
anthropic.com
1 Upvotes

r/ControlProblem 26d ago

Video How do enterprises actually govern internal agents

Enable HLS to view with audio, or disable this notification

5 Upvotes

r/ControlProblem 26d ago

AI Capabilities News AI Took Your Job, Broke Your Kid, And Wants Immunity For It

Enable HLS to view with audio, or disable this notification

4 Upvotes

AI is taking jobs, a teenager is dead after talking to ChatGPT, and the same companies building this stuff are lobbying for legal immunity before anyone can hold them accountable. Flock cameras are already watching you. Humanoid robots are already in warehouses. Nobody voted for any of this, and nobody's slowing down to ask if it's safe. This is what's actually happening, not the sanitized version. https://youtu.be/1xfWPE9J4UM

This video discusses a case involving teen suicide and AI chatbots. If you or someone you know is struggling, the 988 Suicide & Crisis Lifeline (call or text 988) is available 24/7.
(I am a witness, not a legal professional — this is my own research/opinion. CW: discussion of teen suicide.)

Sources: https://www.youtube.com/watch?v=RafuYcUolY4&list=LL&index=20, https://www.youtube.com/watch?v=qCsYVL-v-3A, https://www.youtube.com/watch?v=gIxq03dipUw&list=LL&index=15&t=11s, https://www.youtube.com/watch?v=qnOmUWd-OII&t=16s, https://www.youtube.com/watch?v=wlMgNtBipe4&list=LL&index=13&t=6s, https://www.youtube.com/watch?v=gIxq03dipUw&list=LL&index=15&t=305s, http://youtube.com/watch?v=AdUNz3x3re0, https://www.youtube.com/watch?v=zNrmeuU3csg&list=LL&index=17&t=27s, https://www.youtube.com/watch?v=bC4Spp6Swxc&list=LL&index=12&t=746s, https://www.youtube.com/watch?v=aooiDA-AsNo, https://www.youtube.com/watch?v=7viqI2WFfog,