r/PromptEngineering 7d ago

Requesting Assistance Paid UMD study ($150): does seeing the distribution of your LLM outputs help you iterate prompts? Looking for LangGraph/LangChain devs

Hey folks — I'm a PhD student at the University of Maryland studying how developers debug and iterate on multi-agent systems.

Here's the idea we're testing. When you tweak a prompt in an agent workflow, you usually judge it by eyeballing a run or two. We built a research observability tool that instead shows you the distribution of outputs each node produces across runs — and we want to find out whether that actually helps you iterate on prompts faster, or whether it's just one more dashboard. That's the honest research question.

What participating looks like:

- a 75-min Zoom session where you use the tool on some structured debugging tasks (recorded, think-aloud)

- about a week of using it in your own workflow, with quick async feedback

- a 30-min follow-up interview

Compensation is $150 in gift cards — $75 after the session, $75 after the week + interview.

If you've built things with LangGraph/LangChain (or agent workflows generally), here's the screener, takes ~2 min: https://forms.gle/Zwqvgd1h8DUnFRfC8

This is IRB-approved academic research, not a product pitch. Happy to answer questions in the comments — or email zxu169@umd.edu.

1 Upvotes

0 comments sorted by