Lightnews — Scholar-powered news

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

There’s a lot more detail in the full paper, and I would love to hear your thoughts and feedback on it!

Check out the preprint here: arxiv.org/pdf/2506.10150

arxiv.org

June 17, 2025 at 3:14 PM

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

Huge thanks to my amazing collaborators: Fai Poungpeth, @diyiyang.bsky.social, Erina Farrell, @brucelambert.bsky.social, and @mattgroh.bsky.social 🙌

June 17, 2025 at 3:14 PM

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

LLMs, when benchmarked against reliable expert judgments, can be reliable tools for overseeing emotionally sensitive AI applications.

Our results show we can use LLMs-as-judge to monitor LLMs-as-companion!

June 17, 2025 at 3:14 PM

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

For example, in one of the conversations in our dataset, a response that an expert saw as "dismissing” the speaker’s emotions, a crowdworker interpreted as "validating" their emotions instead!

June 17, 2025 at 3:14 PM

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

These misjudgments from crowdworkers have huge implications for AI training and deployment❌

If we use flawed evaluations to train and monitor "empathic" AI, we risk creating systems that propagate a broken standard of what good communication looks like.

June 17, 2025 at 3:14 PM

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

So why the gap between experts/LLMs and crowds?

Crowdworkers often
- have limited attention
- rely on heuristics like “it’s the thought that counts”
- focusing on intentions rather than actual wording
show systematic rating inflation due to social desirability bias

June 17, 2025 at 3:14 PM

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

And when experts disagree, LLMs struggle to find a consistent signal too.

Here’s how expert agreement (Krippendorff's alpha) varied across empathy sub-components:

June 17, 2025 at 3:14 PM

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

But here’s the catch: LLMs are reliable when experts are reliable.

The reliability of expert judgments depends on the clarity of the construct. For nuanced, subjective components of empathic communication, experts often disagree.

June 17, 2025 at 3:14 PM

aakriti1kumar.bsky.social

@aakriti1kumar.bsky.social

We analyzed thousands of annotations from LLMs, crowdworkers, and experts on 200 real-world conversations

And specifically looked at 21 sub-components of empathic communication from 4 evaluative frameworks

The result? LLMs consistently matched expert judgments better than crowdworkers did! 🔥

June 17, 2025 at 3:14 PM

Reposted

Abhishek Sharma

@abhishekshar.bsky.social

Our paper: Decision-Point Guided Safe Policy Improvement
We show that a simple approach to learn safe RL policies can outperform most offline RL methods. (+theoretical guarantees!)

How? Just allow the state-actions that have been seen enough times! 🤯

arxiv.org/abs/2410.09361

Decision-Point Guided Safe Policy Improvement

Within batch reinforcement learning, safe policy improvement (SPI) seeks to ensure that the learnt policy performs at least as well as the behavior policy that generated the dataset. The core challeng...

arxiv.org

January 23, 2025 at 6:23 PM

Add to Home Screen

Light up
your news

Add to Home Screen

Light upyour news

Sign in to Lightnews

Sign up to start reading

Connect Bluesky

Connect with Bluesky

Light up
your news