Search papers, labs, and topics across Lattice.
Instituto Polit茅cnico Nacional (IPN)
4
0
3
Sentiment drift in RLHF can strip emotional nuance from summaries, but a new regularization technique can mitigate this effect without sacrificing quality.
Classifying political sentiment on social media is more complex than expected, with leading models achieving modest F1-scores that underscore the challenge.
Sentiment drift in RLHF-based summarization can suppress emotional expressiveness, revealing a critical trade-off in alignment methods that researchers must address.
AI's current limitations in adaptability stem from its reliance on psychological learning theories, suggesting a need for representational architectures where systematic behavior is inherent, not accidental.