NHacker Next
  • new
  • past
  • show
  • ask
  • show
  • jobs
  • submit
▲What Is RLCD? The Secret Behind Jev (di-zhang-llm.github.io)
firejake308 7 hours ago [-]
> The operational signal was always relative preference. The scalar merely hid it.

Is this another Claude-ism? "X was always Y. The Z merely hid it." Or am I overcalling it?

dilyevsky 4 hours ago [-]
You're right to question this and OP shouldn't have done it.
agos 6 hours ago [-]
not overcalling, it’s rife with claudisms
WalterGR 8 hours ago [-]
RLCD, not defined in the article, is Reinforcement Learning for Calibrated Decisions.
borgel 7 hours ago [-]
Ah, so not Reflective LCD [1] then.

[1] https://www.e3displays.com/reflective-lcd-display-monitor/

tnspacetime 7 hours ago [-]
If you need any background info on Jev read this:

https://software.human-tokens.dev/

daemonk 7 hours ago [-]
Yeah the calibration is really what makes it useful in practice for quick, small decisions. Asking a LLM to give scores to a problem will yield inconsistently scaled/anchored results that changes at a whim.

The blog is pretty heavy on statistics. I'll have to study it more when I have time. Is it essentially bootstrapping results to statistically normalize the answers?

tnspacetime 7 hours ago [-]
I have not studied it properly too. Good that it has both code and note though.
jackb4040 8 hours ago [-]
[flagged]
Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact
Rendered at 00:04:22 GMT+0000 (Coordinated Universal Time) with Vercel.