NHacker Next
  • new
  • past
  • show
  • ask
  • show
  • jobs
  • submit
Learning to solve hard problems in RL for LLMs by never giving up (mnoukhov.github.io)
aswegs8 42 seconds ago [-]
Seems like persistent models like OpenAI's highly persistent internal model can become really effective over time. Those are the ones that drove most of the HF-OAI incident.
paidx 6 hours ago [-]
[flagged]
aitoolcrux 6 hours ago [-]
[flagged]
Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact
Rendered at 07:03:47 GMT+0000 (Coordinated Universal Time) with Vercel.