news.volyx.in

DeepSeek: Inference-Time Scaling for Generalist Reward Modeling (arxiv.org)

163 points by tim_sw · 485 days ago · 35 comments on HN

The AI summary for this story hasn't been generated yet — it's produced hourly. Check back soon. Meanwhile, read the discussion on HN.