Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
half-kh-hacker
5 months ago
|
parent
|
context
|
favorite
| on:
A sufficiently detailed spec is code
How does post-training via reinforcement learning factor in? Does every evaluated judgement count as 'the training data' ?
abcde666777
5 months ago
|
next
[–]
I guess I'd place both within a broader umbrella: human generated input. So it still holds that they're regurgitating the decisions made by humans.
internet_points
5 months ago
|
prev
[–]
yes
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: