Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
mordae
15 days ago
|
parent
|
context
|
favorite
| on:
Handbook.md shows that long policy documents do no...
To run at decent speed, all models try hard to use only most likely relevant part of the context and most likely relevant weights (MoE) to predict the next token. Doing the math in full is unfeasible.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: