Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
christianqchung
on April 14, 2025
|
parent
|
context
|
favorite
| on:
GPT-4.1 in the API
But Llama 4 Scout does badly on long context benchmarks despite claiming 10M. It scores 1 slot above Llama 3.1 8B in this one[1].
[1]
https://github.com/adobe-research/NoLiMa
omneity
on April 14, 2025
[–]
Indeed, but it does not take away the fact that long context is not trained through long content but by scaling short content instead.
Consider applying for YC's Fall 2026 batch!
Applications
are open till July 27.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search:
[1] https://github.com/adobe-research/NoLiMa