Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
|
from
login
JWST Confirms a Runaway Supermassive Black Hole via Its Supersonic Bow Shock
(
arxiv.org
)
2 points
by
gnabgib
9 months ago
|
past
SimpleQA Verified: Reliable Factuality Benchmark to Measure Parametric Knowledge
(
arxiv.org
)
3 points
by
handfuloflight
9 months ago
|
past
Latent Collaboration in Multi-Agent Systems
(
arxiv.org
)
1 point
by
walterbell
9 months ago
|
past
Can large language models generalize analogy solving like children can? [pdf]
(
arxiv.org
)
2 points
by
Gys
9 months ago
|
past
|
1 comment
Sudoku-Bench: Evaluating Creative Reasoning With Sudoku Variants
(
arxiv.org
)
1 point
by
optimalsolver
9 months ago
|
past
Let's (Not) Just Put Things in Context: Test-Time Training for Long-Context LLMs
(
arxiv.org
)
1 point
by
elashri
9 months ago
|
past
Apriel-H1: Towards Efficient Enterprise Reasoning Models
(
arxiv.org
)
1 point
by
guiriduro
9 months ago
|
past
|
1 comment
High-Speed On-Chip Photonic Memory and Compute Systems
(
arxiv.org
)
2 points
by
beardyw
9 months ago
|
past
Weird Generalization and Inductive Backdoors: New Ways to Corrupt LLMs
(
arxiv.org
)
7 points
by
bediger4000
9 months ago
|
past
|
1 comment
Cisco Integrated AI Security and Safety Framework Report
(
arxiv.org
)
2 points
by
takira
9 months ago
|
past
AI-Triggered Delusional Ideation as Folie a Deux Technologique
(
arxiv.org
)
10 points
by
kelseyfrog
9 months ago
|
past
|
2 comments
Beaver: An Efficient Deterministic LLM Verifier
(
arxiv.org
)
1 point
by
tshanmu
9 months ago
|
past
|
1 comment
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs
(
arxiv.org
)
2 points
by
chengchang316
9 months ago
|
past
|
2 comments
Detailed balance in large language model-driven agents
(
arxiv.org
)
48 points
by
Anon84
9 months ago
|
past
|
5 comments
Large language models are not about natural language
(
arxiv.org
)
2 points
by
50kIters
9 months ago
|
past
When Quantum Federated Learning Meets Blockchain in 6G Networks
(
arxiv.org
)
1 point
by
JoachimS
9 months ago
|
past
Predicting the Emergence of the EV Industry
(
arxiv.org
)
1 point
by
burekqueen
9 months ago
|
past
Audio Frequency-Time Dual Domain Evaluation on Depression Diagnosis
(
arxiv.org
)
2 points
by
internetguy
9 months ago
|
past
A quarter of US-trained scientists eventually leave
(
arxiv.org
)
155 points
by
bikenaga
9 months ago
|
past
|
172 comments
Detailed balance in large language model-driven agents
(
arxiv.org
)
1 point
by
gmays
9 months ago
|
past
Gradient Descent Algorithm Survey
(
arxiv.org
)
3 points
by
PaulHoule
9 months ago
|
past
|
1 comment
Batched Ranged Random Integer Generation
(
arxiv.org
)
1 point
by
ibobev
9 months ago
|
past
Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in LLMs
(
arxiv.org
)
3 points
by
trueduke
9 months ago
|
past
The End of Concept Nativism
(
arxiv.org
)
1 point
by
usingla
9 months ago
|
past
Reviving, reproducing, and revisiting Axelrod's second tournament
(
arxiv.org
)
3 points
by
m-hodges
9 months ago
|
past
AI Agents vs. Pentesters
(
arxiv.org
)
2 points
by
takira
9 months ago
|
past
Weird Generalization and Inductive Backdoors: New Ways to Corrupt LLMs
(
arxiv.org
)
3 points
by
_tk_
9 months ago
|
past
Baseline: Operation-Based Evolution and Versioning of Data
(
arxiv.org
)
2 points
by
todsacerdoti
9 months ago
|
past
An open-source energy and latency profiler for LLMs: ELANA
(
arxiv.org
)
1 point
by
nkko
9 months ago
|
past
Stronger Normalization-Free Transformers
(
arxiv.org
)
4 points
by
mfiguiere
9 months ago
|
past
More
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: