Hacker Newsnew | past | comments | ask | show | jobs | submit | fromlogin
JWST Confirms a Runaway Supermassive Black Hole via Its Supersonic Bow Shock (arxiv.org)
2 points by gnabgib 9 months ago | past
SimpleQA Verified: Reliable Factuality Benchmark to Measure Parametric Knowledge (arxiv.org)
3 points by handfuloflight 9 months ago | past
Latent Collaboration in Multi-Agent Systems (arxiv.org)
1 point by walterbell 9 months ago | past
Can large language models generalize analogy solving like children can? [pdf] (arxiv.org)
2 points by Gys 9 months ago | past | 1 comment
Sudoku-Bench: Evaluating Creative Reasoning With Sudoku Variants (arxiv.org)
1 point by optimalsolver 9 months ago | past
Let's (Not) Just Put Things in Context: Test-Time Training for Long-Context LLMs (arxiv.org)
1 point by elashri 9 months ago | past
Apriel-H1: Towards Efficient Enterprise Reasoning Models (arxiv.org)
1 point by guiriduro 9 months ago | past | 1 comment
High-Speed On-Chip Photonic Memory and Compute Systems (arxiv.org)
2 points by beardyw 9 months ago | past
Weird Generalization and Inductive Backdoors: New Ways to Corrupt LLMs (arxiv.org)
7 points by bediger4000 9 months ago | past | 1 comment
Cisco Integrated AI Security and Safety Framework Report (arxiv.org)
2 points by takira 9 months ago | past
AI-Triggered Delusional Ideation as Folie a Deux Technologique (arxiv.org)
10 points by kelseyfrog 9 months ago | past | 2 comments
Beaver: An Efficient Deterministic LLM Verifier (arxiv.org)
1 point by tshanmu 9 months ago | past | 1 comment
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs (arxiv.org)
2 points by chengchang316 9 months ago | past | 2 comments
Detailed balance in large language model-driven agents (arxiv.org)
48 points by Anon84 9 months ago | past | 5 comments
Large language models are not about natural language (arxiv.org)
2 points by 50kIters 9 months ago | past
When Quantum Federated Learning Meets Blockchain in 6G Networks (arxiv.org)
1 point by JoachimS 9 months ago | past
Predicting the Emergence of the EV Industry (arxiv.org)
1 point by burekqueen 9 months ago | past
Audio Frequency-Time Dual Domain Evaluation on Depression Diagnosis (arxiv.org)
2 points by internetguy 9 months ago | past
A quarter of US-trained scientists eventually leave (arxiv.org)
155 points by bikenaga 9 months ago | past | 172 comments
Detailed balance in large language model-driven agents (arxiv.org)
1 point by gmays 9 months ago | past
Gradient Descent Algorithm Survey (arxiv.org)
3 points by PaulHoule 9 months ago | past | 1 comment
Batched Ranged Random Integer Generation (arxiv.org)
1 point by ibobev 9 months ago | past
Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in LLMs (arxiv.org)
3 points by trueduke 9 months ago | past
The End of Concept Nativism (arxiv.org)
1 point by usingla 9 months ago | past
Reviving, reproducing, and revisiting Axelrod's second tournament (arxiv.org)
3 points by m-hodges 9 months ago | past
AI Agents vs. Pentesters (arxiv.org)
2 points by takira 9 months ago | past
Weird Generalization and Inductive Backdoors: New Ways to Corrupt LLMs (arxiv.org)
3 points by _tk_ 9 months ago | past
Baseline: Operation-Based Evolution and Versioning of Data (arxiv.org)
2 points by todsacerdoti 9 months ago | past
An open-source energy and latency profiler for LLMs: ELANA (arxiv.org)
1 point by nkko 9 months ago | past
Stronger Normalization-Free Transformers (arxiv.org)
4 points by mfiguiere 9 months ago | past

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: