| user: | veryluckyxyz |
| created: | September 30, 2014 |
| karma: | 549 |
| 1. | 32 days ago | discuss |
| 2. | |
| 3. | 116 days ago | discuss |
| 4. | |
| 5. | Hidden drivers of HRM's performance on ARC-AGI(arcprize.org) |
| 6. | |
| 7. | Deep Think with Confidence(jiaweizzhao.github.io) |
| 8. | |
| 9. | Easily Understand Rdma Technology(naddod.com) |
| 10. | |
| 11. | |
| 12. | Building and better understanding vision-language models (2024)(huggingface.co) |
| 13. | HF smolagents computer-agent demo(huggingface.co) |
| 14. | |
| 15. | |
| 16. | Retrieval with Learned Similarities(arxiv.org) |
| 17. | The Curse of Depth in Large Language Models(arxiv.org) |
| 18. | Looking Back at Speculative Decoding(research.google) |
| 19. | Long-Context GRPO(unsloth.ai) |
| 20. | |
| 21. | |
| 22. | Process Reinforcement Through Implicit Rewards(curvy-check-498.notion.site) |
| 23. | |
| 24. | Phi-4 Technical Report(arxiv.org) |
| 25. | Alignment Faking in LLMs [pdf](assets.anthropic.com) |
| 26. | |
| 27. |