| 1. | | Claude.md is good for taste and project context.It's a weak place for invariants (tesseracted-labs-blog.vercel.app) |
| 3 points by kkm 3 days ago | past | discuss |
|
| 2. | | Moving coding-agent guardrails from prompts to hooks (tesseracted-labs-blog.vercel.app) |
| 4 points by kkm 4 days ago | past | discuss |
|
| 3. | | Improving Throughput by Optimising KV Cache Efficiency for Agentic Workloads (j9s.io) |
| 3 points by kkm 5 days ago | past | 1 comment |
|
| 4. | | Guide to the Kimi DeltaNet Family of linear attention (doubleword.ai) |
| 3 points by kkm 54 days ago | past |
|
| 5. | | Forensic Analysis of Container Snapshot Chains for Post-Event Reconstruction [pdf] (radostin.io) |
| 2 points by kkm 60 days ago | past |
|
| 6. | | The Agent swarm that designs itself (peterbhabra.com) |
| 1 point by kkm 81 days ago | past |
|
| 7. | | Don't Build a Router. Train the Small Model to Know When to Defer (distillabs.ai) |
| 2 points by kkm 81 days ago | past | 1 comment |
|
| 8. | | The gap between open weights LLMs and closed source LLMs (doubleword.ai) |
| 306 points by kkm 85 days ago | past | 250 comments |
|
| 9. | | InfiniBand, RoCE, and All That (fergusfinn.com) |
| 5 points by kkm 3 months ago | past |
|
| 10. | | 2678x Faster Matrix Multiplication with a GPU (0mean1sigma.com) |
| 2 points by kkm 3 months ago | past |
|
| 11. | | UCCL-EP: DeepEP-style expert parallelism on any NIC, no GPU-initiated comms (fergusfinn.com) |
| 9 points by kkm 3 months ago | past |
|
| 12. | | Hacking Google with A.I. For $500k (brutecat.com) |
| 1 point by kkm 3 months ago | past |
|
| 13. | | How to setup a local coding agent on macOS (ikyle.me) |
| 507 points by kkm 3 months ago | past | 127 comments |
|
| 14. | | Anatomy of a high-performance EP kernel (fergusfinn.com) |
| 16 points by kkm 3 months ago | past | 1 comment |
|
| 15. | | No Token Left Behind: Demystifying Token-in-Token-Out in Miles (lmsys.org) |
| 2 points by kkm 3 months ago | past |
|
| 16. | | MoE expert co-activations: Reordering inputs yields easy throughput gains (doubleword.ai) |
| 2 points by kkm 3 months ago | past |
|
| 17. | | The Economics of Speculative Decoding (fergusfinn.com) |
| 30 points by kkm 3 months ago | past | 6 comments |
|
| 18. | | Speculative KV coding: losslessly compressing KV cache by up to ~4Ă— (fergusfinn.com) |
| 155 points by kkm 3 months ago | past | 48 comments |
|
| 19. | | 70x faster cold(ish) starts for SGLang (fergusfinn.com) |
| 1 point by kkm 3 months ago | past |
|
| 20. | | Bringing Up DeepSeek-V4-Flash on AMD MI300X (fergusfinn.com) |
| 120 points by kkm 3 months ago | past | 25 comments |
|
| 21. | | Brave AI privacy:LLMs on NEAR AI Nvidia-Backed Trusted Execution Environments (brave.com) |
| 1 point by kkm 10 months ago | past |
|
| 22. | | How fast can an LLM go? (fergusfinn.com) |
| 2 points by kkm 10 months ago | past |
|
| 23. | | FHE can be leveraged for LLMs such as ChatGPT in a privacy-preserving manner (huggingface.co) |
| 4 points by kkm on Aug 13, 2024 | past |
|
| 24. | | Harnessing the Power of Large Language Models for Insightful Review Analysis (holidaycheck.com) |
| 1 point by kkm on April 16, 2024 | past |
|
| 25. | | A Privacy-First approach to use AI for understanding our Customers Better (holidaycheck.com) |
| 1 point by kkm on March 12, 2024 | past |
|
| 26. | | How to make LLMs go fast (vgel.me) |
| 2 points by kkm on Dec 19, 2023 | past |
|
| 27. | | Leveraging Large Language Models for Sentiment Classification in Hotel Reviews (holidaycheck.com) |
| 1 point by kkm on Nov 7, 2023 | past |
|
| 28. | | Managers Should Think More Like Hackers (hbr.org) |
| 3 points by kkm on April 9, 2023 | past |
|
| 29. | | Activists park burnt-out tank outside Russian Embassy in Berlin (thelocal.com) |
| 3 points by kkm on Feb 25, 2023 | past |
|
| 30. | | Exposing 185M+ Indians’ Personal Information and much more (blog.robinjust.in) |
| 3 points by kkm on Feb 21, 2023 | past |
|
|
| More |