Hacker Newsnew | past | comments | ask | show | jobs | submit | kkm's submissionslogin
1.Claude.md is good for taste and project context.It's a weak place for invariants (tesseracted-labs-blog.vercel.app)
3 points by kkm 3 days ago | past | discuss
2.Moving coding-agent guardrails from prompts to hooks (tesseracted-labs-blog.vercel.app)
4 points by kkm 4 days ago | past | discuss
3.Improving Throughput by Optimising KV Cache Efficiency for Agentic Workloads (j9s.io)
3 points by kkm 5 days ago | past | 1 comment
4.Guide to the Kimi DeltaNet Family of linear attention (doubleword.ai)
3 points by kkm 54 days ago | past
5.Forensic Analysis of Container Snapshot Chains for Post-Event Reconstruction [pdf] (radostin.io)
2 points by kkm 60 days ago | past
6.The Agent swarm that designs itself (peterbhabra.com)
1 point by kkm 81 days ago | past
7.Don't Build a Router. Train the Small Model to Know When to Defer (distillabs.ai)
2 points by kkm 81 days ago | past | 1 comment
8.The gap between open weights LLMs and closed source LLMs (doubleword.ai)
306 points by kkm 85 days ago | past | 250 comments
9.InfiniBand, RoCE, and All That (fergusfinn.com)
5 points by kkm 3 months ago | past
10.2678x Faster Matrix Multiplication with a GPU (0mean1sigma.com)
2 points by kkm 3 months ago | past
11.UCCL-EP: DeepEP-style expert parallelism on any NIC, no GPU-initiated comms (fergusfinn.com)
9 points by kkm 3 months ago | past
12.Hacking Google with A.I. For $500k (brutecat.com)
1 point by kkm 3 months ago | past
13.How to setup a local coding agent on macOS (ikyle.me)
507 points by kkm 3 months ago | past | 127 comments
14.Anatomy of a high-performance EP kernel (fergusfinn.com)
16 points by kkm 3 months ago | past | 1 comment
15.No Token Left Behind: Demystifying Token-in-Token-Out in Miles (lmsys.org)
2 points by kkm 3 months ago | past
16.MoE expert co-activations: Reordering inputs yields easy throughput gains (doubleword.ai)
2 points by kkm 3 months ago | past
17.The Economics of Speculative Decoding (fergusfinn.com)
30 points by kkm 3 months ago | past | 6 comments
18.Speculative KV coding: losslessly compressing KV cache by up to ~4Ă— (fergusfinn.com)
155 points by kkm 3 months ago | past | 48 comments
19.70x faster cold(ish) starts for SGLang (fergusfinn.com)
1 point by kkm 3 months ago | past
20.Bringing Up DeepSeek-V4-Flash on AMD MI300X (fergusfinn.com)
120 points by kkm 3 months ago | past | 25 comments
21.Brave AI privacy:LLMs on NEAR AI Nvidia-Backed Trusted Execution Environments (brave.com)
1 point by kkm 10 months ago | past
22.How fast can an LLM go? (fergusfinn.com)
2 points by kkm 10 months ago | past
23.FHE can be leveraged for LLMs such as ChatGPT in a privacy-preserving manner (huggingface.co)
4 points by kkm on Aug 13, 2024 | past
24.Harnessing the Power of Large Language Models for Insightful Review Analysis (holidaycheck.com)
1 point by kkm on April 16, 2024 | past
25.A Privacy-First approach to use AI for understanding our Customers Better (holidaycheck.com)
1 point by kkm on March 12, 2024 | past
26.How to make LLMs go fast (vgel.me)
2 points by kkm on Dec 19, 2023 | past
27.Leveraging Large Language Models for Sentiment Classification in Hotel Reviews (holidaycheck.com)
1 point by kkm on Nov 7, 2023 | past
28.Managers Should Think More Like Hackers (hbr.org)
3 points by kkm on April 9, 2023 | past
29.Activists park burnt-out tank outside Russian Embassy in Berlin (thelocal.com)
3 points by kkm on Feb 25, 2023 | past
30.Exposing 185M+ Indians’ Personal Information and much more (blog.robinjust.in)
3 points by kkm on Feb 21, 2023 | past

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: