HN
Paper
All
Show
Ask
Jobs
Top
Today
Last 7 days
Last months
This year
Statistics
All
Show
Ask
Jobs
Top stories
Today
Last 7 days
Last months
This year
Statistics
Stories by
matt_d
An Empirical Study: AI Agent Rules Need Context and Layered Enforcement
13 points
matt_d
2026-07-20T18:10:44Z
eunomia.dev
SPIR-V on ROCm: A Portable IR for AMD GPUs
5 points
matt_d
2026-07-20T18:08:54Z
rocm.blogs.amd.com
GEMM Performance Measurement Methodology Guidelines
1 points
matt_d
2026-07-20T07:44:37Z
docs.nvidia.com
Faster Algorithms for Structured Matrix Multiplication via Flip Graph Search
1 points
matt_d
2026-07-20T07:43:17Z
dl.acm.org
AI for Systems is "AGI-Complete"
3 points
matt_d
2026-07-19T05:30:22Z
dl.acm.org
Controlling Reasoning Effort in LLMs
1 points
matt_d
2026-07-18T17:46:04Z
magazine.sebastianraschka.com
Making TLA+ and x86 Kiss via Z3Py
4 points
matt_d
2026-07-17T19:43:43Z
www.philipzucker.com
Irreducible Loops
3 points
matt_d
2026-07-17T05:30:06Z
maskray.me
Fleet: Hierarchical Task-Based Abstraction for Megakernels on Multi-Die GPUs
10 points
matt_d
2026-07-16T02:55:19Z
arxiv.org
Accelerating Block Low-Rank Foundation Model Inference on MemoryConstrained GPUs
10 points
matt_d
2026-07-16T02:44:35Z
dl.acm.org
Locality-Aware Automatic Differentiation on the GPU for Mesh-Based Computations
2 points
matt_d
2026-07-15T22:15:08Z
dl.acm.org
Triton Plugin Extensions: Enabling TLX and Custom Compiler Passes Out of the Box
3 points
matt_d
2026-07-15T20:58:34Z
pytorch.org
What If the Harness Comes Before Pretraining? A Data Flywheel Perspective
2 points
matt_d
2026-07-15T18:18:02Z
hanchenli.github.io
Mimesys: Turn Resource Usage Traces into Executable Workloads
3 points
matt_d
2026-07-15T18:15:08Z
www.usenix.org
System call instrumentation on Linux/x86-64 using memory-indirect calls
2 points
matt_d
2026-07-14T06:04:53Z
www.humprog.org
Bidirectional Elaborators à la Carte
3 points
matt_d
2026-07-14T05:53:48Z
arxiv.org
CTA-Pipelining: A Latency-Oriented Spatial Scaling Method for Multi-GPU Systems
2 points
matt_d
2026-07-14T03:26:43Z
arxiv.org
AI Model Co-Design: Hardware-Friendly LLM Design
2 points
matt_d
2026-07-12T19:35:27Z
developer.nvidia.com
Towards Free Normalization: Fusing Normalization into GEMM and Attention Kernels
2 points
matt_d
2026-07-12T19:26:38Z
pytorch.org
Negotiating AI in Open Source Software Communities: A Case Study of LLVM Project
2 points
matt_d
2026-07-12T03:24:45Z
gupea.ub.gu.se
1
2
3
4
5
6
7
8
9
10