Hacker Newsnew | past | comments | ask | show | jobs | submit | fromlogin
PyTorch: A Reference Language (pytorch.org)
80 points by matt_d 6 days ago | past | 8 comments
Bringing PyTorch Monarch to AMD GPUs (pytorch.org)
81 points by gmays 9 days ago | past | 7 comments
Triton Plugin Extensions (pytorch.org)
2 points by zer0zzz 18 days ago | past
Triton Plugin Extensions: Enabling TLX and Custom Compiler Passes Out of the Box (pytorch.org)
3 points by matt_d 18 days ago | past
Towards Free Normalization: Fusing Normalization into GEMM and Attention Kernels (pytorch.org)
2 points by matt_d 22 days ago | past
Miles: A PyTorch-Native Stack for Large-Scale LLM RL Post-Training (pytorch.org)
1 point by gmays 26 days ago | past
From Minutes to Seconds: LLM-Guided Autotuning for Helion Kernels (pytorch.org)
3 points by matt_d 46 days ago | past
PyTorch's playbook for AI coding, as of May 2026 (pytorch.org)
3 points by matt_d 62 days ago | past
When does fragmentation occur in the CUDA caching allocator? (pytorch.org)
15 points by matt_d 63 days ago | past | 1 comment
PyTorch 2.12 Release (pytorch.org)
7 points by gmays 74 days ago | past
Running PyTorch Models on Apple Silicon GPUs with the ExecuTorch MLX Delegate (pytorch.org)
2 points by tosh 75 days ago | past
Running PyTorch Models on Apple Silicon GPUs with the ExecuTorch MLX Delegate (pytorch.org)
2 points by 0bytematt 76 days ago | past
PyTorch 2.12 Released (pytorch.org)
1 point by 0bytematt 81 days ago | past
In-Kernel Broadcast Optimization: Co-Designing Kernels for RecSys Inference (pytorch.org)
2 points by gmays 83 days ago | past
PyTorch DevLog (pytorch.org)
2 points by matt_d 84 days ago | past
SMG: The Case for Disaggregating CPU from GPU in LLM Serving (pytorch.org)
3 points by gmays 89 days ago | past
AutoSP from PyTorch (pytorch.org)
1 point by gmays 3 months ago | past
AutoSP: Long-Context LLM Training via Compiler-Based Sequence Parallelism (pytorch.org)
1 point by matt_d 3 months ago | past
A Primer on LLM Post-Training (pytorch.org)
2 points by hyperpape 3 months ago | past
Optimizing Effective Training Time for Meta's Recommendation/Ranking Workloads (pytorch.org)
1 point by gmays 3 months ago | past
Monarch: An API to Your Supercomputer (pytorch.org)
1 point by gmays 3 months ago | past
SOTA Normalization Performance with Torch.compile (pytorch.org)
1 point by salkahfi 3 months ago | past
PyTorch 2.11 Released (pytorch.org)
7 points by 0bytematt 4 months ago | past
TorchSpec: Speculative Decoding Training at Scale (pytorch.org)
2 points by zagwdt 4 months ago | past
Generalized Dot-Product Attention: Tackling Real-World Challenges in GPU Kernels (pytorch.org)
1 point by matt_d 4 months ago | past
PyTorch Broadcasting Semantics (pytorch.org)
1 point by tosh 5 months ago | past
Mooncake Joins PyTorch Ecosystem (pytorch.org)
1 point by mji 5 months ago | past
PyTorch Now Uses Pyrefly for Type Checking (pytorch.org)
5 points by ocamoss 5 months ago | past
Building Highly Efficient Inference System for Recommenders Using PyTorch (pytorch.org)
2 points by mfiguiere 5 months ago | past
Warp Specialization in Triton: Design and Roadmap (pytorch.org)
2 points by matt_d 6 months ago | past

Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: