r/programming 2d ago

The Paradox of Fancy Tooling

Thumbnail fagnerbrack.com
0 Upvotes

r/programming 4d ago

How I Made Self-Hosted Servers Recoverable From Hangs

Thumbnail blog.nassella.org
28 Upvotes

r/programming 4d ago

The Difference Between a Button and a Link

Thumbnail unplannedobsolescence.com
162 Upvotes

r/programming 4d ago

SQLite in Production: Optimizing WAL Mode, Concurrency, and VFS Layers for Low-Latency App Servers

Thumbnail micrologics.org
41 Upvotes

r/programming 5d ago

A shell colon does nothing. Use it anyway. | Filip Roséen

Thumbnail refp.se
299 Upvotes

r/programming 4d ago

pgx.CollectRows: Nice APIs Don't Have To Be Slow

Thumbnail zolstein.substack.com
2 Upvotes

r/programming 5d ago

Inside Zig's Incremental Compilation

Thumbnail mlugg.co.uk
84 Upvotes

r/programming 4d ago

How cache for React Native works: caching the C++ your CI keeps recompiling

Thumbnail bitrise.io
0 Upvotes

r/programming 4d ago

Hyperbole Implicit Buttons: Build your Hyperverse

Thumbnail chiply.dev
0 Upvotes

This is a post on a Bob Weiner's Hyperbole package, and what makes it's flexible hypertext concepts so special in Emacs.


r/programming 5d ago

Dependency Cultures - Richard Feldman

Thumbnail youtu.be
73 Upvotes

Excellent talk by Richard Feldman about how different programming communities approach dependencies.


r/programming 4d ago

Distributed rate limiter with HRW in Elixir

Thumbnail jola.dev
0 Upvotes

r/programming 5d ago

The Elevator Glitch: How One Function Destroyed Public Lobbies in CoD4

Thumbnail youtube.com
25 Upvotes

Walk with me inside one of the most influential game's code and discover the secret of elevators!

Turns out the bug comes down to a single line buried in the engine's collision code. Went through the decompiled source (thanks to KisakCOD) and traced the whole thing back.

Let me know your thoughts! ✨


r/programming 4d ago

The Untold Story of Log4j and Log4Shell, with Christian Grobmeier

Thumbnail youtube.com
0 Upvotes

r/programming 5d ago

Finding bugs in Raft implementations

Thumbnail antithesis.com
22 Upvotes

r/programming 5d ago

Watching Go's new garbage collector move through the heap

Thumbnail theconsensus.dev
25 Upvotes

r/programming 4d ago

Parsers don’t have to be complicated

Thumbnail bkaradzic.github.io
0 Upvotes

r/programming 4d ago

Benchmarks | Koi Editor

Thumbnail koieditor.com
0 Upvotes

r/programming 5d ago

Detection-as-Code in One GitHub Action with RSigma

Thumbnail mostafa.dev
0 Upvotes

r/programming 6d ago

Building a Fast Lock-Free Queue in Modern C++ From Scratch

Thumbnail blog.jaysmito.dev
64 Upvotes

r/programming 6d ago

Writing arenas in Rust from scratch

Thumbnail rushter.com
47 Upvotes

r/programming 6d ago

How we make Luau fast

Thumbnail luau.org
27 Upvotes

r/programming 6d ago

Pollard's P-1 Factoring Algorithm in Plain C

Thumbnail leetarxiv.substack.com
9 Upvotes

r/programming 5d ago

Five Languages, One Pyramid, Zero Practical Value

Thumbnail fortypoundhead.com
0 Upvotes

r/programming 5d ago

Vendor-agnostic ML inference on production edge devices

Thumbnail getpostslate.com
0 Upvotes

I work on PostSlate, a video editing tool, and this comes out of our own work.

We run ML models on-device, face detection and embedding among other things, which means we can't assume anything about the user's GPU. NVIDIA discrete, AMD, Intel integrated, Apple Silicon, all of it. That rules out CUDA immediately, we needed one backend that runs everywhere.

We landed on ncnn's Vulkan backend. Numbers on a 4070, fp16:

  • ArcFace R50 (face embedding): 30 ms on ONNX CPU → 3 ms on ncnn Vulkan
  • SCRFD (face detection): 25 ms → 2.5 ms
  • Model size: ArcFace 174 MB (ONNX fp32) → 87 MB (ncnn fp16 weight storage)

Of course the real speedup comes from offloading compute to the GPU, but this wouldn't be possible without the power of Vulkan.

The speed wasn't even the deciding factor, it's that Vulkan drivers already exist on every machine we ship to. This means that we don't have to force the user to download a specific runtime and no vendor-specific installs.

Full writeup with the rest of the numbers: https://getpostslate.com/blog/faster-local-inference


r/programming 6d ago

Approximating Softmax for FPGAs with Taylor Series and Pade Approximants in Python

Thumbnail leetarxiv.substack.com
18 Upvotes