喵喵小喵喵
1.34K subscribers
1.52K photos
36 videos
8 files
10.8K links
大喵喵和小喵喵的转发频道。 No endorsement implied

投喂请前往附属群
Download Telegram
Forwarded from xkcd
https://github.com/xoreaxeaxeax/asm-hall-of-shame

Instruction latency analysis usually focuses on performance optimization—making code run as fast as possible. The Assembly Hall of Shame takes the opposite approach: searching for the absolute floor of single-instruction performance.
🏆 Current Champions 🏆

x86: fxrstor64

Strategy: Use fxrstor64 to load 512-byte FPU/MMX/XMM state from a high-latency MMIO region in the PCIe fabric, then starve the fabric while the load is in flight — a fleet of hammer cores pounds a different high-latency MMIO register with tight 4-byte reads, saturating the PCIe root complex and endpoint with non-posted transactions, so CPU 0's 512-byte fxrstor64 must queue behind all that contending traffic.

🏆 Score: 198,002,498,236 cycles

🏆 Time: 62 seconds
😈2
太伟大了,Typst 允许我写这种东西,还能编译到 HTML 里
🔥3
Media is too big
VIEW IN TELEGRAM
学习了一下怎么用 media query 从 layout flow 里面摘出来一个元素扔到 sidebar 上,然后加 sticky
scribble.meow.plus 加了个目录
2026 年写 CSS 太爽了,有 parent selector
🥰1