Fetching from the wire…
Skills2026-04-11 · source-backed
The 60% performance regression in cuBLAS dispatches the wrong kernel for all batched FP32 workloads on RTX GPUs. Profile your local inference with nsys or ncu to see if you're hitting the simt_sgemm_128x32_8x5 kernel path. If so, you're leaving 40-60% performance on the table.
Each link below shares sources, entities, or timing with this story.
The top r/MachineLearning post of the day traces the creep: argparse boilerplate, then plotting, then experiment scaffolding, dataloader refactors, first-pass training-run debugging, analysis scripts, with the author mostly reading diffs and approving. Throughput is up. But th...
The repo turns a machine into a local AI server with LLM inference, chat UI, voice, agents, workflows, RAG and image generation, at 5,206 stars, with a review queue fourteen times larger than its bug queue and a last tagged release of v2.6.0 from July 28 despite being pushed t...
After decades as a Cornell-hosted service, arXiv is establishing itself as an independent nonprofit (288↑, 66 comments). Independence could affect moderation policies, access models, and integration with downstream research tools. They're hiring a CEO at ~$300K.
Up from $4,299.99 in June against a $1,999 launch MSRP, with Korean listings at $5,112 (r/LocalLLaMA). Memory is now over 80% of a GPU's bill of materials, with 16GB of GDDR7 climbing from about $65-80 per card in mid-2025 to over $200 by year end. Local inference economics ch...
PR #19378 landed in llama.cpp this week, and I think most people are underselling what it means. Backend-agnostic tensor parallelism via --split-mode tensor makes multi-GPU inference work across AMD, Intel, and Apple Silicon. Not just CUDA. Everything. For context: llama.cpp h...
Keogh posted to r/MachineLearning (369 upvotes) that on most of the benchmark datasets used across NeurIPS, SIGKDD and VLDB time-series anomaly detection papers, plain SPC matches or beats the published SOTA, scoring perfectly on the ECG trace he shows and trivially on the TAO...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.