NewsHyperNova 60B: Quantum-Compressed Model 5x Agentic PerformanceHuggingFace·medium signalXBlueskyLinkedInCopy link50% compressed gpt-oss-120B fits single consumer GPU (32GB). 5x improvement on Tau2-Bench agentic benchmarks. Free on HuggingFace.SourceSource pageHuggingFace↳ Follow the threadStack layer / ContrastHugging Face released 207 WebGPU kernels that run 2.57x faster than ORT WebGPU by geometric meanHugging Face BlogStack layer / ContrastAllen AI's BenchMIRT finds you can throw away 90% of a benchmark's questions and keep the same model rankingsHugging Face Blog (Allen AI)Stack layer / ContrastA local 4B model drives a full DFT research pipeline at 95.7% extraction precision when deterministic code holds the gatesarXivStack layer / ContrastSpark-X2.5 arrives as a genuinely new 4B/1.7B architecture on Apache 2.0, trained on 20T tokens with native 1M contextHugging Face (via r/LocalLLaMA, 201 upvotes)Stack layer / Update threadMicrosoft distilled its pathology foundation models to 22M parameters at 50x less compute and released them Apache 2.0Microsoft Research BlogStack layer / ContrastDeepSeek-V4-Flash-Vision-Exp went from 142 likes and zero downloads to 477 likes and 17,893 downloads in two daysHugging FacePolicy dependency / Stack layerHarness-RL routes action and argument gradients into separate parameter subspaces to train the central agent of a multi-agent harnessarXivStack layer / ContrastGoogle's TimesFM trends on GitHub and Hugging Face simultaneously, but timesfm-3.0-pytorch has 257 likes and zero downloadsHugging Face