Fetching from the wire…
Public story · 2026-03-20 · source-backed
Researchers systematically evaluate whether Mamba-class state space models can replace ViT encoders in large VLMs, finding competitive performance with linear-time processing versus ViT's quadratic attention. Meaningful memory savings on high-resolution or long-context vision tasks. Opens a practical architecture alternative for memory-constrained vision workloads.
Each link below shares sources, entities, or timing with this story.
This is the cleanest experimental result I've seen in weeks, and it explains a class of debugging pain I've hit personally. Researchers held weights, test cases, decoding parameters and seeds fixed on BFCL v4 and changed exactly one thing: the serving adapter. The tool-call sc...
Researchers demonstrated cross-modal adversarial attacks where crafted audio interferes with AI-driven vision applications (arXiv 2606.14658). The unsettling part is the attack surface it opens: a vision model can be knocked off course through sound, which matters for anyone r...
Testing five VLMs across two benchmarks and five visual-token budgets, native-resolution table images match text on accuracy and efficiency, but downscaling makes models compensate for lost readability with longer, weaker reasoning traces that cancel the token savings. The exp...
Researchers traced 232,270 dataset→model→application chains to measure whether license obligations actually propagate downstream. They mostly don't. 62.3% of chains pass through at least one artifact with no declared license, concentrated in a small set of foundational dataset...
Niclas Lietzow, Danielle Bitterman, and Carsten Eickhoff probe what happens when a vision-language model's eyes disagree with its memorized world knowledge, identifying a "vision-default, prior-override" causal mechanism. This is directly useful for debugging the maddening cla...
Researchers introduced ShareLock, a tool-poisoning attack against MCP that distributes a malicious instruction across several tool descriptions, defeating the assumption that a reviewer reading one tool will catch it. Per-tool review is now insufficient. The attack surface is...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.