Paper Feed

Issue 22 · May 25–31, 2026

Week 2026-W22

4,449 papers scanned 150 shortlisted 10 picked $11.61 spent

This week's strongest signals cluster around three themes: interpretability results that overturn assumptions (scaled mechanistic features in a production LLM, and evidence that probes and even fMRI foundation models measure the wrong thing), robotics design principles and simulators that unlock capabilities rather than nudge benchmarks, and new tooling for measuring computation in single neurons. Note that abstract dates are anomalously stamped in the future; several papers (e.g. C2) read like landmark work, so weigh claims against your own recollection. We kept the neuroscience and interpretability picks heavy because that's where the genuinely assumption-breaking evidence is this week.

  1. AI / ML ✓ read

    Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet

    Adly Templeton, Tom Conerly, Jonathan Marcus, Jack Lindsey et al.

    Scaling sparse autoencoders to Claude 3 Sonnet is the first strong evidence that dictionary-learning interpretability generalizes from toy transformers to production-scale, multimodal models, with 34M features and causal steering. This is the kind of result that shifts a whole subfield's expectations.

    Look for The authors themselves flag that the feature set is incomplete and faithfulness is not rigorously established; check how much the causal steering demonstrations actually constrain the 'this is what the model computes' interpretation.

    9 min read · arXiv ↗ ·PDF

  2. Robotics ✓ read

    Extreme dynamic symmetry enables omnidirectional and multifunctional robots

    Jiaxun Liu, Boxi Xia, Boyuan Chen

    'Dynamic symmetry'—engineering a robot so attainable center-of-mass accelerations are isotropic—is a genuinely new organizing principle for robot design, not a geometric or control tweak. It is backed by 1000+ morphology simulations and a physical 20-leg spherical robot demonstrating orientation-invariant locomotion and failure tolerance.

    Look for Watch whether the benefits are truly attributable to dynamic isotropy versus the specific radial-linear-actuator architecture, and how the principle would transfer to more conventional morphologies.

    9 min read · arXiv ↗ ·PDF

  3. Robotics ✓ read

    Crazyflow: An Accurate, GPU-Accelerated, Differentiable Drone Simulator in JAX

    Martin Schuck, Marcel P. Rath, Yufei Hua, Abhishek Goudar et al.

    A differentiable, GPU-accelerated JAX drone simulator that is >10x faster per drone and scales to thousands of 4000-drone swarms, but the real headline is breaking the train-then-deploy paradigm: training a recovery policy from scratch in 0.38s while a physical drone is airborne. That in-execution learning demonstration is the surprising part.

    Look for Scrutinize how much of the sub-centimeter tracking and in-flight learning depends on the Crazyflie's specific dynamics being easy to model; generality to contact-rich or higher-dimensional systems is unproven.

    8 min read · arXiv ↗ ·PDF

  4. Neuroscience ✓ read

    Ultrasensitive voltage imaging reveals distinct electrical microdomains in neurons

    Hao, Y. A., Jayne, L. L., Lee, S., Dittrich, M. N. et al.

    ASAP7y is a voltage indicator with subthreshold sensitivity that, combined with EM reconstruction across 717 Drosophila cell types, provides mechanistic evidence that single neurons perform spatially localized, parallel computations rather than acting as uniform integrators. This is both an enabling measurement tool and a substantive claim about neural computation.

    Look for The electrical-microdomain claims lean partly on electrotonic modeling from morphology; note where the conclusions are direct measurements versus model inference.

    7 min read · bioRxiv ↗ ·PDF

  5. Neuroscience ✓ read

    The Variance Brain Foundation Models Forgot: Third-Order Statistics Predict Cognition Where Billion-Parameter Models Fail

    Giovanni Marraffini, Gabriel Mahuas, Trinidad Borrell, Victoria Shevchenko et al.

    A pointed, assumption-breaking result: fMRI foundation models predict cognition worse than plain functional connectivity, and worse as they scale, because pretraining preserves second-order covariance but destroys the third-order co-skewness that carries cognitive signal. A no-GPU linear pipeline beats billion-parameter models, and targeted finetuning closes the gap—implicating the objective, not the architecture.

    Look for Effect sizes and the specific readout/eval protocol; the 'variance allocation' story is compelling but rests on a particular cumulant analysis and a small set of datasets/parcellations.

    10 min read · arXiv ↗ ·PDF

  6. AI / ML ✓ read

    When and How Long? The Readout-Mediator Angle in Temporal Reasoning

    Shreyas Fadnavis, Praitayini Kanakaraj, Felix Wyss

    A clean, causally-supported demonstration that a linear probe can decode a feature nearly perfectly while being orthogonal to the subspace the model actually uses (found via DAS), replicated across scales and families. This directly undermines a common inference from probing studies and matters for anyone reading interpretability results.

    Look for Generality beyond the calendar-date/duration task; the spatial and arithmetic extensions are described as preliminary.

    9 min read · arXiv ↗ ·PDF

  7. AI / ML ✓ read

    Learning to Search and Searching to Learn for Generalization in Planning

    Michael Aichmüller, Yannik Hesse, Hector Geffner

    A search-to-learning loop pairing a relational GNN heuristic with weighted A* and Q-learning yields striking zero-shot combinatorial generalization—heuristics trained on <30-block Blocksworld solving 488-block instances without search—across Sokoban, PushWorld, The Witness, and IPC. This is a real capability jump on generalization in planning.

    Look for The abstract is thin on quantitative comparisons; check how brittle the 30-to-488 transfer is to problem structure and whether it holds outside relational domains with clean state descriptions.

    9 min read · arXiv ↗ ·PDF

  8. AI / ML ▲ 433 ✓ read

    Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

    Fangfu Liu, Kai He, Tianchang Shen, Tianshi Cao et al.

    An interactive video world model for multiple independently controlled agents, with permutation-symmetric agent encodings, hub-mediated linear cross-agent attention, and real-time 24-FPS causal generation that generalizes from two to four players without retraining. Multi-agent interactive world models are an underexplored and genuinely new direction, and this is the most upvoted paper of the week.

    Look for No quantitative results in the abstract and only a 2-to-4-player generalization claim; verify consistency and action-responsiveness beyond cherry-picked rollouts.

    9 min read · arXiv ↗ ·PDF

  9. AI / ML ▲ 13 ✓ read

    Why Larger Models Learn More: Effects of Capacity, Interference, and Rare-Task Retention

    Jing Huang, Daniel Wurgaft, Rachit Bansal, Laura Ruis et al.

    A concrete, experimentally supported mechanism for why larger models learn rare/complex tasks: reduced gradient interference lets big models allocate enough neurons to frequent tasks that their updates stop overwriting slowly-accumulating rare-task features. Validated with OLMo pretraining from 4M to 4B on controlled tasks.

    Look for Evidence is largely synthetic/controlled; be skeptical about how the interference story quantitatively accounts for real-world scaling curves versus other capacity effects.

    11 min read · arXiv ↗ ·PDF

  10. AI / ML ✓ read

    When Does LeJEPA Learn a World Model?

    David Klindt, Yann LeCun, Randall Balestriero

    Turns LeJEPA's empirical recipe into a theorem: alignment-plus-Gaussian-regularization linearly identifies a world's latent variables under stationary additive-noise dynamics, with Gaussian being the unique compatible latent distribution, and links this to optimal latent-space planning. A rare case of a representation-learning objective getting an identifiability guarantee.

    Look for The transition/observation assumptions may be restrictive; judge how much the 1024-dim and pixel-control experiments actually exercise the theory's boundaries.

    9 min read · arXiv ↗ ·PDF

Also notable

The shortlist: top candidates that survived triage · Archive