Searcharxiv⌕ Search

arXiv · 2610.09980

Perspective: Content-addressable memories as a computing primitive for today's AI and beyond

Abstract

Modern artificial intelligence is predominantly executed on computing architectures optimized for dense linear algebra. While this has enabled the success of contemporary neural networks and motivated compute-in-memory (CIM) architectures, a growing class of artificial intelligence (AI) workloads depends on associative retrieval, identifying stored information by content or similarity rather than by explicit memory addresses. Such operations are central to transformer attention, tree-based inference, genomic search, and other retrieval-intensive applications, yet remain inefficiently supported by conventional memory systems. In this Perspective, we argue that content-addressable memories (CAMs) provide a complementary hardware primitive for associative processing in AI. We review their ability to perform massively parallel in-memory matching, discuss how emerging memory technologies can improve density and energy efficiency, and identify hierarchical search, application-specific architectures, hardware-aware learning, and heterogeneous integration with CIM as key directions for scalable associative computing. Together, these developments suggest that associative retrieval should complement linear algebra as a foundational computing primitive for future AI systems.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Paul-Philipp Manea, Bo Wen, Peiyi He, Can Li, John Paul Strachan. 2026-10-07. Perspective: Content-addressable memories as a computing primitive for today's AI and beyond. https://arxiv.org/abs/2610.09980

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

An RRAM-based Hardware Implementation of a Radial Basis Function Neuron for Edge Classifiers

The deployment of modern machine learning (ML) solutions on resource-constrained edge devices highlights implementation challenges. This is especially true for extreme edge applications that include safety-critical components, such as autonomous navigation tasks. This paper demonstrates an artificial neural network (ANN) design leveraging Metal-Oxide Resistive RAM (RRAM) -based Analogue Content Addressable Memory (ACAM) as an efficient hardware substrate for performing metric-based classification and online adaptation on the edge. The proposed design is based on a custom Template piXeL (TXL) cell used for building the ACAM module, where each TXL cell acts as a configurable receptive field neuron. These cells employ a Radial Basis activation function to calculate the distance of an input from the programmed receptive field. The TXL can be organised into dense arrays for calculating the distance of a high-dimensional input against all stored prototypes, effectively performing fast and energy efficient similarity search. This hardware engine enables on-the-fly learning, where the receptive field parameters can be tuned to track domain shift. Through simulation of the proposed TXL-RBF classifier we can achieve 89.1\% accuracy on the MNIST dataset while consuming 185fJ per cell per operation when operating at 10MHz.

cs.ET↗

Analytical Modeling of Open- and Closed-Loop Dispersive Molecular Communication Channels with Pulsatile Flow

Molecular communication (MC) is a communication paradigm in which information is conveyed through the release, propagation, and reception of molecules. Many envisioned healthcare applications of MC are expected to operate inside the human body, where the cardiovascular system (CVS) may serve as the physical propagation environment and molecular transport is governed by diffusion and blood flow. Although blood flow is inherently pulsatile, most analytical MC channel models assume steady flow. In this paper, we develop a time-variant analytical model for dispersive MC channels with pulsatile flow. We derive the straight-duct response as a Normal distribution with time-variant mean and variance, capturing the combined effects of diffusion and pulsatile flow, and extend it to closed-loop channels through a wrapped-Normal representation. The model is validated against three-dimensional (3D) particle-based simulations (PBSs) for synthetic and physiologically motivated velocity waveforms. We further derive a closed-form first-order approximation for the first-arrival peak time and introduce the nondimensional indicator $S_{\mathrm{Rx}}$ for assessing the applicability of the reduced one-dimensional (1D) model. Our results show that pulsatility has the strongest influence when molecular transport is predominantly advective and the temporal flow variations are not averaged out, whereas stronger diffusion and faster pulsations reduce its impact on the received signal. Moreover, a PBS-based parameter sweep supports $S_{\mathrm{Rx}} \ge 2$ as a practical criterion for the applicability of the proposed analytical model. Finally, relating the analytical assumptions to representative blood-vessel classes shows that model applicability depends not only on the transport conditions but also on physiological properties such as vessel rigidity, geometric uniformity, and blood rheology.

cs.ET↗

Infrastructure-Native Computing with Electric Power Grids

Computing is conventionally implemented by hardware engineered for information processing. Here we investigate infrastructure-native computing: the use of a physical system built for another primary function as a fixed computational operator. In time-domain simulations of an IEEE 14-bus electrical network, Kirchhoff's current law and Ohm's law relate voltage-reference perturbations applied at distributed controllable nodes interfaced by power electronics converters to current responses through a topology-dependent transformation. A trained digital encoder and decoder exploit this transformation for image classification, reaching 91.5% accuracy on MNIST and 82.25% on Fashion-MNIST. The modeled operator is represented by 933 surrogate parameters, compared with 12,340 task-trained parameters for an accuracy-matched fully connected core transformation. Current superposition further supports concurrent spatial sharing of the operator and sequential temporal reuse, with per-stream accuracies above 85% and 93%, respectively, in surrogate-model evaluations. Evaluations on CIFAR-10 and repeated 10-class tasks sampled from a Butterflies-and-Moths dataset show that the incremental utility of the physical operator depends on the representation supplied by upstream digital feature extraction. These results provide a simulation-based proof of concept for infrastructure-native computing with electrical networks and identify topology, accessible control channels, and input representation as determinants of its computational utility.

cs.ET↗