arXiv Artificial Intelligence

MoME: Mixture-of-Memory Embeddings for Context-Aware Sparse Lookup

MoME: Mixture-of-Memory Embeddings for Context-Aware Sparse Lookup

Quick summary

arXiv:2609.15126v1 Announce Type: cross Abstract: Scaling large language models efficiently has motivated sparse capacity mechanisms such as Mixture-of-Experts and, more recently, conditional memory: token-indexed embedding tables that augment the backbone with cheap parametric lookups. Existing memory-embedding methods retrieve via a deterministic function of the surface form, which collapses different contextual senses of the same token (e.g., python the language vs. the animal) into a single fixed entry. We introduce Mixture of Memory Embeddings (MoME), a context-aware memory mechanism that

Key takeaways

  • arXiv:2609.15126v1 Announce Type: cross Abstract: Scaling large language models efficiently has motivated sparse capacity mechanisms such as Mixture-of-Experts and, more recently, conditional memory: token-indexed embedding tables that augment the backbone with cheap parametric lookups.
  • Existing memory-embedding methods retrieve via a deterministic function of the surface form, which collapses different contextual senses of the same token (e.g., python the language vs.
  • the animal) into a single fixed entry.

Why it matters

The importance of “MoME: Mixture-of-Memory Embeddings for Context-Aware Sparse Lookup” will be measured by what changes in practice. User behavior, access conditions, verifiable performance and responsible-use outcomes are the signals worth following.

Kaynak sitede devamını oku: arXiv Artificial Intelligence ↗