---
title: Efficient Sequential Recommendation
url: https://www.emergentmind.com/topics/efficient-sequential-recommendation
type: topic
---

# Efficient Sequential Recommendation

Efficient sequential recommendation denotes the class of modeling strategies that enable sequential recommenders to deliver high-quality, low-latency, and scalable predictions under realistic constraints, such as limited compute, large catalog sizes, and dynamic user histories. The challenge is to simultaneously maximize recommendation accuracy, minimize computation and memory requirements at both training and inference, and ensure the model is deployable at scale, even with billions of items or long temporal sequences. Recent research leverages innovations in efficient attention/memory architectures, compressed output heads, hybrid linear/nonlinear sequence modeling, and adapter- or token-based parameter efficiency to address these goals.

## 1. Design Principles and Efficiency Criteria

A sequential recommender is considered efficient if it satisfies several criteria:
- **Linear or near-linear

Source: https://www.emergentmind.com/topics/efficient-sequential-recommendation