Papers for

advertising platform developers

Papers whose findings have a practical use for this group, as judged from the abstract. Open a paper to read what it means in practice.

Embedding subspaces enable flexible multi-goal recommendations at scale

Embedding Subspace Partitioning for Dynamic Multi-Objective Retrieval

Abstract: Modern industrial recommender systems must optimize across competing objectives, balancing semantic relevance with business metrics such as engagement and revenue. While bi-encoders dominate large-scale retrieval due to their efficiency, they collapse these heterogeneous signals into a single static embedding space. This design creates a fundamental limitation: once trained, the retriever cannot adapt to shifting objective priorities at serving time without retraining. Moreover, joint optimization with multi-objective losses often induces interference between objectives, leading to suboptimal trade-offs. We propose Embedding Subspace Partitioning (ESP), a retrieval framework that decomposes the embedding into task-aware subspaces and replaces the single dot product with a weighted sum of per-subspace similarities, whose weights are tunable at serving time. For Transformer bi-encoders, ESP uses the model's native end-of-sequence token as a segment delimiter, with segment-aware attention masking and position encoding resets to guarantee subspace isolation in a single forward pass. Serving is performed via GPU-accelerated exhaustive kNN over one concatenated index, eliminating the need for per-objective Approximate Nearest Neighbor (ANN) infrastructure required by multi-head approaches. We evaluate ESP on an open-source benchmark built from MS MARCO. A single ESP model traces a broad Pareto frontier, consistently outperforming strong multi-task baselines across diverse operating points. In LinkedIn's job matching platform (70M+ weekly users), ESP enabled dynamic retrieval reconfiguration and delivered significant key business metric lifts.

Thu 24 SeptInformation Retrieval
The gist
Recommender systems often need to balance different goals like showing relevant items and increasing business value, but traditional methods mix all goals into one fixed understanding. The authors propose splitting the understanding into separate parts, each focused on a specific goal, so the system can adjust priorities on the fly without retraining. Their method works efficiently on large-scale systems by doing one search over combined data rather than multiple searches. Tests show this approach better balances goals and improved user engagement in a large job matching platform.
Open → 2609.30601v1

Parameter inheritance improves scaling and efficiency of recommendation models

Inherit4Rec: Parameter Inheritance for Efficient Scaling of Recommendation Models

Abstract: Scaling model capacity has emerged as an effective approach to overcoming performance bottlenecks in industrial recommender systems. However, repeatedly training larger dense models from scratch demands substantial data and time, while their growing computation conflicts with the strict serving budgets of industrial systems. Parameter inheritance provides a promising route for both dense model growth and sparse conversion, yet existing methods are primarily designed for static corpora and can suffer sharp performance drops under dynamically evolving recommendation data. To address these challenges, we propose Inherit4Rec, a parameter-inheritance framework that supports both Dense-to-Dense (D2D) growth and Dense-to-Sparse (D2S) conversion. Inherit4Rec-D2D combines hybrid growth with asymmetric training to preserve the forward function at expansion and maintain update continuity. Inherit4Rec-D2S constructs SMoE networks through co-activation-aware partitioning and a load-balancing loss, preserving dense-model capabilities while promoting balanced expert activation. Experiments on KuaiRand-1K and an industrial short-video recommendation dataset show that both transformations consistently outperform the evaluated inheritance baselines across all prediction objectives. These results demonstrate the effectiveness of Inherit4Rec for continual capacity expansion and computation-efficient sparse conversion in industrial recommender systems.

Sat 19 SeptInformation Retrieval
The gist
Recommender systems often need bigger models to improve their suggestions but training large models from scratch takes a lot of time and computing power. The authors propose a method called Inherit4Rec that helps grow models or make them sparser by reusing parameters from smaller models, so training is faster and requires less data. Their approach also handles changing recommendation data better than previous methods. Tests on public and industrial datasets show Inherit4Rec works well to keep improving recommendations while saving computation.
Open → 2609.23111v1

Fractional assignment methods ensure fair and strategic object distributions

Fractional Assignment with $\ell_1$ Preferences

Abstract: We study a fractional assignment setting where $n$ objects are to be assigned to $n$ agents with unit capacity, and each agent specifies an ideal distribution over the objects. Unlike in classic random assignment, these ideal distributions are not necessarily degenerate, as agents may prefer a mixture of objects rather than any single object. We assume that agents seek to minimize the $\ell_1$ distance between their ideal distribution and the distribution they receive, which is equivalent to maximizing the overlap between the two distributions. We propose two mechanisms, one based on water filling (WF) and the other on quadratic programming (QP), and show that both mechanisms are utilitarian-optimal (and hence Pareto efficient), envy-free, strategyproof, and satisfy equal treatment of equals. Moreover, we highlight a distinct advantage of each mechanism: while the WF mechanism satisfies the stronger property of group-strategyproofness, the QP mechanism is more robust in terms of egalitarian overlap welfare.

Wed 16 SeptComputer Science and Game Theory
The gist
This paper looks at a situation where multiple people want to share items in parts rather than getting just one whole item. Instead of each person wanting only one item, they may want a mix, and the goal is to match these preferences as closely as possible. The authors provide two ways to do this so that everyone feels treated fairly, no one envies others’ shares, and people cannot benefit by lying about their preferences. One method is better at preventing groups from cheating, while the other improves overall fairness among individuals.
Open → 2609.18299v1