CONSONANCE.for your information
Monday, 5 October 2026frenvi

Worth reading closely

01 — inference 4 upvotes

CARD: Cluster-level Adaptation with Reward-guided Decoding for Personalized Text Generation

QUESTION — How can large language models be adapted to individual users for fine-grained personalization while maintaining scalable deployment?

The authors present CARD, a hierarchical framework addressing the tension between fine-grained personalization and scalable deployment in large language models. CARD clusters users by shared stylistic patterns to learn group-specific LoRA adapters, and applies an implicit preference learning mechanism to infer individual user preferences without manual annotation. During inference, the base model remains frozen while personalization is injected exclusively via lightweight preference vectors and low-rank logit corrections. Experiments on the LaMP and LongLaMP benchmarks show that CARD achieves superior generation quality and significantly improves serving efficiency and scalability.

CARD combines user clustering with group-specific LoRA adapters to ensure robust generalization in low-resource settings.

An implicit preference learning mechanism allows the model to infer user-specific style preferences without manual annotation.

CARD injects personalization exclusively at decoding time by keeping the base model frozen and applying lightweight user preference vectors and low-rank logit corrections.

hulehule · 20 Sept 2026 read the original ↗
↑