LONG CONTEXT · EFFICIENT ATTENTION
Foundation Model for Generative Recommendation
Designed a Transformer/HSTU-based long-context model for long-range dependencies, interest exploration, and long-tail behavior in recommendation sequences. Introduced variable-length Flash Attention and dynamic batching to reduce padding overhead.
