← Search

Haoyang Ruan

1 accepted papers

2026

Textual Self-Attention Network: Test-Time Preference Optimization Through Textual Gradient-Based Attention

AAAI 2026technical

Large Language Models (LLMs) have demonstrated remarkable generalization capabilities, but aligning their outputs with human preferences typically requires expensive supervised fine-tuning. Recent test-time methods leverage textual feedback to overcome this, but they often critique and revise a sing

Cited by 0SourcePDFScholar