Learning Text Styles: A Study on Transfer, Attribution, and Verification

📅 2025-07-22
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This work addresses the computational modeling and controllable manipulation of textual style, systematically tackling three core tasks: Text Style Transfer (TST), Author Attribution (AA), and Author Verification (AV). We propose a parameter-efficient fine-tuning framework built upon large language models, integrating contrastive learning with instruction tuning to achieve disentangled style representations and explicit content–style separation. Crucially, we unify TST and AV under a single, interpretable contrastive disentanglement paradigm—enhancing transfer fidelity and attribution reliability. Empirically, our method achieves state-of-the-art performance across multiple standard benchmarks: it preserves content fidelity in TST while attaining SOTA accuracy on both AA and AV tasks. Moreover, the approach demonstrates strong generalization across domains and styles, alongside inherent interpretability through disentangled latent representations.

Technology Category

Natural Language Processing: Sentiment Analysis, Stylistic Analysis, and Argument MiningMachine Learning: Large Multimodal Models (LMMs)Computer Vision: Large Vision Models

Application Category

Web Mining and Content Analysis: Large pretrained models with web dataSearch and Retrieval-Augmented AI: Web learning to rank, online learning, and counterfactual learning for rankingUser Modeling, Personalization and Recommendation: Attacks and countermeasures in recommendation systems
📝 Abstract
This thesis advances the computational understanding and manipulation of text styles through three interconnected pillars: (1) Text Style Transfer (TST), which alters stylistic properties (e.g., sentiment, formality) while preserving content; (2)Authorship Attribution (AA), identifying the author of a text via stylistic fingerprints; and (3) Authorship Verification (AV), determining whether two texts share the same authorship. We address critical challenges in these areas by leveraging parameter-efficient adaptation of large language models (LLMs), contrastive disentanglement of stylistic features, and instruction-based fine-tuning for explainable verification.
Problem

Research questions and friction points this paper is trying to address.

Altering text styles while preserving content
Identifying authors via stylistic fingerprints
Verifying shared authorship between texts
Innovation

Methods, ideas, or system contributions that make the work stand out.

Parameter-efficient adaptation of large language models
Contrastive disentanglement of stylistic features
Instruction-based fine-tuning for explainable verification
🔎 Similar Papers
No similar papers found.