ON THIS PAGE

Application of Large Language Models and Natural Language Generation in Content Creation: A Retrieval-Augmented and Preference-Aligned Framework

Jian Zhang1,2, Yuxuan Zheng2,3, Feng Ye1,2, Xin Liu3, An Zeng3
1School of Information Science and Technology, University of Science and Technology of China, Hefei 230000, Anhui, China
2Product R&D Center, Communication Brain Technology (Zhejiang) Co., Ltd., Hangzhou 310000, Zhejiang, China
3School of Computer Science and Technology, East China Normal University, Shanghai 200333, China

Abstract

Large language models (LLMs) have substantially advanced natural language generation (NLG), but high-quality content creation remains constrained by hallucination, limited controllability, retrieval noise, domain shift, and computational cost. This study proposes Retrieval-Augmented and Preference-Aligned Generation (RAP-Gen), a unified framework that integrates retrieval-augmented generation, factual preference alignment, in-context learning, and denoising pre-training. The framework combines a dense semantic retriever with an encoder–decoder generator, introduces pairwise factual-preference supervision for controllable generation, aligns retrieval and generation distributions, and uses structured text corruption to improve robustness to incomplete or noisy inputs. Experiments are reported on summarization, long-form creative writing, and knowledge-intensive question answering using CNN/DailyMail, WritingPrompts, Natural Questions, and general language-modeling data. The reported results show that RAP-Gen improves ROUGE, BERTScore, perplexity, FactScore, and human-rated factuality relative to the baselines listed in the manuscript. Ablation results indicate that removing retrieval, preference alignment, or denoising degrades performance, supporting the complementary roles of external knowledge, preference signals, and robust pre-training. Additional analyses examine the number of retrieved documents, attention over retrieved evidence, robustness under input perturbation, model scaling, and in-context learning. Because the supplied material does not include raw predictions, code, repeated-run variance, annotator agreement statistics, or complete benchmark configurations, the numerical results are interpreted as reported experimental outcomes rather than independently reproducible evidence. The study contributes an integrated design for knowledge-grounded, preference-aware content generation and identifies evaluation and reproducibility requirements for future work.

Related Articles
Liudmyla Shlieina1, Anatolii Furman2, Mariia Zaitseva3, Uliana Maraieva3, Ruslan Lavlinskyy4
1Department of Ukrainian Studies/Department of Social Sciences and Humanities, Educational and Scientific Institute of General University Training, Dmytro Motornyi Tavria State Agrotechnological University, Zaporizhzhia, Ukraine
2Department of Psychology and Social Work, West Ukrainian National University, Ternopil, Ukraine
3Department of Philosophy, Faculty of Social Sciences, Uzhhorod National University, Uzhhorod, Ukraine
4Department of Psychology, Interregional Academy of Personnel Management, Kyiv, Ukraine
Silvia Jakabová1, Veronika Michvocíková2, Leoš Stanek1
1DTI University, Sládkovičova 533/20, 018 41 Dubnica nad Váhom, Slovakia
2University of Ss. Cyril and Methodius in Trnava, Nám. J. Herdu 2, 917 01 Trnava, Slovakia
Xiaokai Duan1
1Faculty of Humanities, Zhejiang Guangsha Vocational and Technical University of Construction, Dongyang City, Zhejiang Province 322100, China
Yevhen Kryvokhyzha1, Liudmyla Melko2, Olesia Dolynska3, Volodymyr Velykochyy4, Maryna Kryvoberets5
1Department of Food Technologies, Hotel and Restaurant Services, Chernivtsi Institute of Trade and Economics of the State University of Trade and Economics, Chernivtsi, Ukraine
2Department of Tourism, KROK University, Kyiv, Ukraine
3Department of Tourism, Theory and Methods of Physical Education, and Valeology, Khmelnytskyi Humanitarian-Pedagogical Academy, Khmelnytskyi, Ukraine
4Faculty of Tourism, Vasyl Stefanyk Carpathian National University, Ivano-Frankivsk, Ukraine
5Interregional Academy of Personnel Management, Kyiv, Ukraine
Chenchen Li1, Ge Song2, Linshan Song3
1School of Urban Construction and Design, Urban Vocational College of Sichuan, Chengdu 610000, Sichuan, China
2School of Art and Technology, Chengdu College of University of Electronic Science and Technology of China, Chengdu 610000, Sichuan, China
3Office of Industry-Education Integration, Urban Vocational College of Sichuan, Chengdu 610000, Sichuan, China

Citation

Jian Zhang, Yuxuan Zheng, Feng Ye, Xin Liu, An Zeng. Application of Large Language Models and Natural Language Generation in Content Creation: A Retrieval-Augmented and Preference-Aligned Framework[J], Archives Des Sciences, Volume 76, Issue 3, 2026. 111-123. DOI: https://doi.org/10.68304/as/76313.