Sitemap
A list of all the posts and pages found on the site. For you robots out there, there is an XML version available for digesting as well.
Pages
Posts
Blog Post number 4
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 3
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 2
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 1
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
portfolio
Portfolio item number 1
Published:
Short description of portfolio item number 1
Portfolio item number 2
Published:
Short description of portfolio item number 2 
projects
SAPGraph: Structure-aware Scientific Document Summarization
Published:
Structure-aware extractive summarization for scientific papers using heterogeneous graph neural networks
Recommended citation: S Qi, L Li, Y Li, J Jiang, D Hu, Y Li, Y Zhu, Y Zhou, M Litvak, N Vanetik. (2022). "SAPGraph: Structure-aware extractive summarization for scientific papers with heterogeneous graph." Proceedings of the 2nd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics.
Download Paper
Awesome-Hallu-Eval: A Comprehensive Collection of Hallucination Evaluation Methods
Published:
A curated list of evaluators designed to assess model hallucination in language models
Recommended citation: S Qi. (2024). "Awesome-Hallu-Eval: A Comprehensive Collection of Hallucination Evaluation Methods." GitHub Repository.
Download Paper
FHSumBench: Evaluating LLMs’ Assessment of Mixed-Context Hallucination Through the Lens of Summarization
Published:
Data and code for evaluating LLMs assessment of mixed-context hallucination through summarization
Recommended citation: S Qi, R Cao, Y He, Z Yuan. (2025). "Evaluating LLMs Assessment of Mixed-Context Hallucination Through the Lens of Summarization." arXiv preprint arXiv:2503.01670.
Download Paper
publications
CIST@CL-SciSumm 2020, LongSumm 2020: Automatic scientific document summarization
Published in SDP Workshop @ EMNLP 2020, 2020
Automatic scientific document summarization for CL-SciSumm 2020 and LongSumm 2020 shared tasks.
Recommended citation: L Li, Y Xie, W Liu, Y Liu, Y Jiang, S Qi, X Li. (2020). "CIST@CL-SciSumm 2020, LongSumm 2020: Automatic scientific document summarization." Proceedings of the First Workshop on Scholarly Document Processing. 225-234.
Download Paper
Subjective bias in abstractive summarization
Published in arXiv preprint, 2021
Investigating subjective bias in abstractive summarization and its impact on summary quality.
Recommended citation: L Li, W Liu, M Litvak, N Vanetik, J Pei, Y Liu, S Qi. (2021). "Subjective bias in abstractive summarization." arXiv preprint arXiv:2106.10084.
Download Paper
SAPGraph: Structure-aware extractive summarization for scientific papers with heterogeneous graph
Published in AACL-IJCNLP 2022, 2022
Structure-aware extractive summarization for scientific papers using heterogeneous graph neural networks.
Recommended citation: S Qi, L Li, Y Li, J Jiang, D Hu, Y Li, Y Zhu, Y Zhou, M Litvak, N Vanetik. (2022). "SAPGraph: Structure-aware extractive summarization for scientific papers with heterogeneous graph." Proceedings of the 2nd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics.
Download Paper
A Survey of Automatic Hallucination Evaluation on Natural Language Generation
Published in arXiv preprint, 2024
Comprehensive survey of automatic hallucination evaluation methods in natural language generation.
Recommended citation: S Qi, L Gui, Y He, Z Yuan. (2024). "A Survey of Automatic Hallucination Evaluation on Natural Language Generation." arXiv preprint arXiv:2404.12041.
Download Paper
Can LLMs Simulate L2-English Dialogue? An Information-Theoretic Analysis of L1-Dependent Biases
Published in ACL 2025, 2025
Information-theoretic analysis of L1-dependent biases in LLM simulation of L2-English dialogue.
Recommended citation: R Gao, X Wu, T Kuribayashi, M Ye, S Qi, C Roever, Y Liu, Z Yuan, JH Lau. (2025). "Can LLMs Simulate L2-English Dialogue? An Information-Theoretic Analysis of L1-Dependent Biases." Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (ACL 2025).
Download Paper
EnigmaToM: Improve LLMs’ Theory-of-Mind Reasoning Capabilities with Neural Knowledge Base of Entity States
Published in Findings of ACL 2025, 2025
Improving LLMs theory-of-mind reasoning capabilities using neural knowledge base of entity states.
Recommended citation: H Xu, S Qi, J Li, Y Zhou, J Du, C Catmur, Y He. (2025). "EnigmaToM: Improve LLMs Theory-of-Mind Reasoning Capabilities with Neural Knowledge Base of Entity States." Findings of the Association for Computational Linguistics: ACL 2025.
Download Paper
Evaluating LLMs’ Assessment of Mixed-Context Hallucination Through the Lens of Summarization
Published in Findings of ACL 2025, 2025
Evaluation of LLMs assessment capabilities for mixed-context hallucination in summarization tasks.
Recommended citation: S Qi, R Cao, Y He, Z Yuan. (2025). "Evaluating LLMs Assessment of Mixed-Context Hallucination Through the Lens of Summarization." Findings of the Association for Computational Linguistics: ACL 2025.
Download Paper
NOVER: Incentive Training for Language Models via Verifier-Free Reinforcement Learning
Published in EMNLP 2025, 2025
Incentive training for language models using verifier-free reinforcement learning approach.
Recommended citation: W Liu, S Qi, X Wang, C Qian, Y Du, Y He. (2025). "NOVER: Incentive Training for Language Models via Verifier-Free Reinforcement Learning." Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP 2025).
Download Paper
Beyond Perplexity: Let the Reader Select Retrieval Summaries via Spectrum Projection Score
Published in AAAI 2026, 2025
A label-free metric for aligning retrieved summaries with reader models in retrieval-augmented generation.
Recommended citation: Z Hu, Q Zhu, S Qi, Y He, H Yan, L Gui. (2025). "Beyond Perplexity: Let the Reader Select Retrieval Summaries via Spectrum Projection Score." Proceedings of the AAAI Conference on Artificial Intelligence (AAAI 2026).
Download Paper
When Thinking Backfires: Mechanistic Insights into Reasoning-Induced Misalignment
Published in ICLR 2026, 2025
Mechanistic analysis of how strengthening reasoning capabilities can induce misalignment in LLMs.
Recommended citation: H Yan, H Xu, S Qi, S Yang, Y He. (2026). "When Thinking Backfires: Mechanistic Insights into Reasoning-Induced Misalignment." International Conference on Learning Representations (ICLR 2026).
Download Paper
Position: Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain
Published in ICML 2026 (Position Paper Track), 2026
A position paper on why self-evolving LLM systems plateau and how to sustain self-improvement.
Recommended citation: W Liu, S Qi, Y Du, Y He. (2026). "Position: Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain." Forty-third International Conference on Machine Learning (ICML 2026), Position Paper Track.
Download Paper
HiMe: Real-Time Self-Hosted Personal Agent Platform for Health Insights with Wearable Devices
Published in EMNLP 2026 (System Demonstrations), 2026
A privacy-preserving, locally deployable LLM agent platform for real-time health insights from wearables.
Recommended citation: W Liu, S Qi, L Zhang, L Tudor Car, Y He. (2026). "HiMe: Real-Time Self-Hosted Personal Agent Platform for Health Insights with Wearable Devices." Proceedings of the 2026 Conference on Empirical Methods in Natural Language Processing: System Demonstrations (EMNLP 2026 Demo).
Download Paper
Detecting Contextual Hallucinations in Large Language Models with Frequency-Aware Attention
Published in ICML 2026, 2026
A frequency-aware view of attention dynamics for lightweight contextual hallucination detection.
Recommended citation: S Qi, Y Chen, R Zhao, Q Zhu, Z Hu, W Liu, Y He, Z Yuan, L Gui. (2026). "Detecting Contextual Hallucinations in Large Language Models with Frequency-Aware Attention." Forty-third International Conference on Machine Learning (ICML 2026).
Download Paper
talks
Talk 1 on Relevant Topic in Your Field
Published:
This is a description of your talk, which is a markdown file that can be all markdown-ified like any other post. Yay markdown!
Conference Proceeding talk 3 on Relevant Topic in Your Field
Published:
This is a description of your conference proceedings talk, note the different field in type. You can put anything in this field.
teaching
Teaching experience 1
Undergraduate course, University 1, Department, 2014
This is a description of a teaching experience. You can use markdown like any other post.
Teaching experience 2
Workshop, University 1, Department, 2015
This is a description of a teaching experience. You can use markdown like any other post.
