惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

阮一峰的网络日志
阮一峰的网络日志
The GitHub Blog
The GitHub Blog
酷 壳 – CoolShell
酷 壳 – CoolShell
雷峰网
雷峰网
U
Unit 42
Y
Y Combinator Blog
I
InfoQ
P
Proofpoint News Feed
Engineering at Meta
Engineering at Meta
量子位
Microsoft Security Blog
Microsoft Security Blog
B
Blog
The Cloudflare Blog
F
Fortinet All Blogs
Google DeepMind News
Google DeepMind News
MyScale Blog
MyScale Blog
C
Check Point Blog
S
SegmentFault 最新的问题
爱范儿
爱范儿
博客园 - 叶小钗
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Hugging Face - Blog
Hugging Face - Blog
罗磊的独立博客
T
Tailwind CSS Blog

Cheriton School of Computer Science

Master's Thesis Presentation • Computer Graphics • VR GAViewer: Immersive Visualisation and Direct Manipulation of the Conformal Model in Virtual Reality | Cheriton School of Computer Science | University of Waterloo Seminar • Algorithms and Complexity • Lower Bounds for Private Optimization Via Reconstruction Attacks | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Data Systems • Efficient Oblivious Query Processing for Property Graph Databases | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Artificial Intelligence | Machine Learning • Inferred Author Gender as a Variable Affecting LLM Behaviour | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Bioinformatics • From Candidates to Evidence: Diagnostics for Trustworthy Biological Discovery | Cheriton School of Computer Science | University of Waterloo PhD Defence • Algorithms and Complexity • Graph Property Testing and the Container Method | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Software Engineering • An Empirical Study of Transitive Vulnerability Exposure in PyPI | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Human–Computer Interaction • The Design and Development of a Virtual Patient System for Medical Education | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Software Engineering • Decoupling CLI Agent Scaffolding to Internalize Planning Across Scaffolds | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Algorithms and Complexity • On the Black-Box Impossibility of Hardness in TFNP from One-Way Functions | Cheriton School of Computer Science | University of Waterloo Seminar • Algorithms and Complexity • Geometric Distances for Curves and Graphs: From Matching to Simplification | Cheriton School of Computer Science | University of Waterloo PhD Defence • Computer Algebra | Symbolic Computation • On the Effective Algebraic Geometry of Determinantal Varieties | Cheriton School of Computer Science | University of Waterloo Seminar • Algorithms and Complexity • Computing with Full Memory in 2026 | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Algorithms and Complexity • Bipartite Density: From Mixing Time to Local Algorithms for Dense Subgraphs | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Cryptography, Security, and Privacy (CrySP) • Upgrading Security Properties for Updatable Public-Key Encryption through Modular Transformations | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Programming Languages • The Defensive Tax: Price of Defenses That Never Defend | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Algorithms and Complexity • Algorithms for Analytic Combinatorics: Positivity Bounds and D-finite Operators | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Cryptography, Security, and Privacy (CrySP) • IPFSCover: Examining Website Fingerprinting Threats in the InterPlanetary File System | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Programming Languages • Reified Generic Types for Scala 3 on the JVM | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Artificial Intelligence | Machine Learning • Abstract Reasoning with Vector Symbolic Algebras | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Artificial Intelligence | Machine Learning • Learning at Test Time: Adapting Models with Synthetic Data and Environment Interaction | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Formal Methods • Counterexample Guided Abstraction and Refinement in Dash Models | Cheriton School of Computer Science | University of Waterloo Master’s Thesis Presentation • Systems and Networking • Runtime Configuration of GPU Workloads for Energy-efficient Execution | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Artificial Intelligence | Machine Learning • Beyond Semantic Similarity: Direct Corpus Interaction for Agentic Search | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Artificial Intelligence | Machine Learning • OpenResearcher: Reproducible Training for Long-Horizon Deep Research Agents | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Software Engineering • SLA-Awareness for AI-assisted coding | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Software Engineering • Context-Aware CodeLLM Eviction for AI-assisted Coding | Cheriton School of Computer Science | University of Waterloo PhD Seminar • Bioinformatics • Recurrent Energy-Based Modeling of Side-Chain Allostery | Cheriton School of Computer Science | University of Waterloo Seminar • Bioinformatics | Artificial Intelligence • Advancing Drug Discovery with FAIR Data and Explainable AI in Biomedical Research | Cheriton School of Computer Science | University of Waterloo PhD Defence • Artificial Intelligence | Machine Learning | Bioinformatics • Generative Synthetic Data for Pre-Clinical Drug Discovery | Cheriton School of Computer Science | University of Waterloo
Master’s Thesis Presentation • Data Systems • Evaluating ...
Joe Petrik · 2026-06-23 · via Cheriton School of Computer Science

Please note: This master’s thesis presentation will take place online.

Shakiba Amirshahi, Master’s candidate
David R. Cheriton School of Computer Science

Supervisors: Professors Charles Clarke, Amira Ghenai

Large language models (LLMs) are increasingly used in applications that rely on externally retrieved evidence, including health question answering, scientific claim verification, and retrieval-augmented generation (RAG). A fundamental question underlies these systems: do LLMs genuinely reason over the evidence they receive, or do they primarily follow the stance expressed in the provided documents? This thesis investigates this question through two complementary empirical studies that examine model behavior under harmful, adversarial, and conflicting evidence conditions across health question answering and claim verification tasks.

Study 1 evaluates RAG robustness in the health domain using expert-annotated collections from the TREC 2020 and 2021 Health Misinformation Tracks. Across six LLMs, eight document types, and three query framing conditions, results show that retrieved evidence strongly shapes model behavior regardless of its reliability. Helpful documents drive ground-truth alignment to near-ceiling levels, whereas adversarial documents generated from scratch can reduce alignment to near-zero. Even a single helpful document within an otherwise adversarial retrieval pool substantially improves robustness, highlighting retrieval composition as a key factor in RAG performance. Models also demonstrate greater robustness on COVID-19 queries than on general health questions, suggesting that resistance to misleading evidence may vary across domains.

Study 2 extends the analysis to explicit claim verification, evaluating five LLMs across two domains: Check-COVID, a scientific verification benchmark, and Emergent, a journalistic rumor dataset. Under both single- and paired-document settings, models frequently reverse their verification decisions when evidence stance is flipped, struggle to maintain stable judgments under conflicting evidence, and exhibit sensitivity to document order. These vulnerabilities persist across both scientific and journalistic domains, suggesting that evidence-driven behavior is not domain-specific but a broader limitation of current verification systems. Across both studies, adversarial documents generated from scratch are consistently more damaging than naturally occurring harmful content.

Taken together, the findings show that strong benchmark performance does not necessarily indicate robust evidence reasoning. Helpful evidence can mask differences between models, whereas adversarial evidence exposes substantial variation in robustness. These results highlight the need for evaluation protocols that explicitly test model behavior under misleading and conflicting evidence, and motivate future evidence-grounded systems that assess evidence credibility rather than simply reproducing its stance.


Attend this master’s thesis presentation virtually on Zoom.