惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Google DeepMind News
Google DeepMind News
B
Blog RSS Feed
量子位
aimingoo的专栏
aimingoo的专栏
V
Visual Studio Blog
Y
Y Combinator Blog
Vercel News
Vercel News
云风的 BLOG
云风的 BLOG
宝玉的分享
宝玉的分享
Engineering at Meta
Engineering at Meta
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
GbyAI
GbyAI
人人都是产品经理
人人都是产品经理
博客园 - 叶小钗
Stack Overflow Blog
Stack Overflow Blog
大猫的无限游戏
大猫的无限游戏
Microsoft Security Blog
Microsoft Security Blog
B
Blog
Last Week in AI
Last Week in AI
有赞技术团队
有赞技术团队
博客园 - 聂微东
腾讯CDC
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
J
Java Code Geeks

Help Net Security

Police arrest 10 suspected members of Black Axe cybercrime gang ShinyHunters claims it stole 1.4 million records from Udemy Sevii unveils Cyber Swarm Defense Mode to stop AI-driven attacks at scale Alleged Chinese hacker extradited to US over cyberattacks targeting COVID-19 research Cequence Agent Personas bring granular control and governance to enterprise AI agents NowSecure MARI gives enterprises evidence-based visibility into third-party mobile app risk The metrics killing your SOC, and what to use instead US state privacy fines reached $3.425 billion in 2025 Canada’s first SMS blaster case leads to three arrests Linux storage management tool Stratis 3.9.0 adds online encryption and cache-less pool startup TLS Connect gives SMBs a right-sized automated tool to manage TLS certificates Aptori expands its platform with autonomous offensive testing to reduce security bottlenecks Your IAM was built for humans, AI agents don’t care The AI criminal mastermind is already hiring on gig platforms 25 open-source cybersecurity tools that don’t care about your budget Product showcase: LuLu reveals unauthorized outbound connections from Mac apps Week in review: Claude Mythos finds 271 Firefox flaws, Vercel breach Users advised to drop passwords and make room for passkeys - Help Net Security Indirect prompt injection is taking hold in the wild - Help Net Security Compromised everyday devices power Chinese cyber espionage operations - Help Net Security New Cisco firewall malware can only be killed by pulling the plug - Help Net Security Meta is overhauling how you sign in, manage settings, and protect your accounts - Help Net Security Ubuntu 26.04 LTS delivers memory-safe system tools and live patching for Arm servers - Help Net Security OpenAI’s GPT-5.5 is out with expanded cybersecurity safeguards - Help Net Security AI is speeding up nation-state cyber programs - Help Net Security A study of 1,000 Android apps finds a privacy policy logging gap - Help Net Security IT spending to hit $6.31 trillion record, thanks to AI - Help Net Security Where AI in CI/CD is working for engineering teams - Help Net Security With AI's help, North Korean hackers stumbled into a near-undetectable attack - Help Net Security Hacker with a special interest in breaching sports institutions ends behind bars - Help Net Security
AI prompt confidentiality and false citations worry resea...
Sinisa Marko · 2026-04-29 · via Help Net Security

Academic researchers using commercial AI tools for literature review and idea generation are sending unpublished research questions, draft hypotheses, and proprietary domain knowledge into systems whose data handling they do not understand.

A think-aloud study of 15 researchers documents the workarounds these users have built to manage what they see as unresolved confidentiality and output verification problems in tools including Research Rabbit and Elicit AI.

AI prompt confidentiality

The study, conducted by researchers at the University of Texas at Austin and Microsoft, observed participants in real-time as they completed literature exploration, synthesis, and ideation tasks. These gaps map closely onto concerns familiar to enterprise security functions managing employee use of generative AI.

Prompt content treated as a disclosure vector

Two of the 15 participants raised direct concerns about the confidentiality of prompt content. One participant said AI platforms “will leverage the prompt you share for training, which has the potential to leak your research question or research data.” Another cited “not knowing how much of my personal data is being stored, where it is being stored, and who has access to it.”

The number is small, but the underlying behavior was widespread across the sample. Participants routinely entered draft research questions, descriptions of work in progress, and unpublished analytical framings into the tools. The study describes this as an institutional answerability problem, where end users have no visible forum through which AI vendors can be held responsible for collected, stored, or repurposed inputs.

For organizations governing employee AI use, the parallel is direct. Staff who paste internal documents, code, or strategic plans into commercial LLMs are exposed to the same opacity around retention, training reuse, and access controls.

Output verification gaps drive heavy manual review

Nine of the 15 participants reported difficulty establishing where AI-generated content came from. Retrieval pipelines, training data coverage, and curation logic were opaque, making it impossible to confirm sources. One participant described the black-box nature of the tools as a limitation for rigorous work, since sources and underlying data can’t be reported with certainty.

Seven participants treated hallucinations as a transparency failure rather than a discrete accuracy issue. The study identifies two failure modes. Attribution displacement occurs when accurate information is tied to the wrong source. Synthetic blending integrates fabricated claims alongside legitimate citations in a single output, making verification slow and error-prone.

One researcher described challenging ChatGPT about a non-existent citation and receiving an apology followed by more fabricated references. The same researcher noted a separate failure mode: citations that exist but have no connection to the topic. The study calls this a provenance problem distinct from outright fabrication.

To compensate, all 15 participants developed mitigation strategies, including social credibility heuristics such as recognizing author names or publication venues. Eight defaulted to redundant manual verification, repeatedly checking names, dates, and citations. Ten restricted AI use to low-stakes tasks and kept core analytical work outside the tools.

Implications for enterprise AI governance

The compensatory strategies documented in the study consume time and depend on domain expertise that newer staff may lack. The authors note that early-career researchers are more vulnerable to being misled by confidently stated yet poorly grounded outputs, since they have less baseline knowledge against which to calibrate.

The same dynamic appears in corporate environments where employees use LLMs for tasks outside their expertise. Confident output combined with opaque sourcing creates conditions where errors propagate without detection.

The authors recommend slower, more measured AI adoption supported by verification pipelines, metadata exposure, and clearer data governance disclosures from vendors. The study’s limitations include its small sample, an academic-only participant pool, and the fact that both tools studied have been updated since data collection. The authors call for longer-term, naturalistic research to track how user practices and vendor policies develop.

Read more: