# AI & data breaches?

**URL:** <https://community.fibery.io/t/ai-data-breaches/10550>\
**Category:** Ideas & Features\
**Created:** [March 13, 2026, 1:27pm UTC](https://community.fibery.io/t/ai-data-breaches/10550 "2026-03-13T13:27:34Z")\
**Posts on this page:** 6\
**Page:** 1

<div class="post-metadata">

**Author:** ![myg\_ge](https://sea2.discourse-cdn.com/flex020/user_avatar/community.fibery.io/myg_ge/32/12457_2.png) [@myg\_ge](https://community.fibery.io/u/myg_ge)\
**Post date:** [March 13, 2026, 1:27pm UTC](https://community.fibery.io/t/ai-data-breaches/10550/1 "2026-03-13T13:27:34Z")

</div>

Hey guys, I have a contact who liked Fibery. They work with charities. Charities are very sensitive about donors’ personal/financial data being leaked by using AI. What solutions can we use so that AI doesn’t “steal” the sensitive data? In Fibery AI chat window, I think I could use # context. What else?

---

<div class="post-metadata">

**Author:** ![mdubakov](https://sea2.discourse-cdn.com/flex020/user_avatar/community.fibery.io/mdubakov/32/10_2.png) [@mdubakov](https://community.fibery.io/u/mdubakov)\
**Post date:** [March 13, 2026, 1:29pm UTC](https://community.fibery.io/t/ai-data-breaches/10550/2 "2026-03-13T13:29:16Z")

</div>

The only safe solution is to disable AI

---

<div class="post-metadata">

**Author:** ![myg\_ge](https://sea2.discourse-cdn.com/flex020/user_avatar/community.fibery.io/myg_ge/32/12457_2.png) [@myg\_ge](https://community.fibery.io/u/myg_ge)\
**Post date:** [March 13, 2026, 1:38pm UTC](https://community.fibery.io/t/ai-data-breaches/10550/3 "2026-03-13T13:38:17Z")

</div>

Or self-hosted AI?

---

<div class="post-metadata">

**Author:** ![myg\_ge](https://sea2.discourse-cdn.com/flex020/user_avatar/community.fibery.io/myg_ge/32/12457_2.png) [@myg\_ge](https://community.fibery.io/u/myg_ge)\
**Post date:** [March 13, 2026, 1:48pm UTC](https://community.fibery.io/t/ai-data-breaches/10550/4 "2026-03-13T13:48:25Z")

</div>

From ClickUp:

> ClickUp AI is not trained on data from your Workspace. We’ve secured licensing with our partners to ensure they do not access your data for training purposes. We also have zero data retention agreements with all of the large language model (LLM) organizations we partner with. The agreements require our partners not to retain any data from your Workspace after your data is input and processed through the LLM. Additionally, we use in-context learning (ICL) to ensure that our models are not learning from data.

What do you think about this, how truthful is this?

---

<div class="post-metadata">

**Author:** ![mdubakov](https://sea2.discourse-cdn.com/flex020/user_avatar/community.fibery.io/mdubakov/32/10_2.png) [@mdubakov](https://community.fibery.io/u/mdubakov)\
**Post date:** [March 13, 2026, 2:30pm UTC](https://community.fibery.io/t/ai-data-breaches/10550/5 "2026-03-13T14:30:52Z")

</div>

I have no idea to be honest, but in the nutshell it is all about trust to OpenAI and Anthropic. In general they indeed state that data is not used for training, etc. Same for Fibery. But it is your choice to trust them or not.

local LLM is not an option so far.

---

<div class="post-metadata">

**Author:** ![papertroll](https://sea2.discourse-cdn.com/flex020/user_avatar/community.fibery.io/papertroll/32/14157_2.png) [@papertroll](https://community.fibery.io/u/papertroll)\
**Post date:** [July 24, 2026, 2:55am UTC](https://community.fibery.io/t/ai-data-breaches/10550/6 "2026-07-24T02:55:22Z")

</div>

Just adding a +1 on this topic.

Given that AI Search is much more powerful than standard Fibery search, it is unfortunate that there is no way to run it through a self-hosted (by user or by Fibery) model as @myg_ge suggested, or to ensure zero data retention. Currently, I cannot use Fibery’s AI search due to org policies. It would be great if there were “_solutions […] so that AI doesn’t “steal” the sensitive data?_” as @myg_ge said, that aren’t just to disable AI.

**To clarify my request:** it would be very useful to have either a) ability to use a self-hosted model for semantic search or, as a less ideal but probably easier-to-implement solution, b) Fibery self-hosts the model for semantic search with zero data retention OR gets a zero data retention agreement with the subprocessors.

I would assume many other Fibery users (or potential Fibery users) would like a solution to this too, but maybe I’m being selfish 😅

Thanks for considering!

* * *

**Other relevant comments, for reference**

For clarify, I can see this has been discussed before here:

> [@Fibery Product Report: September 2024](https://community.fibery.io/t/fibery-product-report-september-2024/7572/1):
>
> Grand idea is to host own model for embeddings and get rid of third-party services like OpenAI. In this case we can index databases by default and provide great search experience for all accounts. However, in the first steps we will try to improve search quality and usability, and only then think about self-hosting.

And here:

> [@Feature Request: local large model for AI Search](https://community.fibery.io/t/feature-request-local-large-model-for-ai-search/5953/3):
>
> Not in near future unfortunately. We want to do it at some point though

And in the [Expose Semantic Search through MCP](https://community.fibery.io/t/expose-semantic-search-through-mcp/10999) request
