Guide7 min read

Is It Safe for French Journalists to Use AI?

AI helps summarise and translate. But pasting details that identify a source exposes that source to a third party, outside the newsroom.

By Pierre de ONYRI

The answer fits in one line. AI can help you summarise, translate or structure, but never hand it what identifies a source. A name. A number. A meeting place. A document shared in trust. Pasted into a consumer AI, these details leave the newsroom. That is exposure to a third party. Yet source protection is set by law. And the newsroom remains the data controller under the GDPR. There is a clean method: anonymise before any prompt, then restore the values locally.

The problem: identifying details pasted into AI

The habit spreads fast in newsrooms. You ask an AI to summarise a long interview. You have it translate an exchange with a foreign contact. You submit raw notes to draw up an outline. To save time, you paste the text as is. That text often holds enough to trace a source. This is where the risk begins, not in the tool itself.

A source is not betrayed only by a name. Here is what a poorly prepared prompt can expose.

  • The name, alias, number or address of a contact.
  • A place, date or time of meeting that narrows the circle of suspects.
  • A role or a rare detail that points to a single person.
  • An internal document whose formatting betrays its origin.

The stake: source protection and the GDPR

Source protection is not a mere ethics rule. It has a legal basis. Article 2 of the Law of 29 July 1881 on freedom of the press protects it in the exercise of the mission to inform the public. That protection was reinforced by Law No. 2010-1 of 4 January 2010. The principle is clear: a journalist must be able to promise contacts that they will not be identified.

This protection is strong, but it is not absolute. Article 2 says so itself. It can be set aside only if an overriding requirement of public interest justifies it. And the measures must stay strictly necessary and proportionate to the aim pursued. In other words, secrecy yields in framed cases, never through simple carelessness. That is one more reason not to expose it yourself.

One point is often forgotten. The law also covers indirect breaches. Trying to identify a source by investigating the journalist's usual circle is already a breach of source protection. This reasoning applies to any trace left outside the newsroom. A prompt handed to an outside service is one such trace. The content can be retained, reviewed or reused to train the model. Exposure alone is enough to be a problem, even without a public leak.

The GDPR adds to the secrecy duty. The newsroom is the data controller for the personal data it handles. The CNIL, France's data protection authority, states that AI gets no exemption. As soon as an AI system processes personal data, the regulation applies fully. That also covers the data contained in prompts. Responsibility is not transferred to the model provider.

Should AI be banned in the newsroom?

No. No rule forbids a journalist from using AI. The tool can summarise, translate or produce an outline in seconds. What causes trouble is exposing the identifying details, not using AI itself. So the question is not “should we give up AI?”. The question is “how do we use it without exposing a source?”. The answer holds in one word: minimisation.

AssumptionThe reality
“A summary pasted into AI stays between us”It is exposure to a third party, and a trace left outside the newsroom's control
“Source protection is absolute”It is strong but can yield to an overriding requirement of public interest
“The GDPR doesn't apply to AI”The CNIL states there is no exemption: the regulation applies fully
“AI should be banned in the newsroom”No: it is exposing the identifying details that is the problem, not the tool
The risk isn't using AI — it's the identifying details you leave behind in the prompt.

The fix: anonymise before the prompt

The fix matches what the CNIL says. When personal data is not needed for the processing, you strip it upstream. This is minimisation applied to AI. The journalist keeps the tool and saves time. The AI never sees what identifies the source. You stay in control of the secret, on your own machine.

Two-part diagram: at top, a reporter's notebook carries a “source” line in the clear (amber) that reaches an AI card, with a shield standing between a small source figure and the AI; at bottom, the same line is reduced to cobalt token chips followed by a checkmark, and the AI receives only these anonymized tokens.
After the Senate's study on source protection (Article 2 of the 1881 law, Law of 4 January 2010) and the CNIL's AI and GDPR recommendations.

In practice, you proceed step by step. You spot each identifying element. You replace it with a token before sending. The AI reasons about the shape of the text, without ever reading the name or the place. You then restore the real values, locally. Here is the order to follow.

  1. 1Spot the identifying details: names, contact details, places, dates, rare details.
  2. 2Replace them with reversible tokens, in the browser.
  3. 3Send only the anonymized text to the AI.
  4. 4Restore the real values in the reply, locally, then re-read the result.

That's what ONYRI Sanitize is for. The engine detects sensitive details — names, contact details, places, rare facts — and replaces them with reversible tokens before sending. Detection and the mapping stay in your browser. Only anonymized text reaches the model. The AI finds only tokens, never what identifies your source. You get AI's help, while reducing the exposure that source protection and the GDPR ask you to control.

Frequently asked questions

Can a journalist use AI without exposing sources?
Yes, as long as you do not paste details that identify a source. Source protection is set by Article 2 of the Law of 29 July 1881, reinforced by Law No. 2010-1 of 4 January 2010. Handing a name, a place or a document to a consumer AI exposes those details to a third party. AI stays useful on anonymized text: strip the identifying details before any prompt.
Is source protection absolute?
No. It is strong, but Article 2 of the 1881 law sets one exception. It can be set aside if an overriding requirement of public interest justifies it. The measures must then stay strictly necessary and proportionate. The law also covers indirect breaches, such as investigations into the journalist's circle. That is one more reason not to expose those details yourself, outside the newsroom.
Is anonymising before the prompt enough to protect the source?
No, not on its own. Anonymising reduces risk and supports the minimisation principle the CNIL cites. But it does not guarantee source protection and does not make the newsroom “GDPR-compliant”. It is one useful control among others. Your other duties, and your vigilance on identifying details, remain in full.

Sources & references

Keep your sensitive data in your browser

ONYRI Sanitize detects and masks your sensitive data before it reaches the AI, then restores the answer — from names to API keys.

Anonymize my prompt

Read next