- What it is
- A signal woven into the pixels (and the sound) of every frame — invisible to the eye. According to Google, SynthID “remains detectable even when the content is shared or undergoes a range of transformations”; Kling says its watermark stays “traceable through compression, editing, and resharing”.
- How to check
- Google: upload the video to the Gemini app and ask “Was this generated using Google AI?” — Gemini scans the image and sound and states which segments contain SynthID (files up to 100 MB and 90 seconds). Kling: upload the video at kling.ai/detection.
- Limit
- Each detector only recognises content from its own company’s tools.
Niš Fortress
Could AI have made this fire?
One fake video — yes, in half an hour. But not a dozen videos from different angles showing the same fire in the same place, at the same time, with the same smoke, light, crowd and firefighters — posted minutes apart from independent accounts.
Obstacles to a consistent AI fabrication
- Each clip is processed separately: every tool “invents” its own flame shape, height, smoke and fire progression. Matching 10+ videos from different angles is manual work. [source]
- Physics: models do not understand combustion and fluid dynamics (Physics-IQ; VideoPhy-2 — best model 47.7% on the hard subset; PhyGenBench includes combustion). [source]
- Smoke and wind: no tool takes a shared wind direction; smoke direction and density must be matched by hand in every clip.
- Light: a real fire casts flickering orange light on facades, faces and surroundings — at the same intensity in every video at the same moment.
- People: the crowd reacts, looks and points at the fire, and firefighters arrive (S-10). Inserting fire does not change the behaviour of people in the base footage.
- Length: S-08 runs continuously for 56 s — longer than the maximum editing tools process in one pass (5–30 s); joins would show as jumps in the flame shape.
- Sound: crackling, sirens and voices would have to be made separately for every clip.
- There is no published case or tool that produced 10+ mutually consistent videos of one event from different angles within an hour or two.
Leading AI video models
The key question is not “can AI make fire” but whether it can edit existing footage and do so consistently across several recordings from different angles. No tool does that across separate recordings.
| Model | Max clip per pass | Generation time | Edits real footage? | Multi-angle consistency | Watermark / status explained below |
|---|---|---|---|---|---|
| Google Veo 3.1 | 4–8 s (extension up to 148 s, 720p) | 11 s – 6 min per generation | No — creates new video from text/image | No | SynthID (invisible) on every output |
| Gemini Omni Flash (maj 2026.) | 10 s | not disclosed | Yes — conversational video editing | Context within one session, not multiple real cameras | SynthID |
| OpenAI Sora 2 | 10–25 s | — | No — video uploads blocked | No | Visible moving watermark + C2PA. Discontinued: app 26 Apr 2026, API 24 Sep 2026. |
| Runway Aleph / Aleph 2.0 | 5 s (Aleph) / 30 s (Aleph 2.0, 1080p) | ≈ 220 s per job; often several attempts | Yes — adds objects and relights real footage | Within one multi-cut video; weak across separate recordings | Usage policy bans deceptive content |
| Kling 3.0 / O1 Edit | 3–15 s | ≈ up to 2 min | Yes (O1 Edit) | Only between its own generated shots | Invisible watermark + public detector |
| Luma Ray3 Modify / Ray 3.2 | up to 10–18 s | not specified | Yes (video-to-video) | Character reference, not scene | No C2PA found |
| ByteDance Seedance 2.5 | up to 30 s | — | Partly (modify/extend) | Not across real cameras | — |
| Wan 2.2 / VACE (otvoreni kod) | ≈ 5 s, 480/720p | 5 s at 720p: up to 9 min on an RTX 4090 | Yes (VACE) | No | Local — no enforced watermark |
How AI tools mark their videos — and how to check
The “Watermark / status” column in the table above says two things: whether a tool embeds a mark in its videos showing they were made by artificial intelligence (a watermark), and whether the tool is still available at all (e.g. Sora 2 was shut down). There are four situations:
- What it is
- A logo that moves around the frame of every generated video, so anyone can see it.
- How to check
- With the naked eye — watch the whole video.
- Limit
- Tools for removing it appeared within days of launch, so the absence of a logo proves nothing. Status: Sora 2 was shut down in 2026.
- What it is
- A cryptographically signed record inside the file itself: who made the content, with which tool, and what was changed — “like a nutrition label for digital content”.
- How to check
- Upload the file at contentcredentials.org/verify.
- Limit
- Social networks strip or ignore it on upload: the Washington Post uploaded an AI video to 8 social apps and found that major platforms do not use this standard to flag it. A video downloaded from X or Instagram normally has no C2PA, whether it is real or not.
- What it is
- A video made with an open model on one’s own computer does not have to carry any mark at all.
- How to check
- There is no watermark to check.
- Limit
- That is why no detector can confirm that a video is real.
- Watermark found → strong evidence that the video (or part of it) was made with that AI tool.
- No watermark found → only means it was not made with that company’s tools — not that the video is real.
A watermark can therefore prove that something is fake, but not that something is real. That is why this site relies on evidence that does not depend on watermarks: server posting times, the agreement of several camera positions and the intersection of their sightlines, the firefighters’ and prosecutor’s confirmations, and the soot found by the police.
We have not run the videos through detectors: copies downloaded from social networks are re-encoded files, and a negative result would prove nothing. Whoever claims the videos are AI can show it — for example by finding a SynthID watermark in an original.
- Get the original file from the author — not a screen recording or a copy downloaded from social media.
- Google tools: in the Gemini app upload the video and ask “Was this generated using Google AI?” (up to 100 MB and 90 s).
- Kling: kling.ai/detection.
- Certificate of origin (C2PA): contentcredentials.org/verify.
- Visible watermark: watch the whole video for a moving logo.
How long fabricating 10–12 consistent videos would take
All values in the table are our estimate, derived from data on the tools (time per job, clip length, number of retries) and typical VFX work durations.
| Step | Best case | Realistic | VFX grade |
|---|---|---|---|
| Filming base footage from 10–12 positions | 10–20 min (people pre-positioned) | 45–90 min | half a day |
| Inserting fire, with retries (2–4 min per job, 3–5 attempts) | 40–60 min, in parallel | 5–15 h | some angles never converge |
| Matching flames, smoke, light and timing across clips | 30–60 min (fails side-by-side) | 3–8 h | days (simulation + camera tracking) |
| Compositing and sound | 10–20 min per clip | 1–4 h per clip | several days per shot |
| Posting via 10+ accounts and outlets | minutes — only if every account is controlled | independent outlets with their own reporters would have to be complicit | same |
| Total for a consistent set of 10–12 clips | ≈ 2–4 h | ≈ 1–3 days | ≈ 1–4 weeks |
That does not explain the evidence either
- The videos would have to match the exact time, place, weather, crowd, flags and fireworks of that rally — and nobody disputes the fireworks and flares (M-01, I-01).
- Soot on the rampart was found by the police investigation (P-03, P-06), and N1’s reporter saw “the remains of where it burned” (P-02).
- Firefighters confirmed the fire to N1 the same evening (M-05); the prosecutor qualified it as a criminal offence (P-07).
- The sightlines of all 18 videos intersect at one point on the southern rampart of the Fortress, within ±15 m — from different places they all film the same flames. Pre-made videos would also have to hit exactly that spot from every position (map).
- Južne vesti had its own reporter and photos from the scene (M-02).