Sifft analyses audio and video together — connecting people, words and events across every source, with every finding traceable to its source. Briefed in plain language. On watch across archive and live. Built in the UK.
Two hundred hours of CCTV from one incident. The options: scrub it by hand, or commission a brittle pipeline that answers one fixed question. When the next question arrives — was there a red car? did they meet before? — you start over.
So footage gets searched once, badly, then never again. The audio is rarely touched at all. And live feeds are watched by motion sensors that cry wolf, or by people who can't watch forever.
Search tools retrieve clips. Sifft is an audio-visual evidence exploitation and intelligence platform — it connects what was said to who appears, where, and when, across every source in a case. Investigation, not retrieval.
Uploads or live streams. Open-vocabulary detection finds whatever you describe in plain words — no retraining, no fixed classes — alongside tracking, re-identification and word-timestamped transcription. Gunshots, alarms, raised voices: always heard, whether or not anyone asked.
Everything observed becomes structured evidence — entities, timelines, and a knowledge graph of relationships: who met whom, where, what was said, what came next. Full provenance on every record. Same footage, same evidence, every time. Nothing is invented.
Ask in plain language — "who did she meet between 3 and 5?" — and get cited, verifiable answers. Task agents like analysts: profile a subject, map a contact network, deliver the morning brief. On live feeds, standing watches raise verified alerts, not noise.
Standing watches over live site feeds — "alert me if anyone approaches the east fence after dark" — verified before they reach an operator, evidence attached. A briefed guard force, not a motion sensor.
"What happened overnight? Who came and went? Anything unusual?" — surveyed, compiled and delivered to the morning shift, unprompted, every day.
Ingest the archive once; answer every question in seconds. Follow a person across cameras, link what they said to who they met, build the dossier — cited back to frame and timecode.
Persistent understanding of activity around bases, convoys and installations — patterns of life, repeated visitors, anomalies.
Cross-camera, cross-case search over bodycam, CCTV, seized media and intercepted audio — who, where, when, with whom, saying what — at evidential standard.
Continuous monitoring of high-consequence sites where a missed event is not an option and cloud connectivity is not a given.
Originals preserved untouched. Every finding links to source and timecode; observed fact stays distinct from machine inference; every analysis records how it was made, so an investigator can verify any result — byte for byte, months later.
Standing investigations run to an authorised brief — defined objectives, controlled sources, escalation rules, audit trail. Never open-ended surveillance. Watchlists and alert rules are authored by operators, never by the system.
Cloud, on-premise or fully air-gapped — one codebase. Evidence never leaves your estate: no external cloud, no third-party APIs, no licence that expires mid-operation. Information control, by architecture.
We brief defence, security and critical-infrastructure teams — and the investors who back them. Bring representative footage and we'll show you what Sifft finds.