Delivery status against a Video Evidence Intelligence & Forensics programme: what is running today, what was verified on real Hyderabad junction footage, what remains, and the GPU capacity a pilot needs. The working system is at suraag.xyz/console.
Proof-of-concept scenarios: six of seven pass today. The seventh — the changed-plate anomaly — is built and test-verified; it waits only for footage in which one vehicle carries two plates. A requirement-by-requirement compliance matrix is available on request.
Today the entire vision stack runs on one graphics processor, started on demand — the right economics for demonstrations, the wrong shape for operations, for four concrete reasons.
A serious case lands as dozens of DVR exports at once. One GPU processes them serially; "searchable within one to two times its duration" needs parallel analysis workers.
A GPU started cold spends its first minutes loading models — measured at 10–20 minutes. Operations need a resident GPU during working hours so the first clip of an emergency is processed in seconds.
Question parsing currently uses a vetted external text service — question text only, never imagery. A tender-grade deployment self-hosts the language model, which needs one high-memory (96 GB-class) GPU and makes the platform fully self-contained.
Helmet, face-covering and garment recognition improve by retraining on the deployment's own camera conditions — dedicated training time, separate from serving.
| Stage | GPU footprint | What it buys | Indicative monthly |
|---|---|---|---|
| Demonstration (today) | 1 × 24 GB on demand | Full analysis of demo footage; zero idle cost | US$ 50–125 in hours |
| Station pilot | 1 × 24 GB resident + on-demand burst | Instant triage, parallel surge analysis, nightly retraining window | US$ 400–700 |
| Operations | N × 24 GB workers + 1 × 96 GB language host | City-scale concurrency; fully self-hosted; no external service in the loop | sized on measured throughput |
The measured baseline: a busy 10-second clip analyses end-to-end in ~75 seconds on one warm GPU; a 90-second junction export in roughly three minutes. Raw video never goes to the GPU provider — frames are processed in memory and only the analysis returns.
suraag.xyz/console · support@suraag.xyz