This repository is deprecated. All of its content and history has been moved to googleapis/google-cloud-node.
-
Updated
Jul 13, 2023
This repository is deprecated. All of its content and history has been moved to googleapis/google-cloud-node.
Use the Moondream 2 model to detect faces and their gaze directions in videos.
A powerful video summarization tool that utilizes Moondream alongside multiple AI models to provide comprehensive video understanding through audio transcription, intelligent frame selection, visual description, and content summarization.
Context based video seek and search
Uses Video Intelligence to analyse and edit a video based on a given sentence.
Foundational framework for mission-critical surveillance, autonomous video intelligence, and situational awareness.
Media Info Comparison
This tool uses Moondream 2B, a powerful yet lightweight vision-language model, to detect and redact objects from videos. Moondream can recognize a wide variety of objects, people, text, and more with high accuracy while being much smaller than most vision models.
Real-time AI agent for querying live courtroom video with sub-500ms latency. Multimodal search combining video intelligence, speech-to-text, and hybrid search. Built with Stream, Twelve Labs, Deepgram, and Gemini Live API.
Multimodal video dossiers for agents: transcripts, frames, OCR, evidence, and RAG-ready knowledge.
Universal video intelligence: turn any video + a prompt into a timestamped timeline, evidence-grounded findings, and structured reports — ready for AI agents.
VideoMind AI - AI Video Intelligence OS: collection → ASR → AI analysis → reports. Cross-platform desktop app.
Local-first video intelligence orchestrator using Intel OpenVINO, Tauri, gRPC, SQLite, and agentic AI routing.
Protent is a San Francisco based Y Combinator (Winter 2026) startup building a real-time video intelligence platform for law enforcement and public safety teams. Its software monitors many live surveillance, CCTV, and bodycam feeds at once, automatically prioritizing critical streams when escalation or volatility is detected, and lets teams run…
UnReel is an AI-powered Video Intelligence engine built to decode the context of any short-form content. It is designed for users who encounter language barriers, missed situational context, or struggle to find resources mentioned in a video via a dedicated video analysis pipeline.
BRI — empathetic video intelligence with production Streamlit, FastAPI MCP, SQLite durability, and multimodal ML tooling
Coactive is a multimodal AI platform that delivers a contextual intelligence layer for modern media, turning unstructured images and video into structured, searchable intelligence. The platform ingests visual assets at scale and provides agentic and semantic search, dynamic tagging, concept training and classification, celebrity and face…
AI-powered video intelligence platform that understands, transcribes, summarizes, and extracts knowledge from screen recordings and videos.
High-performance video intelligence for live and recorded streams.
TruVideo is a video-intelligence and omnichannel communication platform for service businesses, built for the automotive service market and expanding into aviation, insurance, and commercial trucking.
Add a description, image, and links to the video-intelligence topic page so that developers can more easily learn about it.
To associate your repository with the video-intelligence topic, visit your repo's landing page and select "manage topics."