This repository is deprecated. All of its content and history has been moved to googleapis/google-cloud-node.
-
Updated
Jul 13, 2023
This repository is deprecated. All of its content and history has been moved to googleapis/google-cloud-node.
Use the Moondream 2 model to detect faces and their gaze directions in videos.
A powerful video summarization tool that utilizes Moondream alongside multiple AI models to provide comprehensive video understanding through audio transcription, intelligent frame selection, visual description, and content summarization.
Context based video seek and search
Uses Video Intelligence to analyse and edit a video based on a given sentence.
Foundational framework for mission-critical surveillance, autonomous video intelligence, and situational awareness.
Multimodal video dossiers for agents: transcripts, frames, OCR, evidence, and RAG-ready knowledge.
Media Info Comparison
Universal video intelligence: turn any video + a prompt into a timestamped timeline, evidence-grounded findings, and structured reports — ready for AI agents.
This tool uses Moondream 2B, a powerful yet lightweight vision-language model, to detect and redact objects from videos. Moondream can recognize a wide variety of objects, people, text, and more with high accuracy while being much smaller than most vision models.
Real-time AI agent for querying live courtroom video with sub-500ms latency. Multimodal search combining video intelligence, speech-to-text, and hybrid search. Built with Stream, Twelve Labs, Deepgram, and Gemini Live API.
Learn public speaking from talks you admire — ask how they do it, watch captioned clips cut from the video. Built with VideoDB.
The missing middle between raw video and reasoning models. Turn video into structured intelligence. Citable, queryable, AI-ready. CLI + MCP server. No Docker. No GPU.
Local-first video intelligence orchestrator using Intel OpenVINO, Tauri, gRPC, SQLite, and agentic AI routing.
VideoMind AI - AI Video Intelligence OS: collection → ASR → AI analysis → reports. Cross-platform desktop app.
Memories — independent third-party profile of a public API surface, by API Evangelist. Memories.ai (MAVI) is a video-intelligence platform that turns raw video into searchable, agent-ready understanding. It ships three products on a shared video-understanding stack: Visual Intelligence (stateless REST APIs for transcription, captioning, frame descr
UnReel is an AI-powered Video Intelligence engine built to decode the context of any short-form content. It is designed for users who encounter language barriers, missed situational context, or struggle to find resources mentioned in a video via a dedicated video analysis pipeline.
High-accuracy video intelligence for object tracking, identity-aware replay, visual search, and interactive scene analysis.
ML and computer-vision based video analysis and virality prediction platform.
BRI — empathetic video intelligence with production Streamlit, FastAPI MCP, SQLite durability, and multimodal ML tooling
To associate your repository with the video-intelligence topic, visit your repo's landing page and select "manage topics."