V- Modal AI: MultiModal Video Search - SDK Flutter
-
Updated
Aug 3, 2026 - Dart
V- Modal AI: MultiModal Video Search - SDK Flutter
PyTorch source code for "Stacked Cross Attention for Image-Text Matching" (ECCV 2018)
Meshed-Memory Transformer for Image Captioning. CVPR 2020
TOMM2020 Dual-Path Convolutional Image-Text Embedding with Instance Loss 🐾 https://arxiv.org/abs/1711.05535
Show, Control and Tell: A Framework for Generating Controllable and Grounded Captions. CVPR 2019
Code for "Learning the Best Pooling Strategy for Visual Semantic Embedding", CVPR 2021 (Oral)
Video Search SDK for Android Kotlin. Integrate in any video app
(RSS 2018) LoST - Visual Place Recognition using Visual Semantics for Opposite Viewpoints across Day and Night
PyTorch library for Visual-Semantic tasks
Image-Text Matching Model Zoo
🔍 Discover and scan vulnerable Next.js instances across your infrastructure to address critical security threats effectively.
SDK for visual search on Images, Videos,
📊 Record LLM API costs and token usage accurately in CI/CD, local development, and production. Monitor expenses without estimation for better cost control.
Add a description, image, and links to the visual-semantic topic page so that developers can more easily learn about it.
To associate your repository with the visual-semantic topic, visit your repo's landing page and select "manage topics."