Rumi Labs logo

Rumi Labs

Rumi Labs is building a decentralized Vision-Language Model (VLM) network that powers interactive media and contextual advertising. It focuses on helping AI understand human stories by analyzing millions of hours of IP media content that centralized AI labs cannot access, targeting media companies, advertisers, and content platforms.
Distributed

Description

Rumi Labs is developing a decentralized network of Vision-Language Models (VLMs) designed to decode narrative and emotional context within video and IP media content — a domain largely inaccessible to centralized AI labs due to IP constraints. Their proprietary VLM architecture, powered by a 'Narrative Intelligence' Mixture-of-Experts (MoE) approach, claims to surpass models like Gemini 2.5 Pro on narrative comprehension tasks while being significantly smaller. The company's decentralized infrastructure enables compliant, cost-effective indexing of media content in real time, gaining deeper insight into how consumers interact with media. Rumi Labs also provides APIs and SDKs that let AI identify, understand, and enrich video content, enabling interactive, personalizable, and shoppable media experiences. Their offerings are aimed at powering a new era of contextual advertising, on-device IP content customization, dynamic interactions, and viewership analytics, serving media companies and advertisers as primary clients.

Grant Funding

VC Funding

None
2025

$0

$4.7M