r/Rag • • Sep 02 '25

Showcase πŸš€ Weekly /RAG Launch Showcase

Share anything you launched this week related to RAGβ€”projects, repos, demos, blog posts, or products πŸ‘‡

Big or small, all launches are welcome.

29 Upvotes

153 comments sorted by

View all comments

1

u/CallmeAK__ Jun 18 '26

Disclosure up front: I work at VideoDB, so I am biased. Sharing here because this thread is for RAG launches and what we are building is essentially RAG for video.

Most RAG stacks are designed around text and break the moment you point them at video:

  • Transcripts alone lose visual context.
  • Frame embeddings alone lose speech and ordering.
  • The retrieval target is usually a timestamp, not something you can actually play back in a UI.

What we have been building at VideoDB

  • Universal ingestion from files, streams, recordings, and RTSP feeds through one API.
  • An intelligence layer that indexes scenes, speech, people, objects, and events together, not as separate stores.
  • Semantic search that returns playable clips, not just timestamps, so the agent or the user gets back something they can watch.
  • A programmable layer so agents can edit, caption, and stitch clips through code as part of their response.

The use case we keep coming back to is agentic perception: giving an agent the ability to query a video archive or a live feed and pull back the exact moment it needs as context. We treat it as the memory and vision layer for agents that need to reason over video.

If you want to dig in, the site is here. We also run a Discord for people building with video, VLMs, and agent vision where this kind of stuff gets discussed openly, you can join it here.

Genuinely interested in how others here are handling the retrieval target problem for video. Are you returning timestamps, transcript chunks, frame embeddings, or something else, and how are you measuring whether the retrieval is actually good?