Projects
OmniFind
In progressSemantic search and chat over your entire filesystem. A multimodal embedding model indexes every file — images, video, docs — so you can find anything by what's inside it. Spotlight, but it actually understands your files.
An autonomous agent that does the tedious 80% of a video edit on a shared timeline — you keep taste and polish the rest by hand.
Turns a handful of product images plus a prompt into a finished marketing video — analysis, scene planning, parallel frame generation, QA, and render.
Chat with your database in plain English. Natural-language-to-SQL with auto-generated visualizations and saved query presets.
Temporally-aware video embeddings
ResearchResearch: a temporally-aware video embedding model powering an agentic text-to-video evaluation benchmark — better correlated with human preference than FVD and CLIPScore for current generators.