Native-video memory for vision-language-action models, using timestamped visual history and exact streaming inference for long-horizon robot manipulation.
-
Updated
Sep 24, 2026 - Python
Native-video memory for vision-language-action models, using timestamped visual history and exact streaming inference for long-horizon robot manipulation.
MEMO, a visual memory assistant that prioritizes useful uncertainty over confident guesses.
Codex插件,能将图片沉淀为可复用的视觉记忆,助你持续创作出美学风格一致的图片。The Codex plugin can turn images into reusable visual memories, helping you consistently create images with a unified aesthetic style.
Compare code reviews across models in Claude Code with ranked todos and subagent dispatch via MCP.
Natural-language robot navigation with multi-agent planning, persistent visual memory, and visual goal verification. Built for ROS 2 and Nav2.
DINOv2-based fixed-budget video memory selection through temporal representation dynamics.
AI visual memory, visual second brain, video memory for LLMs, screen recording for AI, local-first AI memory, visual RAG, deterministic visual memory, seeded source spine, cryptographic video archive, multimodal memory, frame sampling for AI, AI module attachments, and reproducible screen memory.
Ctrl+F for the physical world. A camera on a SiMa.ai Modalix that remembers everything in a room and answers out loud: where is it, who took it, what happened. Segmentation + pose + tracking + a vision-language model on one sub-10 W chip. No cloud.
A private AI memory for your real life. Continuously capture and remember what you see and hear using local VLMs.
To associate your repository with the visual-memory topic, visit your repo's landing page and select "manage topics."