HyperVL is a small but smart model that understands images and text, designed to run fast on phones and tablets.
InfiniteVL is a vision-language model that mixes two ideas: local focus with Sliding Window Attention and long-term memory with a linear module called Gated DeltaNet.