search
vision-language models
Trends
- 1Apple Releases LensVLM-9B Model for Compressed Documents●Apple Releases LensVLM-9B, the Model That Reads Compressed Documents
Apple has released LensVLM-9B, an artificial intelligence model designed to read and understand compressed documents. The release suggests Apple is expanding its work in document-focused vision-language systems, and details about the model's capabilities and availability are now circulating among AI watchers. Further information on benchmarks, licensing and intended use was not immediately available.
- 2New Benchmark Tests Memory of Vision-Language Robots▼New Benchmark Measures Memory Capabilities of VLM-Powered Robots | Newswise
Researchers have introduced a new benchmark designed to measure how well robots powered by vision-language models retain and use memory across tasks. The tool aims to standardize evaluation of robotic memory, a key weakness in current systems, and could help developers compare progress and push robots toward more reliable real-world behaviour.