papersSEP 10 04:00 UTC
LogiScope-VQA: A Benchmark for Vision-Language Models on Warehouse Hazard Detection
Researchers have released LogiScope-VQA, a benchmark that evaluates whether large multimodal models can perceive, understand, and reason about safety hazards in industrial warehouse environments at a level comparable to human experts. The work addresses the lack of domain-specific evaluation data for deploying such models in logistics settings. The paper appears on arXiv with cross-listings in artificial intelligence and computational linguistics.