VLMsAreBlind reasoning A vision-language benchmark that probes blind spots and brittle reasoning in multimodal models. Leaderboard Showing 4 of 4 results Export CSV Graph Rank Qwen3.5-35B-A3B 97.0% iQwen3.6-27B 97.0% iQwen3.5-27B 96.9% iQwen3.5-122B-A10B 96.7% i