Introducing ConTextual: How well can your Multimodal model jointly reason over text and image in text-rich scenes?T1Covered by 1 source · first reported by Hugging Face Blog at 05 Mar 2024, 00:00 UTCSourcesT1Introducing ConTextual: How well can your Multimodal model jointly reason over text and image in text-rich scenes?Hugging Face Blog05 Mar 2024, 00:00