Introducing ConTextual: How well can your Multimodal model jointly reason over text and image in text-rich scenes?
IgnoreHugging Face · 2024-03-05 00:00 UTC
Not analyzed yet
Eligible for automatic cleanup in 3 day(s) unless marked Must Read.
Content
No content snippet available in the feed.