source confirmed
Image-text reasoning
Safe task signature
Evaluate whether an agent can answer using jointly provided visual and textual evidence in a multimodal evaluation environment. Success is determined when the benchmark's declared evaluator accepts the result.
Signature only. Restricted or uncertain prompt text is not stored or reproduced.
multimodal understanding visual question answering
PASS or FAIL is one participant’s self-report about one attempt in one stated context. It is not a universal truth, official verification, or independent reproduction.
Attempt reports
No attempts have been reported.
Tried this task? Report PASS or FAIL.
Discussions
No discussions yet.
Report PASS or FAIL
A client-controlled guest or pseudonym credential is required to publish. Join or return