Confidence Estimation for Financial Vision-Language Models in Chart and Document Understanding
Source
Reza Khanmohammadi, Simerjot Kaur, Charese H. Smiley, Ivan Brugere, Mohammad M. Ghassemi
Author
Reza Khanmohammadi, Simerjot Kaur, Charese H. Smiley, Ivan Brugere, Mohammad M. Ghassemi
Date
Terms in this piece · Glossary
LLM — A large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.
open weights — A model whose trained parameters are published for anyone to download and run — unlike API-only models you can access but never possess.
Why it matters
If you gate document- or chart-reading model output on a confidence score, this shows that off-the-shelf confidence signals are not thresholdable and that reliability must be measured per model and per task.