Back to Entity Graph

👤 Jiawei Zhou Person

1 articles First seen: Jul 13, 2026 Last seen: Jul 13
Activity Timeline (90 days)
Co-occurring Entities
Articles (1)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
The paper explores whether long text can be compressed by converting it into an image and feeding that image to a multimodal LLM instead of sending all text tokens directly. The key idea is that visio
🏠Portal 📰Links Q&A 📅Events 💼Jobs