Vision-only encoders for language representation directly in pixel space.
-
On the Design Fundamentals of Pixel Text Representation Learning
Paper • 2609.01147 • Published • 24 -
Pixel-Linguist/Pixel-Linguist-II-Pretrain
Feature Extraction • 0.7B • Updated -
Pixel-Linguist/Pixel-Linguist-II-Midtrain
Feature Extraction • 0.7B • Updated -
Pixel-Linguist/Pixel-Linguist-II-Midtrain-Only
Feature Extraction • 0.7B • Updated