3 papers
cs.CV2024
H2OVL-Mississippi Vision Language Models Technical Report
Shaikat Galib, Shanshan Wang, Guanshuo Xu +4
Smaller vision-language models (VLMs) are becoming increasingly important for privacy-focused, on-device applications due to their ability to run efficiently on consumer hardware f…
cs.CL2024
H2O-Danube3 Technical Report
Pascal Pfeiffer, Philipp Singer, Yauhen Babakhin +3
We present H2O-Danube3, a series of small language models consisting of H2O-Danube3-4B, trained on 6T tokens and H2O-Danube3-500M, trained on 4T tokens. Our models are pre-trained…
cs.CL2024
H2O-Danube-1.8B Technical Report
Philipp Singer, Pascal Pfeiffer, Yauhen Babakhin +4
We present H2O-Danube, a series of small 1.8B language models consisting of H2O-Danube-1.8B, trained on 1T tokens, and the incremental improved H2O-Danube2-1.8B trained on an addit…