Search CORE

2 research outputs found

PaLI-X: On Scaling up a Multilingual Vision and Language Model

We present the training recipe and results of scaling up PaLI-X, a multilingual vision and language model, both in terms of size of the components and the breadth of its training task mixture. Our model achieves new levels of performance on a wide-range of varied and complex tasks, including multiple image-based captioning and question-answering tasks, image-based document understanding and few-shot (in-context) learning, as well as object detection, video question answering, and video captioning. PaLI-X advances the state-of-the-art on most vision-and-language benchmarks considered (25+ of them). Finally, we observe emerging capabilities, such as complex counting and multilingual object detection, tasks that are not explicitly in the training mix

arXiv.org e-Print Archive

Elucidating the function of the gene adducing 3 (gamma) in T-cell acute lymphoblastic leukemia.

Author: Montgomery Ceslee D
Publication venue: The Mouseion at the JAXlibrary
Publication date: 07/01/2009
Field of study

The Jackson Laboratory: The Mouseion at the JAXlibrary