Search CORE

7 research outputs found

PaLI-3 Vision Language Models: Smaller, Faster, Stronger

Author: Alabdulmohsin Ibrahim
Beyer Lucas
Chen Xi
Goodman Sebastian
Keysers Daniel
Kolesnikov Alexander
Mustafa Basil
Padlewski Piotr
Pavetic Filip
Rong Keran
Salz Daniel
Soricut Radu
Vlasic Daniel
Voigtlaender Paul
Wang Xiao
Wu Jialin
Xiong Xi
Yu Tianli
Zhai Xiaohua
Publication venue
Publication date: 17/10/2023
Field of study

This paper presents PaLI-3, a smaller, faster, and stronger vision language model (VLM) that compares favorably to similar models that are 10x larger. As part of arriving at this strong performance, we compare Vision Transformer (ViT) models pretrained using classification objectives to contrastively (SigLIP) pretrained ones. We find that, while slightly underperforming on standard image classification benchmarks, SigLIP-based PaLI shows superior performance across various multimodal benchmarks, especially on localization and visually-situated text understanding. We scale the SigLIP image encoder up to 2 billion parameters, and achieves a new state-of-the-art on multilingual cross-modal retrieval. We hope that PaLI-3, at only 5B parameters, rekindles research on fundamental pieces of complex VLMs, and could fuel a new generation of scaled-up models

arXiv.org e-Print Archive

PaLI: A Jointly-Scaled Multilingual Language-Image Model

Effective scaling and a flexible task interface enable large language models to excel at many tasks. PaLI (Pathways Language and Image model) extends this approach to the joint modeling of language and vision. PaLI generates text based on visual and textual inputs, and with this interface performs many vision, language, and multimodal tasks, in many languages. To train PaLI, we make use of large pretrained encoder-decoder language models and Vision Transformers (ViTs). This allows us to capitalize on their existing capabilities and leverage the substantial cost of training them. We find that joint scaling of the vision and language components is important. Since existing Transformers for language are much larger than their vision counterparts, we train the largest ViT to date (ViT-e) to quantify the benefits from even larger-capacity vision models. To train PaLI, we create a large multilingual mix of pretraining tasks, based on a new image-text training set containing 10B images and texts in over 100 languages. PaLI achieves state-of-the-art in multiple vision and language tasks (such as captioning, visual question-answering, scene-text understanding), while retaining a simple, modular, and scalable design

arXiv.org e-Print Archive

PaLI-X: On Scaling up a Multilingual Vision and Language Model

We present the training recipe and results of scaling up PaLI-X, a multilingual vision and language model, both in terms of size of the components and the breadth of its training task mixture. Our model achieves new levels of performance on a wide-range of varied and complex tasks, including multiple image-based captioning and question-answering tasks, image-based document understanding and few-shot (in-context) learning, as well as object detection, video question answering, and video captioning. PaLI-X advances the state-of-the-art on most vision-and-language benchmarks considered (25+ of them). Finally, we observe emerging capabilities, such as complex counting and multilingual object detection, tasks that are not explicitly in the training mix

arXiv.org e-Print Archive

Ca2+ sensor-mediated ROS scavenging suppresses rice immunity and is exploited by a fungal effector

Author: Chang Huizhong
Chen Jin
Deng Yiwen
Gao Mingjun
Gong Xiangyu
He Yang
He Zuhua
Huang Yifeng
Li Xiaoyuan
Liu Jiyun
Tharreau Didier
Wang Ertao
Wang Guo-Liang
Wu Yue
Xie Shenghan
Xu Jianlong
Yan Bingxiao
Yang Weibing
Yin Xin
Yue Jiaxing
Zhai Keran
Zhang Guiquan
Zhong Xiangbin
Publication venue: 'Elsevier BV'
Publication date: 01/01/2021
Field of study

Plant immunity is activated upon pathogen perception and often affects growth and yield when it is constitutively active. How plants fine-tune immune homeostasis in their natural habitats remains elusive. Here, we discover a conserved immune suppression network in cereals that orchestrates immune homeostasis, centering on a Ca2+-sensor, RESISTANCE OF RICE TO DISEASES1 (ROD1). ROD1 promotes reactive oxygen species (ROS) scavenging by stimulating catalase activity, and its protein stability is regulated by ubiquitination. ROD1 disruption confers resistance to multiple pathogens, whereas a natural ROD1 allele prevalent in indica rice with agroecology-specific distribution enhances resistance without yield penalty. The fungal effector AvrPiz-t structurally mimics ROD1 and activates the same ROS-scavenging cascade to suppress host immunity and promote virulence. We thus reveal a molecular framework adopted by both host and pathogen that integrates Ca2+ sensing and ROS homeostasis to suppress plant immunity, suggesting a principle for breeding disease-resistant, high-yield crops

Agritrop

HAL-IRD

HAL-CIRAD

A nucleotide-binding site-leucine-rich repeat receptor pair confers broad-spectrum disease resistance through physical association in rice

Author: Bingxiao Yan
Jianyao Shou
Jiyun Liu
Jun Tang
Keran Zhai
Luo M
Meizhong Luo
Qun Li
Sueldo DJ
Xin Wang
Yiwen Deng
Zhen Xie
Zuhua He
Publication venue: 'The Royal Society'
Publication date
Field of study

Crossref

An E3 Ubiquitin Ligase-BAG Protein Module Controls Plant Innate Immunity and Broad-Spectrum Disease Resistance

Author: Alcázar
Bai
Balbi
Bhoj
Bruggeman
Büschges
Cheng
Cheng
Coll
Deng
Dietrich
Donglei Yang
Doukhanina
Duplan
Duttler
Ghag
Gou
Gray
Hoang
Huang
Huang
Huang
Ishikawa
Jianjun Wang
Jingni Wu
Jiyun Liu
Jones
Junzhong Liu
Kabbage
Kaku
Kang
Kawasaki
Keran Zhai
Li
Li
Li
Li
Li
Liu
Liu
Lorrain
Lu
Marino
Marino
Mei
Park
Qi
Qi Xie
Quanyuan You
Qun Li
Robert-Seilaniantz
Serrano
Singh
Spoel
Stuttmann
Takayama
Takayama
Tanaka
Tian
Tong
Vierstra
Wang
Wang
Weibing Yang
Wenbo Pan
Xu
Xudong Zhu
Yamanouchi
Yang
Yang
Yang
Yikun Jian
Yingying Zhang
Yiwen Deng
Yonggen Lou
Yoo
Zebell
Zeng
Zhang
Zhang
Zuhua He
Publication venue: 'Elsevier BV'
Publication date
Field of study

Crossref