Search CORE

7,195 research outputs found

A General Generalization of Jordan's Inequality and a Refinement of L. Yang's Inequality

Author: Cao Jian
Niu Da-Wei
Qi Feng
Publication venue: School of Communications and Informatics, Faculty of Engineering and Science, Victoria University of Technology
Publication date: 01/01/2007
Field of study

Bis[aqua(2,3-naphtho-15-crown-5)sodium] tetrakis(thiocyanato-κN)cobaltate(II)

Author: Li Chengjuan
Li Da-Cheng
Wang Da-Qi
Publication venue: International Union of Crystallography
Publication date: 01/11/2009
Field of study

The title complex, [Na(C18H22O5)(H2O)]2[Co(NCS)4], consists of two aqua(2,3-naphtho-15-crown-5)sodium complex cations and one [Co(NCS)4]2− complex anion, which has crystallographic symmetry. In the anion, the CoII centre is coordinated by the N atoms of four NCS− ligands in a distorted tetrahedral geometry. In the complex cations, the NaI centre is coordinated by five O atoms of the 2,3-naphtho-15-crown-5 ligand and one water O atom. The complex molecules form a two-dimensional network via weak O—H⋯S interactions between adjacent cations and anion

Crossref

Directory of Open Access Journals

PubMed Central

Prompt Switch: Efficient CLIP Adaptation for Text-Video Retrieval

Author: Chen Da
Chen Qi
Deng Chaorui
Qin Pengda
Wu Qi
Publication venue
Publication date: 15/08/2023
Field of study

In text-video retrieval, recent works have benefited from the powerful learning capabilities of pre-trained text-image foundation models (e.g., CLIP) by adapting them to the video domain. A critical problem for them is how to effectively capture the rich semantics inside the video using the image encoder of CLIP. To tackle this, state-of-the-art methods adopt complex cross-modal modeling techniques to fuse the text information into video frame representations, which, however, incurs severe efficiency issues in large-scale retrieval systems as the video representations must be recomputed online for every text query. In this paper, we discard this problematic cross-modal fusion process and aim to learn semantically-enhanced representations purely from the video, so that the video representations can be computed offline and reused for different texts. Concretely, we first introduce a spatial-temporal "Prompt Cube" into the CLIP image encoder and iteratively switch it within the encoder layers to efficiently incorporate the global video semantics into frame representations. We then propose to apply an auxiliary video captioning objective to train the frame representations, which facilitates the learning of detailed video semantics by providing fine-grained guidance in the semantic space. With a naive temporal fusion strategy (i.e., mean-pooling) on the enhanced frame representations, we obtain state-of-the-art performances on three benchmark datasets, i.e., MSR-VTT, MSVD, and LSMDC.Comment: to be appeared in ICCV202

arXiv.org e-Print Archive

Bis(acetato-κ2 O,O′)bis(2-aminopyridine-κN)nickel(II)

Author: Da-Qi Wang
Qiang Wang
Roman
Publication venue: International Union of Crystallography
Publication date: 01/01/2008
Field of study

The title complex, [Ni(C2H3O2)2(C5H6N2)2], has a distorted octahedral geometry around the Ni atom. Intermolecular and intramolecular N—H⋯O hydrogen bonds exist in the crystal structure

Crossref

Directory of Open Access Journals

PubMed Central