most citedUnified Lattice Graph Fusion for Chinese Named Entity Recognition

2 citations · 3 across the 2 of their papers we have counts for

collaborators

6 papers

cs.CL2024

Taiyi-Diffusion-XL: Advancing Bilingual Text-to-Image Generation with Large Vision-Language Model Support

Xiaojun Wu, Dixiang Zhang, Ruyi Gan +6

Recent advancements in text-to-image models have significantly enhanced image generation capabilities, yet a notable gap of open-source models persists in bilingual or Chinese lang…

cs.CL20232 cited

Unified Lattice Graph Fusion for Chinese Named Entity Recognition

Dixiang Zhang, Junyu Lu, Pingjian Zhang

Integrating lexicon into character-level sequence has been proven effective to leverage word boundary and semantic information in Chinese named entity recognition (NER). However, p…

cs.CV2023

iDesigner: A High-Resolution and Complex-Prompt Following Text-to-Image Diffusion Model for Interior Design

Ruyi Gan, Xiaojun Wu, Junyu Lu +8

With the open-sourcing of text-to-image models (T2I) such as stable diffusion (SD) and stable diffusion XL (SD-XL), there is an influx of models fine-tuned in specific domains base…

cs.CL2023

Lyrics: Boosting Fine-grained Language-Vision Alignment and Comprehension via Semantic-aware Visual Objects

Junyu Lu, Dixiang Zhang, Songxin Zhang +6

Large Vision Language Models (LVLMs) have demonstrated impressive zero-shot capabilities in various vision-language dialogue scenarios. However, the absence of fine-grained visual…

cs.CL2023

Ziya2: Data-centric Learning is All LLMs Need

Ruyi Gan, Ziwei Wu, Renliang Sun +11

Various large language models (LLMs) have been proposed in recent years, including closed- and open-source ones, continually setting new records on multiple benchmarks. However, th…

cs.CL20231 cited

Ziya-Visual: Bilingual Large Vision-Language Model via Multi-Task Instruction Tuning

Junyu Lu, Dixiang Zhang, Xiaojun Wu +5

Recent advancements enlarge the capabilities of large language models (LLMs) in zero-shot image-to-text generation and understanding by integrating multi-modal inputs. However, suc…