2 papers
cs.LG2025
Where and How to Enhance: Discovering Bit-Width Contribution for Mixed Precision Quantization
Haidong Kang, Lianbo Ma, Guo Yu +1
Mixed precision quantization (MPQ) is an effective quantization approach to achieve accuracy-complexity trade-off of neural network, through assigning different bit-widths to netwo…
cs.CL2024
A Cascade Dual-Decoder Model for Joint Entity and Relation Extraction
Jian Cheng, Tian Zhang, Shuang Zhang +5
In knowledge graph construction, a challenging issue is how to extract complex (e.g., overlapping) entities and relationships from a small amount of unstructured historical data. T…