activity
20222024
most citedVLN-Trans: Translator for the Vision and Language Navigation Agent

1 citations · 2 across the 8 of their papers we have counts for

collaborators

8 papers

cs.CV2024

Defeasible Visual Entailment: Benchmark, Evaluator, and Reward-Driven Optimization

Yue Zhang, Liqiang Jing, Vibhav Gogate

We introduce a new task called Defeasible Visual Entailment (DVE), where the goal is to allow the modification of the entailment relationship between an image premise and a text hy…

cs.CV2024

Narrowing the Gap between Vision and Action in Navigation

Yue Zhang, Parisa Kordjamshidi

The existing methods for Vision and Language Navigation in the Continuous Environment (VLN-CE) commonly incorporate a waypoint predictor to discretize the environment. This simplif…

cs.CL2024

NavHint: Vision and Language Navigation Agent with a Hint Generator

Yue Zhang, Quan Guo, Parisa Kordjamshidi

Existing work on vision and language navigation mainly relies on navigation-related losses to establish the connection between vision and language modalities, neglecting aspects of…

math.GT2023

Guts and The Minimal Volume Orientable Hyperbolic 3-Manifold with 3 Cusps

Yue Zhang

The minimal volume of orientable hyperbolic manifolds with a given number of cusps has been found for cusps, while the minimal volume of 3-cusped orientable hyperbolic ma…

cs.CL20231 cited

VLN-Trans: Translator for the Vision and Language Navigation Agent

Yue Zhang, Parisa Kordjamshidi

Language understanding is essential for the navigation agent to follow instructions. We observe two kinds of issues in the instructions that can make the navigation task challengin…

cs.LG20231 cited

GLUECons: A Generic Benchmark for Learning Under Constraints

Hossein Rajaby Faghihi, Aliakbar Nafar, Chen Zheng +7

Recent research has shown that integrating domain knowledge into deep learning architectures is effective -- it helps reduce the amount of required data, improves the accuracy of t…