2 papers
cs.CV2026
The Gate Always Closes: On Injecting Auxiliary Signals into Frozen Vision-Language Models
Moshiur Farazi, Sameera Ramasinghe, Bekir Sait Ciftler +2
Auxiliary signal pathways in VLMs are routinely fitted with learnable gates so the optimiser can decide how much of the signal to admit. We find that the optimiser almost always de…
cs.CV2026
HyperVis: Continuous Latent Visual Relational Graphs on the Lorentz Hyperboloid for Compositional Reasoning
Moshiur Farazi, Sameera Ramasinghe, Mahbub Ahmed Turza +1
Vision-Language Models (VLMs) struggle with compositional reasoning that requires understanding inter-object relationships. A natural remedy is to inject explicit scene graph tripl…