◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Philip J. B. Jackson

Centre for Vision, Speech & Signal Processing (CVSSP), University of Surrey

5 papers hereh-index 339 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author4
  • last author1

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.SD4
  • eess.AS1
affiliations
  • Centre for Vision, Speech & Signal Processing (CVSSP), University of Surrey
HomepageORCID 0000-0001-7933-5935
same name
  • Philip J. B. Jackson — 8 papers, h 5
  • Philip J. B. Jackson — 1 paper, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.SDShow all

4 papers · 1 filter

cs.SD2026

Listen, Reason, and Segment: Aligning LALMs with Editorial Judgment for Media Chapterization

Tony Alex, Wish Suharitdamrong, Sara Atito +5

Large Audio Language Models (LALMs) have made rapid progress on standardized benchmarks, yet their deployment in practical media workflows, curation, archival indexing, and content…

cs.SD2026

Hierarchical Activity Recognition and Captioning from Long-Form Audio

Peng Zhang, Qingyu Luo, Philip J. B. Jackson +1

Complex activities in real-world audio unfold over extended durations and exhibit hierarchical structure, yet most prior work focuses on short clips and isolated events. To bridge…

cs.SD2026

PAL: Probing Audio Encoders via LLMs -- Audio Information Transfer into LLMs

Tony Alex, Wish Suharitdamrong, Sara Atito +4

Integration of audio perception into large language models (LLMs) is an emerging research area for enabling machine listening applications, yet efficient transfer of rich audio sem…

cs.SD2025

SSLAM: Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes

Tony Alex, Sara Ahmed, Armin Mustafa +2

Self-supervised pre-trained audio networks have seen widespread adoption in real-world systems, particularly in multi-modal large language models. These networks are often employed…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.