◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

R. Hoory

6 papers hereh-index 222.3k citations62 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author5
  • last author1

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • eess.AS4
  • cs.CL2

identity via Semantic Scholar / OpenAlex

collaborators
Showing eess.ASShow all

4 papers · 1 filter

eess.AS2026

Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS

Hagai Aronowitz, Zvi Kons, Avihu Dekel +2

Speaker-Attributed Automatic Speech Recognition (SAA) enhances traditional ASR systems by incorporating relative speaker identity tags directly into the transcript (e.g., [Speaker…

eess.AS2025

Speech Synthesis From Continuous Features Using Per-Token Latent Diffusion

Arnon Turetzky, Avihu Dekel, Nimrod Shabtay +5

We present SALAD, a zero-shot TTS autoregressive model operating over continuous speech representations. SALAD utilizes a per-token diffusion process to refine and predict continuo…

eess.AS2025

Spoken question answering for visual queries

Nimrod Shabtay, Zvi Kons, Avihu Dekel +3

Question answering (QA) systems are designed to answer natural language questions. Visual QA (VQA) and Spoken QA (SQA) systems extend the textual QA system to accept visual and spo…

eess.AS2025

Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities

George Saon, Avihu Dekel, Alexander Brooks +21

Granite-speech LLMs are compact and efficient speech language models specifically designed for English ASR and automatic speech translation (AST). The models were trained by modali…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.