◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

P. Parrilo

13 papers hereh-index 6225.3k citations284 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author8
  • last author5

Across the 13 of 13 papers where every author was matched, so the position is known.

fields
  • cs.RO6
  • cs.LG3
  • math.OC3
  • cs.AI1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2026

Spectral Scaling Laws of Muon

Gagik Magakyan, Pablo Parrilo, Asuman Ozdaglar

Orthonormalized update rules have rapidly become a leading choice of optimizer for training large language models, with recent open-source state-of-the-art models adopting Muon. To…

cs.LG2026

Collaborative and Efficient Fine-tuning: Leveraging Task Similarity

Gagik Magakyan, Amirhossein Reisizadeh, Chanwoo Park +2

Adaptability has been regarded as a central feature in the foundation models, enabling them to effectively acclimate to unseen downstream tasks. Parameter-efficient fine-tuning met…

cs.LG2024

A Unified Linear Programming Framework for Offline Reward Learning from Human Demonstrations and Feedback

Kihyun Kim, Jiawei Zhang, Asuman Ozdaglar +1

Inverse Reinforcement Learning (IRL) and Reinforcement Learning from Human Feedback (RLHF) are pivotal methodologies in reward learning, which involve inferring and shaping the und…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.