◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Markus Wulfmeier

DeepMInd

13 papers hereh-index 253.1k citations63 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author10
  • last author1

Across the 12 of 13 papers where every author was matched, so the position is known.

fields
  • cs.RO6
  • cs.LG4
  • astro-ph.IM1
  • cs.AI1
  • cs.CL1
affiliations
  • DeepMInd
same name
  • Markus Wulfmeier — 3 papers, h 1

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedImproving cosmological reach of a gravitational wave observatory using Deep Loop Shaping

8 citations · 8 across the 5 of their papers we have counts for

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2025

LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities

Thomas Schmied, Jörg Bornschein, Jordi Grau-Moya +2

The success of Large Language Models (LLMs) has sparked interest in various agentic applications. A key hypothesis is that LLMs, leveraging common sense and Chain-of-Thought (CoT)…

cs.LG2024

Imitating Language via Scalable Inverse Reinforcement Learning

Markus Wulfmeier, Michael Bloesch, Nino Vieillard +13

The majority of language model training builds on imitation learning. It covers pretraining, supervised fine-tuning, and affects the starting conditions for reinforcement learning…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.