◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Purbesh Mitra

4 papers hereh-index 6138 citations12 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL3
  • cs.IT1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators

4 papers

cs.CL2026

Chained Recursive Language Models for Multi-Iteration Reasoning

Purbesh Mitra, Sennur Ulukus

Long context reasoning in large language models (LLMs) is usually constrained by the fact that a single inference trajectory has to simultaneously explore the context, store interm…

cs.CL2025

Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning

Purbesh Mitra, Sennur Ulukus

Long context reasoning in large language models (LLMs) has demonstrated enhancement of their cognitive capabilities via chain-of-thought (CoT) inference. Training such models is us…

cs.CL2025

MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs

Purbesh Mitra, Sennur Ulukus

Recent advancements in the reasoning capabilities of large language models (LLMs) show that employing group relative policy optimization (GRPO) algorithm for reinforcement learning…

cs.IT2024

Distributed Mixture-of-Agents for Edge Inference with Large Language Models

Purbesh Mitra, Priyanka Kaswan, Sennur Ulukus

Mixture-of-Agents (MoA) has recently been proposed as a method to enhance performance of large language models (LLMs), enabling multiple individual LLMs to work together for collab…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.