◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Mantas Mazeika

19 papers hereh-index 2822.4k citations55 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author9
  • last author3

Across the 15 of 19 papers where every author was matched, so the position is known.

fields
  • cs.LG9
  • cs.AI3
  • cs.CL3
  • cs.CR2
  • cs.CY2

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedHumanity's Last Exam

18 citations · 24 across the 8 of their papers we have counts for

collaborators
Showing cs.AIShow all

3 papers · 1 filter

cs.AI2026

SAE-StatSteer: Statistical Consensus Feature Selection for Optimization-Free Activation Steering of Large Language Models

Oshayer Siddique, J. M Areeb Uzair Alam, Md Jobayer Rahman Rafy +3

Activation steering adds a residual-stream direction at inference time, providing lightweight behavioral control without fine-tuning. Sparse autoencoders (SAEs) can make such inter…

cs.AI2025★ 3 cited

A Definition of AGI

Dan Hendrycks, Dawn Song, Christian Szegedy +30

The lack of a concrete definition for Artificial General Intelligence (AGI) obscures the gap between today's specialized AI and human-level cognition. This paper introduces a quant…

cs.AI2025

TextQuests: How Good are LLMs at Text-Based Video Games?

Long Phan, Mantas Mazeika, Andy Zou +1

Evaluating AI agents within complex, interactive environments that mirror real-world challenges is critical for understanding their practical capabilities. While existing agent ben…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.