activity
20222026
most citedOpenThoughts: Data Recipes for Reasoning Models

1 citations · 1 across the 5 of their papers we have counts for

collaborators

6 papers

cs.SE2026

Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces

Mike A. Merrill, Alexander G. Shaw, Nicholas Carlini +82

AI agents may soon become capable of autonomously completing valuable, long-horizon tasks in diverse domains. Current benchmarks either do not measure real-world tasks, or are not…

cs.DC2025

Learned Cost Model for Placement on Reconfigurable Dataflow Hardware

Etash Guha, Tianxiao Jiang, Andrew Deng +2

Mapping a dataflow-graph of an ML model onto a reconfigurable system is difficult, as different mappings have different throughputs and consume resource constraints differently. To…

cs.LG20251 cited

OpenThoughts: Data Recipes for Reasoning Models

Etash Guha, Ryan Marten, Sedrick Keh +47

Reasoning models have made rapid progress on many benchmarks involving math, code, and science. Yet, there are still many open questions about the best training recipes for reasoni…

cs.CV2025

Synthetic Document Question Answering in Hungarian

Jonathan Li, Zoltan Csaki, Nidhi Hiremath +4

Modern VLMs have achieved near-saturation accuracy in English document visual question-answering (VQA). However, this task remains challenging in lower resource languages due to a…

cs.CV2024

BLIP3-KALE: Knowledge Augmented Large-Scale Dense Captions

Anas Awadalla, Le Xue, Manli Shu +13

We introduce BLIP3-KALE, a dataset of 218 million image-text pairs that bridges the gap between descriptive synthetic captions and factual web-scale alt-text. KALE augments synthet…

cs.LG2022

On Accelerated Perceptrons and Beyond

Guanghui Wang, Rafael Hanashiro, Etash Guha +1

The classical Perceptron algorithm of Rosenblatt can be used to find a linear threshold function to correctly classify linearly separable data points, assuming the classes are…