2 papers
cs.AI2025
UpBench: A Dynamically Evolving Real-World Labor-Market Agentic Benchmark Framework Built for Human-Centric AI
Darvin Yi, Teng Liu, Mattie Terzolo +4
As large language model (LLM) agents increasingly undertake digital work, reliable frameworks are needed to evaluate their real-world competence, adaptability, and capacity for hum…
cs.LG2025
GraphMatch: Fusing Language and Graph Representations in a Dynamic Two-Sided Work Marketplace
MikoÅaj Sacha, Hammad Jafri, Mattie Terzolo +2
Recommending matches in a text-rich, dynamic two-sided marketplace presents unique challenges due to evolving content and interaction graphs. We introduce GraphMatch, a new large-s…