3 papers
cs.AI2026
Best-of-Q: Improving VLM agents with Q-function Action Ranking at Inference
Emilien Biré, María Santos, Kai Yuan
Vision-Language Models (VLMs) have become powerful backbones for agents to autonomously operate in digital environments like the web and operating systems. However, these models su…
cs.CL2025
Training Report of TeleChat3-MoE
Xinzhang Liu, Chao Wang, Zhihao Yang +51
TeleChat3-MoE is the latest series of TeleChat large language models, featuring a Mixture-of-Experts (MoE) architecture with parameter counts ranging from 105 billion to over one t…
cs.AI2025
Surfer 2: The Next Generation of Cross-Platform Computer Use Agents
Mathieu Andreux, Märt Bakler, Yanael Barbier +50
Building agents that generalize across web, desktop, and mobile environments remains an open challenge, as prior systems rely on environment-specific interfaces that limit cross-pl…