الباحثون

Adel Bibi

المنشورات 3

نسخة أولية وصول مفتوح

Towards a Unified Misuse Monitoring Benchmark

LLM agents increasingly act in multi-actor environments, exposing them to misuse from multiple sources: decomposition attacks, where a harmful request is split into innocuous sub-requests, and prompt injection attacks, where a compromised tool delivers a malicious instruction. Existing evaluations treat these threats s …

نسخة أولية وصول مفتوح

BazaarBench: Delegation Safety in Decentralized C2C Marketplaces Run by LLM Agents

Ziyan Wang, Shuqing Shi, James Oldfield وآخرون · 2026

In decentralized consumer-to-consumer (C2C) marketplaces, people list goods, negotiate with strangers, and rate one another, so trust rests on reputation. Large language model (LLM) agents now act for users, raising risks to their money, privacy, and reputation. We introduce BazaarBench, a simulated C2C marketplace and …

نسخة أولية وصول مفتوح

Towards Better Exploration in Sequential Test-Time Scaling

Joseph Rance, Fabio Pizzati, Juil Sock وآخرون · 2026

Test-time scaling improves language model reasoning by spending additional compute at inference. However, both classes of existing methods often fail to continue improving over long timescales. Parallel methods repeatedly sample independent answers from the model, scaling poorly on problems the model is unlikely to sol …

المؤلفون المشاركون