الباحثون

James Oldfield

المنشورات 2

نسخة أولية وصول مفتوح

Towards a Unified Misuse Monitoring Benchmark

LLM agents increasingly act in multi-actor environments, exposing them to misuse from multiple sources: decomposition attacks, where a harmful request is split into innocuous sub-requests, and prompt injection attacks, where a compromised tool delivers a malicious instruction. Existing evaluations treat these threats s …

نسخة أولية وصول مفتوح

BazaarBench: Delegation Safety in Decentralized C2C Marketplaces Run by LLM Agents

Ziyan Wang, Shuqing Shi, James Oldfield وآخرون · 2026

In decentralized consumer-to-consumer (C2C) marketplaces, people list goods, negotiate with strangers, and rate one another, so trust rests on reputation. Large language model (LLM) agents now act for users, raising risks to their money, privacy, and reputation. We introduce BazaarBench, a simulated C2C marketplace and …

المؤلفون المشاركون