VAmoS Bench evaluates voice agents through stateful, end-to-end customer support simulations using a seeded PostgreSQL backend. It moves beyond component metrics like WER to measure task containment, determining if agents can resolve complex goals like card cancellations without human intervention.
HOW THIS AFFECTS YOU
●
builderYou can use this to test the actual business utility and task completion rates of your voice bots.
●
founderThis allows you to benchmark your product against real-world containment metrics used by contact centers.