Artificial Analysis open-sources AA-AgentPerf-Local for laptop and workstation inference
Artificial Analysis has open-sourced AA-AgentPerf-Local, which replays recorded agent sessions to measure local inference speed, and published initial 4-bit results for four systems. On models that fit in 32 GB, the GeForce RTX 5090 finished more than 3.5 times faster than the other systems tested.