연구
Token Arena: A Continuous Benchmark Unifying Energy and Cognition in AI Inference
arXiv:2605.00300v1 Announce Type: new Abstract: Public inference benchmarks compare AI systems at the model and provider level, but the unit at which deployment decisions are actually made is the endpoint: the provider, model, stockkeepingunit tuple at which a specific quantization, decoding...
arXiv:2605.00300v1 Announce Type: new Abstract: Public inference benchmarks compare AI systems at the model and provider level, but the unit at which deployment decisions are actually made is the endpoint: the provider, model, stockkeepingunit tuple at which a specific quantization, decoding strategy, region, and serving stack is exposed.
이 콘텐츠는 ArXiv AI 원본 기사의 요약입니다. 전문은 원본 사이트에서 확인해주세요.
원문 기사 보기 →