Skip to main content
Question

GeMM benchmark

  • September 28, 2026
  • 0 replies
  • 4 views

Hi,

I am running voyager-sdk v1.8 and trying to compile a single GeMM model for 4 AIPU cores so that all 4 cores cooperate on a single inference. According to the CompilerConfig reference, MulticoreMode lists cooperative as a valid value, described as cores cooperating on a single inference.

When I set multicore_mode to cooperative with aipu_cores_used set to 4 and resources_used set to 1.0, the compilation fails. I will attach the full traceback.

Using batch mode instead compiles successfully, however the input shape gets promoted from batch 1 to batch 4.

My questions are: what is the intended use of cooperative mode and how does it differ from batch mode in practice? Is it available in v1.8 or is it something still in development? And if cooperative is not the right tool here, is there a recommended way to have all 4 cores work together on a single inference?

 

Attached the python script i am using for compiling. I am using axrunmodel for running the compiled ”model”.

 

Output of axdevice:
Device 0: metis-0:6:0 16GiB metis-pcie flver=1.3.2 bcver=1.4 clock=800MHz mvm=100%