SIGN IN SIGN UP

Arm backend: Add serialised xlarge VKML model suite (#21518)

The DeepSeek-R1-Distill-Qwen layer tests were re-landed after the TOSA
model shard OOM was addressed by a serialised xlarge TOSA suite.

CommitlyExtended still showed memory pressure in the VKML model shard.
The normal VKML model shard uses pytest-xdist auto parallelism, so
xlarge VGF exports can also be resident at the same time.

Add a serialised xlarge VKML model suite while keeping the normal VKML
model shard parallelised. Mark the DeepSeek VGF layer tests as xlarge so
they are excluded from normal VKML and can be routed through the
serialised suite.



cc @digantdesai @freddan80 @per @zingo @oscarandersson8218 @mansnils
@Sebastian-Larsson @robell @rascani
B
Baris committed
d46e62016be1c11a69e249985c714ce94380c2ab
Parent: 9789eea
Committed by GitHub <noreply@github.com> on 8/2/2026, 6:16:40 PM