Out of resource: shared memory

#16

by iszhaoxin - opened Jun 8

Discussion

iszhaoxin

Jun 8

•

edited Jun 8

I got the error message as below:

'triton.runtime.autotuner.OutOfResources: out of resource: shared memory, Required: 135200, Hardware limit: 101376. Reducing block sizes or num_stages may help.'

I tried on both RTX A6000 and RTX 6000.
I guess maybe it is because the model is only trained and tested on specific types GPUs, such as A100?

Satandon1999

Jun 10

Yes, in my experience as well this model works well only on the GPUs listed as 'tested' in the documentation.

LeeStott

Microsoft org Jul 19

The recommended adjustment layer is

"target_modules": [
"o_proj",
"qkv_proj"
]

joker26

Jul 21

@LeeStott how to achieve

ProfLinh

1 day ago

@LeeStott I'm running into this error as well. Can you show us how to adjust the layer?

Satandon1999

1 day ago

If using PEFT you can set this using the "target_modules" parameter in LoraConfig
https://huggingface.co./docs/peft/package_reference/lora#peft.LoraConfig

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.

Tap or paste here to upload images

· Sign up or log in to comment