| deepgrove/Bonsai | awq | A10 | wikitext2-train n=128 | 0 | OSError("/cache/quants/bonsai-0p5b__awq__wikitext2__n128__s0 does not appear to have a file named modeling_qllama.py. Checkout 'https://huggingface.co//cache/quants/bonsai-0p5b__awq__wikitext2__n128__s0/tree/main' for available files.") |
| deepgrove/Bonsai | awq | T4 | wikitext2-train n=128 | 0 | OSError("/cache/quants/bonsai-0p5b__awq__wikitext2__n128__s0 does not appear to have a file named modeling_qllama.py. Checkout 'https://huggingface.co//cache/quants/bonsai-0p5b__awq__wikitext2__n128__s0/tree/main' for available files.") |
| deepgrove/Bonsai | awq | A10 | wikitext2-train n=128 | 1 | OSError("/cache/quants/bonsai-0p5b__awq__wikitext2__n128__s1 does not appear to have a file named modeling_qllama.py. Checkout 'https://huggingface.co//cache/quants/bonsai-0p5b__awq__wikitext2__n128__s1/tree/main' for available files.") |
| deepgrove/Bonsai | awq | T4 | wikitext2-train n=128 | 1 | OSError("/cache/quants/bonsai-0p5b__awq__wikitext2__n128__s1 does not appear to have a file named modeling_qllama.py. Checkout 'https://huggingface.co//cache/quants/bonsai-0p5b__awq__wikitext2__n128__s1/tree/main' for available files.") |
| deepgrove/Bonsai | awq | A10 | wikitext2-train n=128 | 2 | OSError("/cache/quants/bonsai-0p5b__awq__wikitext2__n128__s2 does not appear to have a file named modeling_qllama.py. Checkout 'https://huggingface.co//cache/quants/bonsai-0p5b__awq__wikitext2__n128__s2/tree/main' for available files.") |
| deepgrove/Bonsai | awq | T4 | wikitext2-train n=128 | 2 | OSError("/cache/quants/bonsai-0p5b__awq__wikitext2__n128__s2 does not appear to have a file named modeling_qllama.py. Checkout 'https://huggingface.co//cache/quants/bonsai-0p5b__awq__wikitext2__n128__s2/tree/main' for available files.") |
| deepgrove/Bonsai | none-fp16 | A10 | none n=0 | 0 | a10 batch call failed: RuntimeError("ImportError: cannot import name 'LossKwargs' from 'transformers.utils' (/usr/local/lib/python3.11/site-packages/transformers/utils/__init__.py)") |
| deepgrove/Bonsai | gptq | A10 | wikitext2-train n=128 | 0 | a10 batch call failed: RuntimeError("ImportError: cannot import name 'LossKwargs' from 'transformers.utils' (/usr/local/lib/python3.11/site-packages/transformers/utils/__init__.py)") |
| deepgrove/Bonsai | gptq | A10 | wikitext2-train n=128 | 1 | a10 batch call failed: RuntimeError("ImportError: cannot import name 'LossKwargs' from 'transformers.utils' (/usr/local/lib/python3.11/site-packages/transformers/utils/__init__.py)") |
| deepgrove/Bonsai | gptq | A10 | wikitext2-train n=128 | 2 | a10 batch call failed: RuntimeError("ImportError: cannot import name 'LossKwargs' from 'transformers.utils' (/usr/local/lib/python3.11/site-packages/transformers/utils/__init__.py)") |
| Qwen/Qwen2.5-3B-Instruct | awq | A10 | openhermes-2.5 n=512 | 0 | OutOfMemoryError('CUDA out of memory. Tried to allocate 4.96 GiB. GPU 0 has a total capacity of 22.06 GiB of which 419.44 MiB is free. Process 1 has 21.64 GiB memory in use. Of the allocated memory 16.42 GiB is allocated by PyTorch, and 4.92 GiB is reserved by PyTorch but unallocated. If reserved but unallocated memory is large try setting PYTORCH_CUDA_ALLOC_CONF=expandable_segments:True to avoid |
| Qwen/Qwen2.5-3B-Instruct | awq | A10 | openhermes-2.5 n=512 | 1 | OutOfMemoryError('CUDA out of memory. Tried to allocate 4.87 GiB. GPU 0 has a total capacity of 22.06 GiB of which 4.56 GiB is free. Process 1 has 17.49 GiB memory in use. Of the allocated memory 16.76 GiB is allocated by PyTorch, and 436.95 MiB is reserved by PyTorch but unallocated. If reserved but unallocated memory is large try setting PYTORCH_CUDA_ALLOC_CONF=expandable_segments:True to avoid |
| Qwen/Qwen2.5-3B-Instruct | awq | A10 | openhermes-2.5 n=512 | 2 | OutOfMemoryError('CUDA out of memory. Tried to allocate 4.91 GiB. GPU 0 has a total capacity of 22.06 GiB of which 3.35 GiB is free. Process 1 has 18.70 GiB memory in use. Of the allocated memory 16.86 GiB is allocated by PyTorch, and 1.54 GiB is reserved by PyTorch but unallocated. If reserved but unallocated memory is large try setting PYTORCH_CUDA_ALLOC_CONF=expandable_segments:True to avoid fr |
| Qwen/Qwen2.5-3B-Instruct | awq | A10 | wikitext2-train n=512 | 0 | OutOfMemoryError('CUDA out of memory. Tried to allocate 4.72 GiB. GPU 0 has a total capacity of 22.06 GiB of which 1.27 GiB is free. Process 1 has 20.78 GiB memory in use. Of the allocated memory 15.79 GiB is allocated by PyTorch, and 4.69 GiB is reserved by PyTorch but unallocated. If reserved but unallocated memory is large try setting PYTORCH_CUDA_ALLOC_CONF=expandable_segments:True to avoid fr |
| Qwen/Qwen2.5-3B-Instruct | awq | A10 | wikitext2-train n=512 | 1 | OutOfMemoryError('CUDA out of memory. Tried to allocate 4.71 GiB. GPU 0 has a total capacity of 22.06 GiB of which 427.44 MiB is free. Process 1 has 21.63 GiB memory in use. Of the allocated memory 16.48 GiB is allocated by PyTorch, and 4.85 GiB is reserved by PyTorch but unallocated. If reserved but unallocated memory is large try setting PYTORCH_CUDA_ALLOC_CONF=expandable_segments:True to avoid |
| Qwen/Qwen2.5-3B-Instruct | awq | A10 | wikitext2-train n=512 | 2 | OutOfMemoryError('CUDA out of memory. Tried to allocate 4.62 GiB. GPU 0 has a total capacity of 22.06 GiB of which 411.44 MiB is free. Process 1 has 21.65 GiB memory in use. Of the allocated memory 17.10 GiB is allocated by PyTorch, and 4.25 GiB is reserved by PyTorch but unallocated. If reserved but unallocated memory is large try setting PYTORCH_CUDA_ALLOC_CONF=expandable_segments:True to avoid |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=128 | 0 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=128 | 1 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=128 | 2 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=32 | 0 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=32 | 1 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=32 | 2 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=512 | 0 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=512 | 1 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | openhermes-2.5 n=512 | 2 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | wikitext2-train n=128 | 0 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | wikitext2-train n=128 | 1 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | wikitext2-train n=128 | 2 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | wikitext2-train n=32 | 0 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | wikitext2-train n=32 | 1 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | wikitext2-train n=32 | 2 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | A10 | wikitext2-train n=512 | 0 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | wikitext2-train n=512 | 1 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |
| HuggingFaceTB/SmolLM3-3B | gptq | T4 | wikitext2-train n=512 | 2 | ValueError('No compatible quant linear was found for this module: SmolLM3ForCausalLM') |