adenta 4 hours ago
Meet Jerry- it's literally a guy named Jerry answering your questions.
mattvr an hour ago
keyle 4 hours ago
It can technically be used for a lot of use cases, I'd like people to chime in on ideas on this?
hrpnk 40 minutes ago
Since it's a LoRa on Qwen, I assume this is runnable via llama.cpp. Pity that the PEFT/LoRa->GGUF translation is left to the user. Anyone got past:
$ uv run --with transformers==5.19.0 convert_lora_to_gguf.py ~/Downloads/lora --dry-run --verbose
[...]
File "/Users/user/repos/llama.cpp/conversion/base.py", line 630, in map_tensor_name
raise ValueError(f"Can not map tensor {name!r}")
ValueError: Can not map tensor 'layers.0.linear_attn.in_proj_a.weight'SubiculumCode an hour ago
soltanov 2 hours ago
miguelspizza 4 hours ago
It is just the right mix of size, capability and speed to make it generally useful for adhoc bulk classification tasks.
For those wanting to run it in browser: https://huggingface.co/alxnahas/strands-decider-2B-webgpu
teruakohatu 2 hours ago
davvie 2 hours ago
stephantul an hour ago
yieldcrv 2 hours ago
lin7c 38 minutes ago
Fluid_Mechanics 4 hours ago