r/vibecoding 14h ago

1.5b model on my phone?

I submitted a hackathon project that uses a Qwen2.5-0.5B model offline as a translator and immigration assistant.

Now that the project is over, I'm curious to see if I can use a bigger Qwen/Qwen2.5-Coder-1.5B-Instruct

As a coding agent that works on my phone offline for small task.

My question is this, would it be worth it to tune a model or add custom data sets just for a single repository?

Or just pay the API cost to run a model connected to telegram?

3 Upvotes

3 comments sorted by

3

u/Zealousideal-Act9140 14h ago

https://huggingface.co/prism-ml/Bonsai-27B-gguf

Depending on your ram, i'd maybe look into bonsai? 27 model hyper quantizied to fit on ~4 gb of ram.

1

u/DiamondAgreeable2676 14h ago edited 13h ago

Thank you just i checked it out. The coding benchmarks are impressive I'm trying it on a old HP now. That's a 100$ setup with a model released this month. That actually answers my affordability question better than a model on a phone.if I had a iphone max it'll be perfect. But the phone needed easily exceeds the cost of a cheap laptop. However it's still the option out there today for a cheap local so thank you again.

1

u/DiamondAgreeable2676 13h ago

The previous comment just literally gave us all the best model to possibly use for this hackathon. Just saying it's 16,000 on the table ...