Developer ngxson created a software prototype that allows for the training of AI models directly on a user's machine without an external server. The proof of concept demonstrates that large language models can be fine-tuned in a web browser using WebGPU, backed by the llama.cpp and wllama libraries. This approach eliminates the need for expensive cloud infrastructure during the model adaptation process.

This development shifts the capability of browser-based AI beyond inference, which only runs pre-existing models to generate output. Current efforts are focused on integrating Low-Rank Adaptation, known as LoRA, to further optimize the memory and power required for client-side training.

Sign in to suggest edits

Key sources

  1. SOURCE@ngxson“backed by llama.cpp / wllama”x.com
  2. SUPPORT@maximelabonne“In-browser fine-tuning is coming”x.com
Markdown