WebLLM | Home
This project brings large-language model and LLM-based chatbot to web browsers. Everything runs inside the browser with no server support and accelerated with WebGPU. This opens up a lot of fun opportunities to build AI assistants for everyone and enable privacy while enjoying GPU acceleration. Please check out our GitHub repo to see how we did it. There is also a demo which you can try out. We have been seeing amazing progress in generative AI and LLM recently. Thanks to the open-source efforts like LLaMA, Alpaca, Vicuna and Dolly, we start to see an exciting future of building our own open source language models and personal AI assistant. These models are usually big and compute-heavy. To build a chat service, we will need a large cluster to run an inference server, while clients send requests to servers and retrieve the inference output. We also usually have to run on a specific type of GPUs where popular deep-learning frameworks are readily available. This project is our step to br
WebLLM | Home --> Home GitHub WebLLM: High-Performance In-Browser LLM Inference Engine Get Started Chat with WebLLM Overview We have been seeing amazing progress in generative AI and LLM recently. Thanks to the open-source efforts like LLaMA, Alpaca, Vicuna and Dolly, we start to see an exciting future of building our own open source language models and personal AI assistant. These models are usually big and compute-heavy. To build a chat service, we will need a large cluster to run an inference server, while clients send requests to servers and retrieve the inference output. We also usually h
related reading
- MLC | WebLLM: A High-Performance In-Browser LLM Inference Engineblog.mlc.ai
- Together AI | The AI Native Cloudtogether.ai
- LLM Visualizationbbycroft.net
- Parsed | Custom, interpretable AI systems that continuously learnparsed.com
- AI Model & API Providers Analysis | Artificial Analysisartificialanalysis.ai
- GitHub - open-webui/open-webui: User-friendly AI Interface (Supports Ollama, OpenAI API, ...)github.com
- LangChain: the open agent platform to own your intelligencelangchain.com
- Hugging Face – The AI community building the future.huggingface.co
- Replicate - Run AI with an APIreplicate.com
- API Overview | OpenAI API Referenceplatform.openai.com
- GitHub - x1xhlol/system-prompts-and-models-of-ai-tools: FULL Augment Code, Claude Code, Cluely, CodeBuddy, Comet, Cursor, Devin AI, Junie, Kiro, Leap.new, Lovable, Manus, NotionAI, Orchids.app, Perplexity, Poke, Qoder, Replit, Same.dev, Trae, Traycer AI, VSCode Agent, Warp.dev, Windsurf, Xcode, Z.ai Code, Dia & v0. (And other Open Sourced) System Prompts, Internal Tools & AI Modelsgithub.com
- GitHub - Hannibal046/Awesome-LLM: Awesome-LLM: a curated list of Large Language Modelgithub.com