One-shot full-stack apps with your existing AI coding tool. Puter.js gives you Auth, Storage, DB, AI & more, with up to 90% fewer AI tokens than other backend platforms. Learn more →
Llama.vscode Alternatives
Similar projects and alternatives to llama.vscode
-
-
Puter.js
Puter.js - The Backend for AI-Generated Apps. One-shot full-stack apps with your existing AI coding tool. Puter.js gives you Auth, Storage, DB, AI & more, with up to 90% fewer AI tokens than other backend platforms.
-
-
-
ghostty
👻 Ghostty is a fast, feature-rich, and cross-platform terminal emulator that uses platform-native UI and GPU acceleration.
-
-
-
-
SaaSHub
SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives
-
-
-
-
claude-code-router
One local control plane for every AI agent: route across models, fuse new capabilities, orchestrate tools, and stay fully in control.
-
-
-
-
-
-
-
llama-swap
Reliable model swapping for any local OpenAI/Anthropic compatible server - llama.cpp, vllm, etc
-
-
SaaSHub
SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives
llama.vscode discussion
llama.vscode reviews and mentions
- Run Gemma-4 E2B-it with llama.cpp on Raspberry Pi4
- Local LLM Inference on Windows 11 and AMD GPU using WSL and llama.cpp
-
Microsoft lowers AI software sales quota
I am working on a larger project about containers and isolation stronger than current conventions but short kata etc…
But if you follow the podman instructions for cuda, the llama.cpp shows you how to use their plugin here
https://github.com/ggml-org/llama.vscode
-
Ask HN: Who uses open LLMs and coding assistants locally? Share setup and laptop
how are you running qwen3 with llama-vscode? I am still using qwen-2.5-7b.
There is an open issue about adding support for Qwn3 which I have been monitoring, would love to use Qwen3 if possible. Issue - https://github.com/ggml-org/llama.vscode/issues/55
-
Qwen3-Coder: Agentic Coding in the World
I'm on a m1 max with 64gb ram, but i never use this vscode plugin before. Should I try?
Is this the one? https://github.com/ggml-org/llama.vscode it sems to be built for code completion rather than outright agent mode
-
GitHub Copilot Pro+
- https://github.com/cline/cline
- https://github.com/ggml-org/llama.vscode
I'm sure there are others if you're interesting in switching to something else.
- Local LLM-assisted text completion extension for VS Code
-
Trae: An AI Powered IDE by ByteDance
You can run it all locally: https://github.com/ggml-org/llama.vscode
-
A note from our sponsor - Puter.js
developer.puter.com | 19 Aug 2026
Stats
ggml-org/llama.vscode is an open source project licensed under MIT License which is an OSI approved license.
The primary programming language of llama.vscode is TypeScript.