Tag: llama.cpp
Demo 2: llama-Server AI Android App+Qwen3-VL-8B GGUF(Text GGUF only)(taken base llama-Android Example) with Enhancement via AI Coding to run GGUF model and Serve as Server for AI Agents and Other Android App (demo speed is 10x-long response by Qwen3-VL-8B(15 Minutes).
Demo 1: llama-Server AI Android App+Smollm3 GGUF(taken base llama-Android Example) with Enhancement via AI Coding to run GGUF model and Serve as Server for AI Agents and Other Android App.
(Preview-Custom Code UI-AI-NLP)-Apache Superset (AI-Chat-Custom Page)(Calling All Dashboards in Natural language within AI Chat inline)-Powered with llama-Server & Qwen3.5–9B(Q4_K_M).