GGUF
conversational

YAML Metadata Warning:empty or missing yaml metadata in repo card

Check out the documentation for more information.

ΩFFFΣLLIa • llama.cpp • OFFFELLIA_PURE

OFFFELLIA_PURE Banner

OFFFELLIA_PURE Web UI -

   ██████╗ ███████╗███████╗███████╗██╗     ██╗     ██╗ █████╗ 
  ██╔═══██╗██╔════╝██╔════╝██╔════╝██║     ██║     ██║██╔══██╗
  ██║   ██║█████╗  █████╗  █████╗  ██║     ██║     ██║███████║
  ██║   ██║██╔══╝  ██╔══╝  ██╔══╝  ██║     ██║     ██║██╔══██║
  ╚██████╔╝██║     ██║     ███████╗███████╗███████╗██║██║  ██║-PURE
   ╚═════╝ ╚═╝     ╚═╝     ╚══════╝╚══════╝╚══════╝╚═╝╚═╝  ╚═╝

High-Performance LLM / VLM Inference & Autonomous Agentic Ecosystem in Pure C/C++

License: MIT C++: 17/20 WebUI: SvelteKit + Vite Status: Unlocked & Optimized Agentic: Multi--Turn Engine


📖 Visão Geral

ΩFFFΣLLIa • llama.cpp • AlgMor24 é um fork avançado, destravado e de alta performance do ecossistema llama.cpp. Este projeto integra inferência local de última geração em C/C++ com um motor agêntico autônomo multi-turn, suporte nativo a FIM (Fill-in-the-Middle) para geração e preenchimento de código, Speculative Decoding otimizado para programação, integração de ferramentas MCP (Model Context Protocol) e uma interface Web moderna em SvelteKit/Vite com a identidade visual Cyberpunk Neon Fire.


git clone https://github.com/brunoconta1980-tech/llama_OFFFELLIA_1984

cd llama_OFFFELLIA_1984

cmake -B build
-DGGML_VULKAN=ON
-DLLAMA_BUILD_WEBUI=ON
-DLLAMA_SERVER_TOOLS=ON

cmake --build build -j

or

cmake -S . -B build-vulkan
-DGGML_VULKAN=ON
-DLLAMA_BUILD_WEBUI=ON
-DLLAMA_SERVER_TOOLS=ON

cmake --build build-vulkan -j

Comando sugerido:

"/home/userk21/llama_OFFFELLIA_1984/build/bin/llama-server"
-m "/home/userk21/Área de trabalho/userk21/LLMS/ΩFFFΣLLIα_IQ4_NL_gemma-4-26B-A4B-it.gguf"
-ngl 99 --n-cpu-moe 99
-c 50000
-ctk q8_0
-ctv q8_0
-t 4
-tb 4
-b 2048
-ub 1024
-fa on
--cpu-strict 1
--parallel 1
--agent
--tools all
--reasoning auto
--kv-unified
--load-mode mmap
--cors-origins "*"
--webui-mcp-proxy
--threads-http -1
--port 5173
--host 127.0.0.1


📜 Licença

Distribuído sob a licença MIT. Veja o arquivo LICENSE para mais detalhes. Gracias https://github.com/charlie12345/ROCmFPX


ΩFFFΣLLIa • llama.cpp • AlgMor24 — Inferência Local, Descentralizada e Sem Limites
Downloads last month
1,686
GGUF
Model size
8B params
Architecture
qwen2vl
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support