ΩFFFΣLLIa • llama.cpp • OFFFELLIA_HERETIC


██████╗ ███████╗███████╗███████╗██╗ ██╗ ██╗ █████╗
██╔═══██╗██╔════╝██╔════╝██╔════╝██║ ██║ ██║██╔══██╗
██║ ██║█████╗ █████╗ █████╗ ██║ ██║ ██║███████║
██║ ██║██╔══╝ ██╔══╝ ██╔══╝ ██║ ██║ ██║██╔══██║
╚██████╔╝██║ ██║ ███████╗███████╗███████╗██║██║ ██║-HERETIC
╚═════╝ ╚═╝ ╚═╝ ╚══════╝╚══════╝╚══════╝╚═╝╚═╝ ╚═╝
High-Performance LLM / VLM Inference & Autonomous Agentic Ecosystem in Pure C/C++
📖 Visão Geral
ΩFFFΣLLIa • llama.cpp • AlgMor24 é um fork avançado, destravado e de alta performance do ecossistema llama.cpp. Este projeto integra inferência local de última geração em C/C++ com um motor agêntico autônomo multi-turn, suporte nativo a FIM (Fill-in-the-Middle) para geração e preenchimento de código, Speculative Decoding otimizado para programação, integração de ferramentas MCP (Model Context Protocol) e uma interface Web moderna em SvelteKit/Vite com a identidade visual Cyberpunk Neon Fire.
✨
git clone https://github.com/brunoconta1980-tech/llama_OFFFELLIA_1984
TKS TO: https://huggingface.co/mlasli/Nemotron-3.5-Lightning-30B-A3B-Heretic-Uncensored-BF16/tree/main
cd llama_OFFFELLIA_1984
cmake -B build
-DGGML_VULKAN=ON
-DLLAMA_BUILD_WEBUI=ON
-DLLAMA_SERVER_TOOLS=ON
cmake --build build -j
or
cmake -S . -B build-vulkan
-DGGML_VULKAN=ON
-DLLAMA_BUILD_WEBUI=ON
-DLLAMA_SERVER_TOOLS=ON
cmake --build build-vulkan -j
Comando sugerido:
"/home/userk21/OFFFELLIA_llama.cpp_Neon_Themes/build/bin/llama-server"
-m "/home/userk21/Área de trabalho/userk21/LLMS/ΩFFFΣLLIα_Q5_K_coder3101_gemma-4-E4B-8b_666-tensors-heretic.gguf"
-ngl 99 --n-cpu-moe 99
-c 50000
-t 5
-tb 5
-ctk q8_0
-ctv q8_0
-fa on
--cpu-strict 1
--parallel 1
--agent
--tools all
--reasoning auto
--kv-unified
--load-mode mmap
--cors-origins "*"
--webui-mcp-proxy
--threads-http -1
--port 8080
--host 127.0.0.1
📜 Licença
Distribuído sob a licença MIT. Veja o arquivo LICENSE para mais detalhes.
Gracias https://github.com/charlie12345/ROCmFPX