A Python framework for self-hosted LLM tool-calling and multi-step agentic workflows
-
Updated
Jul 18, 2026 - Python
A Python framework for self-hosted LLM tool-calling and multi-step agentic workflows
Create an IRC chat bot powered by AI, using llamafile, in minutes.
A simple github actions script to build a llamafile and uploads to huggingface
Portable offline AI kit on a USB SSD — run LLM chat + Wikipedia + encrypted vault anywhere. No internet, no install, no subscriptions. Windows/macOS/Linux/Android/iOS. Powered by Qwen3 + llamafile.
This repository demonstrates LLM execution on CPUs using packages like llamafile, emphasizing low-latency, high-throughput, and cost-effective benefits for inference and serving.
👾 Self-Healing AI Terminal Assistant — 100% local, offline-capable, privacy-first. Multi-tier LLM (llamafile/GGUF), auto-fix with FixNet consensus, 80+ commands. No cloud, no API keys needed.
An offline knowledge library on a USB drive. Wikipedia, Stack Exchange, Gutenberg, and Khan Academy, plus a local LLM that answers questions using those sources as context. Works on any computer, no internet required.
Gathering insights from Common Crawl using Apache Spark and LLMs.
Training materials on how to deploy generative AI models locally on your laptop or workstation.
Deterministic local-first AI operator runtime with receipts, replay, rollback, ranked memory, llamafile CPU driven execution, swappable intelligence cognition core, and sandboxed self-improvement.
UAE / Saudi LLM llamafile — Single-file, no-install executables for Falcon, Jais, ALLaM, and more.
Run Claude Code offline on Mac mini M4 — local Qwen3 via llamafile, OS-level sandbox via Safehouse, zero API cost.
Add a description, image, and links to the llamafile topic page so that developers can more easily learn about it.
To associate your repository with the llamafile topic, visit your repo's landing page and select "manage topics."