Multi-tenant fine-tuning for LLMs with Tinker-compatible API
-
Updated
Sep 2, 2026 - Python
Multi-tenant fine-tuning for LLMs with Tinker-compatible API
Turbo Ultimate Field Fare is a MacOS app that lets users run models like Qwen, Gemma, and GPT-OSS models with expert-streaming, allowing for large models on devices without a lot of memory.
CROW - Your AI Agent. MCPs, OpenRouter, Any Model or local. It's your choice.
Notolog Markdown Editor
Fast-ASDLC: 5x TTM with AI-native Agentic SDLC. Local-LLM first, Human-in-the-loop, Spec-driven. Built on DDD, Hexagonal Architecture, C4 Model & MCP. Features Meta-agents for self-improvement, Memory Bank for context persistence, and automated 100% test coverage. Everything-as-Code & Mermaid.js centric to save context window and slash token costs.
Describe images with Ollama
A from-scratch implementation of the language-model training pipeline in PyTorch: tokenization, pretraining, SFT, and preference-based alignment.
🚀 Unified NLP Pipelines for Language Models
Delta: LLM conversation branching
J.A.R.V.I.S: An AI-powered Open Source Intelligence (OSINT) system. It orchestrates deep web scraping and local LLMs to autonomously generate comprehensive intelligence dossiers.
Playground for learning by doing
A Unity package for building open-source AI voice agents that run fully locally. You can use it to build intelligent non-player characters (NPCs), game interfaces, among many other applications.
XR — the secure, self-hosted AI agent. BYOK · local-first · spend-capped · tamper-evident. by rrrtx
Nova Studio - Windows desktop workbench for local LLM inference (vLLM / SGLang / llama.cpp via WSL2), bridging WSL engines to any OpenAI-compatible client
The Operating System for Local Intelligence. ⚙️
🤖 Autonomous AI career engine — scans 28,700+ real tech company ATS boards (Greenhouse, Lever, Ashby, Workday & SmartRecruiters), monitors Wellfound alerts + custom career pages, scores via cosine similarity + LLM rubrics, generates STAR-tailored pitches for 1-click review. 100% Free.
A lightweight CLI to orchestrate Gemini and GPT using your local files as a shared blackboard.
A minimalist terminal script that analyzes your hardware and use cases to recommend the best local AI models you can reliably run. Powered by the free Gemini 3.1 Flash Lite model, it uses your own Google API key at runtime. You can generate your key for free in under 90 seconds in Google AI Studio, and that's completely safe & costs nothing!
Local-first RAG endpoint that answers cheap questions locally and escalates hard ones to the cloud automatically, behind your LLM Proxy
GGUF-Runner - Want to run LLMs locally, use this guide, and run with LLAMA.cpp
To associate your repository with the local-llms topic, visit your repo's landing page and select "manage topics."