
local-llm
A comprehensive guide for building and configuring a high-end local machine to run state-of-the-art LLMs, with detailed hardware choices, BIOS tuning, and Docker-based model serving.
It goes beyond simple OCR API calls to deliver a complete, production-ready system with advanced features like multi-token prediction, zero-copy RAM handoff, KEDA-based autoscaling, and enterprise security controls.
Build, deploy, and scale a production-grade OCR pipeline using Rust, vLLM, Redis, KEDA, and Kubernetes.
Similar projects matched by category, topics, and programming language.

A comprehensive guide for building and configuring a high-end local machine to run state-of-the-art LLMs, with detailed hardware choices, BIOS tuning, and Docker-based model serving.
Nativ is a native macOS app that lets you run AI models locally on Apple silicon, offering chat, model management, performance analytics, and an OpenAI/Anthropic-compatible API server.
Cue is a free, open-source AI copilot that lives on your screen, sees your screen and hears your meetings, and tries to stay hidden in screen shares, using your own API key from providers like OpenAI, Anthropic, or Google Gemini.